Connections
Amazi AI Gateway
Connect OpenAI, browse its model catalog, keep Gateway credentials server-side, and inspect usage.
What the gateway does
Amazi AI Gateway gives applications in a project managed access to OpenAI. Only OpenAI is currently supported. It is separate from the external Integrations catalog: provider connections have their own project tab, plan capability, credentials, model catalog, and usage accounting. This feature controls AI services used by your applications; it does not enable or disable the Amazi agent in project chat. Chat work still requires the applicable project permissions and credits.
Open Gateway in the project and connect it when your role and the project owner's plan allow the action. Amazi enables OpenAI, creates and stores its managed credentials, and shows its environment names. Other providers are not available in the project Gateway list. Disconnecting removes Gateway access and its voice-agent routes without turning providers into normal integrations.
Use the centered tabs below the header: Providers opens first and contains provider credentials and model catalogs. Expand Available endpoints above Available models to see the supported OpenAI HTTP paths and how to append them to the base URL. Expand Available models to see a table with model names, API identifiers, categories and descriptions. Each tab has one information block. SIP Realtime contains voice agents, their Realtime addresses and SIP details. Switching tabs preserves your selected provider and unfinished voice-agent form.
If Gateway cannot load, use Retry. A loading error does not mean Gateway has been disabled for your account; plan and role restrictions have separate notices.
Use it from server code
Gateway credentials are server secrets:
- call the provider from a backend or service application;
- use the exact environment names shown by Amazi;
- never copy a credential into chat, source control, a browser variable, or frontend code.
The agent receives instructions and variable names, not the saved secret value.
Ordinary HTTP requests
In backend code, send provider-compatible requests to the injected base URL and authenticate with the injected key as a bearer token. Select a model from the provider catalog shown by Amazi. It lists model names, exact API identifiers, categories and descriptions. The available providers and models table below is loaded from Amazi's current catalog. The agent can inspect endpoint, format and pricing metadata when selecting a model for your application.
The operation paths below are defined by the OpenAI API. Amazi AI Gateway forwards supported
requests to the corresponding OpenAI endpoints using those same paths.
The OpenAI base URL ends in /openai/v1. Append the operation path from this table to it;
do not append another /v1. Every operation below uses POST:
| Operation | Path to append |
|---|---|
| Responses | /responses |
| Chat completions | /chat/completions |
| Speech transcription | /audio/transcriptions |
| Audio translation | /audio/translations |
| Speech generation | /audio/speech |
| Embeddings | /embeddings |
| Image generation | /images/generations |
| Image editing | /images/edits |
| Moderation | /moderations |
For example, transcription uses <base URL>/audio/transcriptions. Sending to the base
/v1 alone does not select an operation, and /transcribe is not a supported path.
Choose an exact model identifier that supports the operation and input format. A model's category
does not choose the route automatically, and Gateway does not expose every provider endpoint.
For JSON operations, set Content-Type: application/json and send the provider's JSON body.
For example, POST <base URL>/embeddings accepts
{"model":"<embedding-model-from-the-catalog>","input":"Example text"} with the bearer key.
For audio transcription, send file and one text model field as multipart/form-data.
The fields may appear in any order. With FormData, let the HTTP client generate the
Content-Type header and boundary. In this backend example, modelName is the selected
transcription model and audioBlob contains the audio file:
const baseUrl = process.env.AMAZI_AI_GATEWAY_OPENAI_BASE_URL;
const key = process.env.AMAZI_AI_GATEWAY_OPENAI_API_KEY;
const form = new FormData();
form.append("file", audioBlob, "recording.wav");
form.append("model", modelName);
form.append("response_format", "json");
const response = await fetch(baseUrl + "/audio/transcriptions", {
method: "POST",
headers: { Authorization: "Bearer " + key },
body: form,
});
if (!response.ok) throw new Error("Transcription failed: " + response.status);
const transcript = await response.json();
Set the request timeout in your application's HTTP client or provider SDK. Gateway does not impose a model-execution deadline. Your application can cancel a request by closing its connection; otherwise it waits for the provider's result or error.
For OpenAI Responses streams, your backend can request live progress with the
X-Amazi-Stream-Mode: live header and stream: true. Text and reasoning summaries arrive
while the model works. A response marked x-amazi-live-stream: 1 reports its settled cost in
an amazi.billing event containing chargedCredits and requestId, before the completion
event. Count this charge once per request ID; if it is missing, show the cost as unavailable.
Realtime WebSocket
Realtime is available to server applications. Convert the injected base URL from https to wss
(or http to ws), append /realtime?model=<exact-model-name>, and send the same bearer
credential in the WebSocket upgrade headers. Realtime events retain the OpenAI protocol. Do not
request a browser client secret and do not put this server credential into frontend code.
Amazi does not impose a fixed call duration. A session remains connected while the provider is
available and the project owner has credits. Completed responses are metered during the session. An
Amazi transport failure arrives as an amazi.gateway.error event containing an actionable code,
message, requestId, stage, and retryable flag; use these fields in safe backend logs.
Phone voice agents
After Gateway is connected, open SIP Realtime and create one or more OpenAI voice agents. Each agent has an optional name, optional backend configuration and lifecycle webhook URLs, and its own SIP address. Copy that address to the forwarding destination at any compatible SIP operator. Deleting an agent revokes its SIP address without changing other agents. Each card labels the saved configuration endpoint, call webhook, and Realtime WebSocket URLs so you can distinguish their purposes without opening the edit form. If creation reports that the SIP service is unavailable, the agent was not saved. Retry later or contact Support if the problem persists.
Without a configuration URL, the agent uses the default session settings and the first Realtime model
in the current catalog. Add a configuration URL for your own instructions, audio settings or tools.
When one is set, Amazi sends an authenticated GET to it for each session. A preview backend may
be used only for a manual check while its environment is running; an active phone route must use a
production application URL so preview sleep cannot replace the JSON response with a wake response.
The URL must use HTTPS, be reachable from the public internet, and must not rely on browser cookies
or interactive user authentication. Verify its bearer token against the server-only
AMAZI_VOICE_CONFIGURATION_SECRET environment variable and return a provider-native Realtime
session object containing model. Request headers also identify
the session, voice agent, provider, mono PCM16 24 kHz audio contract, and event WebSocket URL. Project
backend code can connect to that URL with the selected provider's injected Gateway key to receive
provider events and return tool results unchanged. This keeps business tools and database access in
the project; Amazi transports the events but does not execute project functions.
Credits and usage
In project chat, type / and choose Check project usage to open the account-wide AI Gateway page with this project selected. The filter stays selected even when the project has no usage yet; you can choose another project or all projects there.
Archived and permanently deleted projects remain in usage statistics with separate Archived and Deleted labels. You can still select them to inspect past usage, and their charges remain included in the totals. Restoring an archived project removes its archive label.
To display the exact cost inside your application, read the x-amazi-charged-credits and
x-amazi-gateway-request-id headers from successful HTTP responses on your server. Sum the
integer credit amounts once per unique request ID. This uses Amazi's settled charge, including
cached tokens and rounding. If the headers are absent, show the cost as unavailable rather than zero.
Gateway requests from a project are charged to the current project owner's Amazi credit balance. They share the account's plan and purchased credits with personal agent work. The account-wide Gateway page identifies whether usage came from an application gateway or the project agent. Usage distinguishes project-agent, image-generation and audio-transcription operations; older Research and Code records remain in history, alongside provider, model, project, status, timing, token totals, and credit usage.
If the owner's plan no longer includes gateway access, existing connection records remain but provider use stops until eligibility returns. Contact Support for a provider-side balance or configuration problem; manage your account credits on Plan.
Available providers and models
OpenAI
| Model | API model | Category | Description |
|---|---|---|---|
| babbage-002 | babbage-002 | Text and reasoning | Replacement for the GPT-3 ada and babbage base models |
| Chat Latest | chat-latest | Text and reasoning | Latest Instant model used in ChatGPT |
| davinci-002 | davinci-002 | Text and reasoning | Replacement for the GPT-3 curie and davinci base models |
| GPT-3.5 Turbo | gpt-3.5-turbo | Text and reasoning | Legacy GPT model for cheaper chat and non-chat tasks |
| gpt-3.5-turbo-instruct | gpt-3.5-turbo-instruct | Text and reasoning | OpenAI model available through OpenAI. |
| GPT-4 | gpt-4 | Text and reasoning | An older high-intelligence GPT model |
| GPT-4 Turbo | gpt-4-turbo | Text and reasoning | An older high-intelligence GPT model |
| GPT-4.1 | gpt-4.1 | Text and reasoning | Smartest non-reasoning model |
| GPT-4.1 mini | gpt-4.1-mini | Text and reasoning | Smaller, faster version of GPT-4.1 |
| GPT-4.1 nano | gpt-4.1-nano | Text and reasoning | Fastest, most cost-efficient version of GPT-4.1 |
| GPT-4o | gpt-4o | Text and reasoning | Fast, intelligent, flexible GPT model |
| GPT-4o mini | gpt-4o-mini | Text and reasoning | Fast, affordable small model for focused tasks |
| GPT-5 | gpt-5 | Text and reasoning | Previous intelligent reasoning model for coding and agentic tasks with configurable reasoning effort |
| GPT-5 mini | gpt-5-mini | Text and reasoning | Near-frontier intelligence for cost sensitive, low latency, high volume workloads |
| GPT-5 nano | gpt-5-nano | Text and reasoning | Fastest, most cost-efficient version of GPT-5 |
| GPT-5 Pro | gpt-5-pro | Text and reasoning | Version of GPT-5 that produces smarter and more precise responses |
| GPT-5.1 | gpt-5.1 | Text and reasoning | The best model for coding and agentic tasks with configurable reasoning effort |
| GPT-5.2 | gpt-5.2 | Text and reasoning | Previous frontier model for professional work with configurable reasoning effort |
| GPT-5.2 Pro | gpt-5.2-pro | Text and reasoning | Previous pro model for professional work that produces smarter and more precise responses. |
| GPT-5.3-Codex | gpt-5.3-codex | Text and reasoning | The most capable agentic coding model to date. |
| GPT-5.4 | gpt-5.4 | Text and reasoning | A more affordable model for coding and professional work. |
| GPT-5.4 mini | gpt-5.4-mini | Text and reasoning | Our strongest mini model yet for coding, computer use, and subagents |
| GPT-5.4 nano | gpt-5.4-nano | Text and reasoning | Our cheapest GPT-5.4-class model for simple high-volume tasks |
| GPT-5.4 Pro | gpt-5.4-pro | Text and reasoning | Version of GPT-5.4 that produces smarter and more precise responses. |
| GPT-5.5 | gpt-5.5 | Text and reasoning | A new class of intelligence for coding and professional work. |
| GPT-5.5 Pro | gpt-5.5-pro | Text and reasoning | Version of GPT-5.5 that produces smarter and more precise responses. |
| GPT-5.6 Luna | gpt-5.6-luna | Text and reasoning | GPT-5.6 model optimized for cost-sensitive workloads |
| GPT-5.6 Terra | gpt-5.6-terra | Text and reasoning | GPT-5.6 model that balances intelligence and cost |
| GPT-5.6 Sol | gpt-5.6-sol | Text and reasoning | Frontier model for complex professional work |
| gpt-audio | gpt-audio | Text and reasoning | For audio inputs and outputs with Chat Completions API |
| gpt-audio-1.5 | gpt-audio-1.5 | Text and reasoning | The best voice model for audio in, audio out with Chat Completions. |
| gpt-audio-mini | gpt-audio-mini | Text and reasoning | A cost-efficient version of GPT Audio |
| o1 | o1 | Text and reasoning | Previous full o-series reasoning model |
| o1-pro | o1-pro | Text and reasoning | Version of o1 with more compute for better responses |
| o3 | o3 | Text and reasoning | Reasoning model for complex tasks, succeeded by GPT-5 |
| o3-mini | o3-mini | Text and reasoning | A small model alternative to o3 |
| o3-pro | o3-pro | Text and reasoning | Version of o3 with more compute for better responses |
| o4-mini | o4-mini | Text and reasoning | Fast, cost-efficient reasoning model, succeeded by GPT-5 mini |
| GPT-6 Astra | gpt-6-astra | Text and reasoning | Advanced model for complex reasoning, coding, research, and document workflows. |
| GPT-6 Luna | gpt-6-luna | Text and reasoning | Efficient reasoning model for focused, high-volume tasks. |
| gpt-3.5-turbo-0125 | gpt-3.5-turbo-0125 | Text and reasoning | Documented API model: gpt-3.5-turbo-0125. Alias or version of gpt-3.5-turbo. |
| gpt-3.5-turbo-1106 | gpt-3.5-turbo-1106 | Text and reasoning | Documented API model: gpt-3.5-turbo-1106. Alias or version of gpt-3.5-turbo. |
| gpt-4-turbo-2024-04-09 | gpt-4-turbo-2024-04-09 | Text and reasoning | Documented API model: gpt-4-turbo-2024-04-09. Alias or version of gpt-4-turbo. |
| gpt-4.1-mini-2025-04-14 | gpt-4.1-mini-2025-04-14 | Text and reasoning | Documented API model: gpt-4.1-mini-2025-04-14. Alias or version of gpt-4.1-mini. |
| gpt-4.1-nano-2025-04-14 | gpt-4.1-nano-2025-04-14 | Text and reasoning | Documented API model: gpt-4.1-nano-2025-04-14. Alias or version of gpt-4.1-nano. |
| gpt-4.1-2025-04-14 | gpt-4.1-2025-04-14 | Text and reasoning | Documented API model: gpt-4.1-2025-04-14. Alias or version of gpt-4.1. |
| gpt-4-0613 | gpt-4-0613 | Text and reasoning | Documented API model: gpt-4-0613. Alias or version of gpt-4. |
| gpt-4o-mini-2024-07-18 | gpt-4o-mini-2024-07-18 | Text and reasoning | Documented API model: gpt-4o-mini-2024-07-18. Alias or version of gpt-4o-mini. |
| gpt-4o-2024-11-20 | gpt-4o-2024-11-20 | Text and reasoning | Documented API model: gpt-4o-2024-11-20. Alias or version of gpt-4o. |
| gpt-4o-2024-08-06 | gpt-4o-2024-08-06 | Text and reasoning | Documented API model: gpt-4o-2024-08-06. Alias or version of gpt-4o. |
| gpt-4o-2024-05-13 | gpt-4o-2024-05-13 | Text and reasoning | Documented API model: gpt-4o-2024-05-13. Alias or version of gpt-4o. |
| gpt-5-mini-2025-08-07 | gpt-5-mini-2025-08-07 | Text and reasoning | Documented API model: gpt-5-mini-2025-08-07. Alias or version of gpt-5-mini. |
| gpt-5-nano-2025-08-07 | gpt-5-nano-2025-08-07 | Text and reasoning | Documented API model: gpt-5-nano-2025-08-07. Alias or version of gpt-5-nano. |
| gpt-5-pro-2025-10-06 | gpt-5-pro-2025-10-06 | Text and reasoning | Documented API model: gpt-5-pro-2025-10-06. Alias or version of gpt-5-pro. |
| gpt-5.1-2025-11-13 | gpt-5.1-2025-11-13 | Text and reasoning | Documented API model: gpt-5.1-2025-11-13. Alias or version of gpt-5.1. |
| gpt-5.2-pro-2025-12-11 | gpt-5.2-pro-2025-12-11 | Text and reasoning | Documented API model: gpt-5.2-pro-2025-12-11. Alias or version of gpt-5.2-pro. |
| gpt-5.2-2025-12-11 | gpt-5.2-2025-12-11 | Text and reasoning | Documented API model: gpt-5.2-2025-12-11. Alias or version of gpt-5.2. |
| gpt-5.4-mini-2026-03-17 | gpt-5.4-mini-2026-03-17 | Text and reasoning | Documented API model: gpt-5.4-mini-2026-03-17. Alias or version of gpt-5.4-mini. |
| gpt-5.4-nano-2026-03-17 | gpt-5.4-nano-2026-03-17 | Text and reasoning | Documented API model: gpt-5.4-nano-2026-03-17. Alias or version of gpt-5.4-nano. |
| gpt-5.4-pro-2026-03-05 | gpt-5.4-pro-2026-03-05 | Text and reasoning | Documented API model: gpt-5.4-pro-2026-03-05. Alias or version of gpt-5.4-pro. |
| gpt-5.4-2026-03-05 | gpt-5.4-2026-03-05 | Text and reasoning | Documented API model: gpt-5.4-2026-03-05. Alias or version of gpt-5.4. |
| gpt-5.5-pro-2026-04-23 | gpt-5.5-pro-2026-04-23 | Text and reasoning | Documented API model: gpt-5.5-pro-2026-04-23. Alias or version of gpt-5.5-pro. |
| gpt-5.5-2026-04-23 | gpt-5.5-2026-04-23 | Text and reasoning | Documented API model: gpt-5.5-2026-04-23. Alias or version of gpt-5.5. |
| gpt-5.6-cyber | gpt-5.6-cyber | Text and reasoning | Restricted API access: requires separate Daybreak approval and provisioning. |
| gpt-5-2025-08-07 | gpt-5-2025-08-07 | Text and reasoning | Documented API model: gpt-5-2025-08-07. Alias or version of gpt-5. |
| gpt-audio-mini-2025-12-15 | gpt-audio-mini-2025-12-15 | Text and reasoning | Documented API model: gpt-audio-mini-2025-12-15. Alias or version of gpt-audio-mini. |
| gpt-audio-2025-08-28 | gpt-audio-2025-08-28 | Text and reasoning | Documented API model: gpt-audio-2025-08-28. Alias or version of gpt-audio. |
| gpt-daybreak-blue-latest | gpt-daybreak-blue-latest | Text and reasoning | Restricted API access: requires separate Daybreak approval and provisioning. |
| gpt-daybreak-red-latest | gpt-daybreak-red-latest | Text and reasoning | Restricted API access: requires separate Daybreak approval and provisioning. |
| o1-pro-2025-03-19 | o1-pro-2025-03-19 | Text and reasoning | Documented API model: o1-pro-2025-03-19. Alias or version of o1-pro. |
| o1-2024-12-17 | o1-2024-12-17 | Text and reasoning | Documented API model: o1-2024-12-17. Alias or version of o1. |
| o3-mini-2025-01-31 | o3-mini-2025-01-31 | Text and reasoning | Documented API model: o3-mini-2025-01-31. Alias or version of o3-mini. |
| o3-pro-2025-06-10 | o3-pro-2025-06-10 | Text and reasoning | Documented API model: o3-pro-2025-06-10. Alias or version of o3-pro. |
| o3-2025-04-16 | o3-2025-04-16 | Text and reasoning | Documented API model: o3-2025-04-16. Alias or version of o3. |
| o4-mini-2025-04-16 | o4-mini-2025-04-16 | Text and reasoning | Documented API model: o4-mini-2025-04-16. Alias or version of o4-mini. |
| gpt-5.6 | gpt-5.6 | Text and reasoning | Documented API model: gpt-5.6. Alias or version of gpt-5.6-sol. |
| gpt-3.5-turbo-completions | gpt-3.5-turbo-completions | Text and reasoning | Legacy Completions API alias. Scheduled shutdown: 2026-10-23. |
| gpt-4-0613-completions | gpt-4-0613-completions | Text and reasoning | Legacy Completions API alias. Scheduled shutdown: 2026-10-23. |
| gpt-4-completions | gpt-4-completions | Text and reasoning | Legacy Completions API alias. Scheduled shutdown: 2026-10-23. |
| gpt-4-turbo-completions | gpt-4-turbo-completions | Text and reasoning | Legacy Completions API alias. Scheduled shutdown: 2026-10-23. |
| gpt-5-search-api | gpt-5-search-api | Text and reasoning | Search-specialized Chat Completions API model. |
| gpt-5.5-cyber | gpt-5.5-cyber | Text and reasoning | Restricted Daybreak API model; account provisioning required. |
| gpt-5.4-cyber | gpt-5.4-cyber | Text and reasoning | Restricted Daybreak API model; account provisioning required. |
| gpt-4-1106-preview | gpt-4-1106-preview | Text and reasoning | Legacy API snapshot. Latest lifecycle notice lists 2026-10-23 shutdown; an older notice conflicts, so account availability needs verification. |
| chatgpt-image-latest | chatgpt-image-latest | Images | Previous image model used in ChatGPT. |
| GPT Image 1 | gpt-image-1 | Images | Our previous image generation model |
| gpt-image-1-mini | gpt-image-1-mini | Images | A cost-efficient version of GPT Image 1 |
| GPT Image 1.5 | gpt-image-1.5 | Images | Our previous image generation model |
| GPT Image 2.5 Flare | gpt-image-2.5-flare | Images | Fast image generation and editing with GPT Image 2.5 |
| GPT Image 2.5 Sunburst | gpt-image-2.5-sunburst | Images | Image generation and precise editing with GPT Image 2.5 |
| GPT Image 2 | gpt-image-2 | Images | State-of-the-art image generation model |
| gpt-image-1.5-2025-12-16 | gpt-image-1.5-2025-12-16 | Images | Documented API model: gpt-image-1.5-2025-12-16. Alias or version of gpt-image-1.5. |
| gpt-image-2-2026-04-21 | gpt-image-2-2026-04-21 | Images | Documented API model: gpt-image-2-2026-04-21. Alias or version of gpt-image-2. |
| sora-2-pro | sora-2-pro | Video | Documented API model: sora-2-pro. |
| sora-2-pro-2025-10-06 | sora-2-pro-2025-10-06 | Video | Documented API model: sora-2-pro-2025-10-06. Alias or version of sora-2-pro. |
| sora-2 | sora-2 | Video | Documented API model: sora-2. |
| sora-2-2025-12-08 | sora-2-2025-12-08 | Video | Documented API model: sora-2-2025-12-08. Alias or version of sora-2. |
| sora-2-2025-10-06 | sora-2-2025-10-06 | Video | Documented API model: sora-2-2025-10-06. Alias or version of sora-2. |
| GPT-Realtime | gpt-realtime | Realtime | Model capable of realtime text and audio inputs and outputs |
| GPT-Realtime-1.5 | gpt-realtime-1.5 | Realtime | The best voice model for audio in, audio out |
| GPT-Realtime-2 | gpt-realtime-2 | Realtime | Reasoning model with tool use |
| GPT-Realtime-2.1 | gpt-realtime-2.1 | Realtime | Reasoning model with tool use |
| GPT-Realtime-2.1 mini | gpt-realtime-2.1-mini | Realtime | Reasoning model with tool use |
| GPT-Realtime mini | gpt-realtime-mini | Realtime | A cost-efficient version of GPT-Realtime |
| GPT-Realtime-Translate | gpt-realtime-translate | Realtime | Streaming speech-to-speech translation model |
| gpt-realtime-mini-2025-12-15 | gpt-realtime-mini-2025-12-15 | Realtime | Documented API model: gpt-realtime-mini-2025-12-15. Alias or version of gpt-realtime-mini. |
| gpt-realtime-2025-08-28 | gpt-realtime-2025-08-28 | Realtime | Documented API model: gpt-realtime-2025-08-28. Alias or version of gpt-realtime. |
| text-embedding-3-large | text-embedding-3-large | Embeddings | Most capable embedding model |
| text-embedding-3-small | text-embedding-3-small | Embeddings | Small embedding model |
| text-embedding-ada-002 | text-embedding-ada-002 | Embeddings | Older embedding model |
| GPT-4o mini TTS | gpt-4o-mini-tts | Speech generation | Text-to-speech model powered by GPT-4o mini |
| TTS-1 | tts-1 | Speech generation | Text-to-speech model optimized for speed |
| TTS-1 HD | tts-1-hd | Speech generation | Text-to-speech model optimized for quality |
| gpt-4o-mini-tts-2025-03-20 | gpt-4o-mini-tts-2025-03-20 | Speech generation | Documented API model: gpt-4o-mini-tts-2025-03-20. Alias or version of gpt-4o-mini-tts. |
| gpt-4o-mini-tts-2025-12-15 | gpt-4o-mini-tts-2025-12-15 | Speech generation | Documented API model: gpt-4o-mini-tts-2025-12-15. Alias or version of gpt-4o-mini-tts. |
| GPT-4o mini Transcribe | gpt-4o-mini-transcribe | Speech recognition | Speech-to-text model powered by GPT-4o mini |
| GPT-4o Transcribe | gpt-4o-transcribe | Speech recognition | Speech-to-text model powered by GPT-4o |
| GPT-4o Transcribe Diarize | gpt-4o-transcribe-diarize | Speech recognition | Transcription model that identifies who's speaking when |
| GPT Live Transcribe | gpt-live-transcribe | Speech recognition | Low-latency streaming speech-to-text model for live audio. |
| GPT-Realtime-Whisper | gpt-realtime-whisper | Speech recognition | Streaming speech-to-text model for realtime transcription |
| GPT Transcribe | gpt-transcribe | Speech recognition | High-accuracy speech-to-text model for files and realtime input transcription. |
| Whisper | whisper-1 | Speech recognition | General-purpose speech recognition model |
| gpt-4o-mini-transcribe-2025-03-20 | gpt-4o-mini-transcribe-2025-03-20 | Speech recognition | Documented API model: gpt-4o-mini-transcribe-2025-03-20. Alias or version of gpt-4o-mini-transcribe. |
| gpt-4o-mini-transcribe-2025-12-15 | gpt-4o-mini-transcribe-2025-12-15 | Speech recognition | Documented API model: gpt-4o-mini-transcribe-2025-12-15. Alias or version of gpt-4o-mini-transcribe. |
| omni-moderation | omni-moderation-latest | Moderation | Identify potentially harmful content in text and images |
| omni-moderation-2024-09-26 | omni-moderation-2024-09-26 | Moderation | Documented API model: omni-moderation-2024-09-26. Alias or version of omni-moderation-latest. |