Connections

Amazi AI Gateway

Connect OpenAI, browse its model catalog, keep Gateway credentials server-side, and inspect usage.

What the gateway does

Amazi AI Gateway gives applications in a project managed access to OpenAI. Only OpenAI is currently supported. It is separate from the external Integrations catalog: provider connections have their own project tab, plan capability, credentials, model catalog, and usage accounting. This feature controls AI services used by your applications; it does not enable or disable the Amazi agent in project chat. Chat work still requires the applicable project permissions and credits.

Open Gateway in the project and connect it when your role and the project owner's plan allow the action. Amazi enables OpenAI, creates and stores its managed credentials, and shows its environment names. Other providers are not available in the project Gateway list. Disconnecting removes Gateway access and its voice-agent routes without turning providers into normal integrations.

Use the centered tabs below the header: Providers opens first and contains provider credentials and model catalogs. Expand Available endpoints above Available models to see the supported OpenAI HTTP paths and how to append them to the base URL. Expand Available models to see a table with model names, API identifiers, categories and descriptions. Each tab has one information block. SIP Realtime contains voice agents, their Realtime addresses and SIP details. Switching tabs preserves your selected provider and unfinished voice-agent form.

If Gateway cannot load, use Retry. A loading error does not mean Gateway has been disabled for your account; plan and role restrictions have separate notices.

Use it from server code

Gateway credentials are server secrets:

  • call the provider from a backend or service application;
  • use the exact environment names shown by Amazi;
  • never copy a credential into chat, source control, a browser variable, or frontend code.

The agent receives instructions and variable names, not the saved secret value.

Ordinary HTTP requests

In backend code, send provider-compatible requests to the injected base URL and authenticate with the injected key as a bearer token. Select a model from the provider catalog shown by Amazi. It lists model names, exact API identifiers, categories and descriptions. The available providers and models table below is loaded from Amazi's current catalog. The agent can inspect endpoint, format and pricing metadata when selecting a model for your application.

The operation paths below are defined by the OpenAI API. Amazi AI Gateway forwards supported requests to the corresponding OpenAI endpoints using those same paths. The OpenAI base URL ends in /openai/v1. Append the operation path from this table to it; do not append another /v1. Every operation below uses POST:

OperationPath to append
Responses/responses
Chat completions/chat/completions
Speech transcription/audio/transcriptions
Audio translation/audio/translations
Speech generation/audio/speech
Embeddings/embeddings
Image generation/images/generations
Image editing/images/edits
Moderation/moderations

For example, transcription uses <base URL>/audio/transcriptions. Sending to the base /v1 alone does not select an operation, and /transcribe is not a supported path. Choose an exact model identifier that supports the operation and input format. A model's category does not choose the route automatically, and Gateway does not expose every provider endpoint.

For JSON operations, set Content-Type: application/json and send the provider's JSON body. For example, POST <base URL>/embeddings accepts {"model":"<embedding-model-from-the-catalog>","input":"Example text"} with the bearer key.

For audio transcription, send file and one text model field as multipart/form-data. The fields may appear in any order. With FormData, let the HTTP client generate the Content-Type header and boundary. In this backend example, modelName is the selected transcription model and audioBlob contains the audio file:

const baseUrl = process.env.AMAZI_AI_GATEWAY_OPENAI_BASE_URL;
const key = process.env.AMAZI_AI_GATEWAY_OPENAI_API_KEY;
const form = new FormData();
form.append("file", audioBlob, "recording.wav");
form.append("model", modelName);
form.append("response_format", "json");
const response = await fetch(baseUrl + "/audio/transcriptions", {
  method: "POST",
  headers: { Authorization: "Bearer " + key },
  body: form,
});
if (!response.ok) throw new Error("Transcription failed: " + response.status);
const transcript = await response.json();

Set the request timeout in your application's HTTP client or provider SDK. Gateway does not impose a model-execution deadline. Your application can cancel a request by closing its connection; otherwise it waits for the provider's result or error.

For OpenAI Responses streams, your backend can request live progress with the X-Amazi-Stream-Mode: live header and stream: true. Text and reasoning summaries arrive while the model works. A response marked x-amazi-live-stream: 1 reports its settled cost in an amazi.billing event containing chargedCredits and requestId, before the completion event. Count this charge once per request ID; if it is missing, show the cost as unavailable.

Realtime WebSocket

Realtime is available to server applications. Convert the injected base URL from https to wss (or http to ws), append /realtime?model=<exact-model-name>, and send the same bearer credential in the WebSocket upgrade headers. Realtime events retain the OpenAI protocol. Do not request a browser client secret and do not put this server credential into frontend code.

Amazi does not impose a fixed call duration. A session remains connected while the provider is available and the project owner has credits. Completed responses are metered during the session. An Amazi transport failure arrives as an amazi.gateway.error event containing an actionable code, message, requestId, stage, and retryable flag; use these fields in safe backend logs.

Phone voice agents

After Gateway is connected, open SIP Realtime and create one or more OpenAI voice agents. Each agent has an optional name, optional backend configuration and lifecycle webhook URLs, and its own SIP address. Copy that address to the forwarding destination at any compatible SIP operator. Deleting an agent revokes its SIP address without changing other agents. Each card labels the saved configuration endpoint, call webhook, and Realtime WebSocket URLs so you can distinguish their purposes without opening the edit form. If creation reports that the SIP service is unavailable, the agent was not saved. Retry later or contact Support if the problem persists.

Without a configuration URL, the agent uses the default session settings and the first Realtime model in the current catalog. Add a configuration URL for your own instructions, audio settings or tools. When one is set, Amazi sends an authenticated GET to it for each session. A preview backend may be used only for a manual check while its environment is running; an active phone route must use a production application URL so preview sleep cannot replace the JSON response with a wake response. The URL must use HTTPS, be reachable from the public internet, and must not rely on browser cookies or interactive user authentication. Verify its bearer token against the server-only AMAZI_VOICE_CONFIGURATION_SECRET environment variable and return a provider-native Realtime session object containing model. Request headers also identify the session, voice agent, provider, mono PCM16 24 kHz audio contract, and event WebSocket URL. Project backend code can connect to that URL with the selected provider's injected Gateway key to receive provider events and return tool results unchanged. This keeps business tools and database access in the project; Amazi transports the events but does not execute project functions.

Credits and usage

In project chat, type / and choose Check project usage to open the account-wide AI Gateway page with this project selected. The filter stays selected even when the project has no usage yet; you can choose another project or all projects there.

Archived and permanently deleted projects remain in usage statistics with separate Archived and Deleted labels. You can still select them to inspect past usage, and their charges remain included in the totals. Restoring an archived project removes its archive label.

To display the exact cost inside your application, read the x-amazi-charged-credits and x-amazi-gateway-request-id headers from successful HTTP responses on your server. Sum the integer credit amounts once per unique request ID. This uses Amazi's settled charge, including cached tokens and rounding. If the headers are absent, show the cost as unavailable rather than zero.

Gateway requests from a project are charged to the current project owner's Amazi credit balance. They share the account's plan and purchased credits with personal agent work. The account-wide Gateway page identifies whether usage came from an application gateway or the project agent. Usage distinguishes project-agent, image-generation and audio-transcription operations; older Research and Code records remain in history, alongside provider, model, project, status, timing, token totals, and credit usage.

If the owner's plan no longer includes gateway access, existing connection records remain but provider use stops until eligibility returns. Contact Support for a provider-side balance or configuration problem; manage your account credits on Plan.

Available providers and models

OpenAI

ModelAPI modelCategoryDescription
babbage-002babbage-002Text and reasoningReplacement for the GPT-3 ada and babbage base models
Chat Latestchat-latestText and reasoningLatest Instant model used in ChatGPT
davinci-002davinci-002Text and reasoningReplacement for the GPT-3 curie and davinci base models
GPT-3.5 Turbogpt-3.5-turboText and reasoningLegacy GPT model for cheaper chat and non-chat tasks
gpt-3.5-turbo-instructgpt-3.5-turbo-instructText and reasoningOpenAI model available through OpenAI.
GPT-4gpt-4Text and reasoningAn older high-intelligence GPT model
GPT-4 Turbogpt-4-turboText and reasoningAn older high-intelligence GPT model
GPT-4.1gpt-4.1Text and reasoningSmartest non-reasoning model
GPT-4.1 minigpt-4.1-miniText and reasoningSmaller, faster version of GPT-4.1
GPT-4.1 nanogpt-4.1-nanoText and reasoningFastest, most cost-efficient version of GPT-4.1
GPT-4ogpt-4oText and reasoningFast, intelligent, flexible GPT model
GPT-4o minigpt-4o-miniText and reasoningFast, affordable small model for focused tasks
GPT-5gpt-5Text and reasoningPrevious intelligent reasoning model for coding and agentic tasks with configurable reasoning effort
GPT-5 minigpt-5-miniText and reasoningNear-frontier intelligence for cost sensitive, low latency, high volume workloads
GPT-5 nanogpt-5-nanoText and reasoningFastest, most cost-efficient version of GPT-5
GPT-5 Progpt-5-proText and reasoningVersion of GPT-5 that produces smarter and more precise responses
GPT-5.1gpt-5.1Text and reasoningThe best model for coding and agentic tasks with configurable reasoning effort
GPT-5.2gpt-5.2Text and reasoningPrevious frontier model for professional work with configurable reasoning effort
GPT-5.2 Progpt-5.2-proText and reasoningPrevious pro model for professional work that produces smarter and more precise responses.
GPT-5.3-Codexgpt-5.3-codexText and reasoningThe most capable agentic coding model to date.
GPT-5.4gpt-5.4Text and reasoningA more affordable model for coding and professional work.
GPT-5.4 minigpt-5.4-miniText and reasoningOur strongest mini model yet for coding, computer use, and subagents
GPT-5.4 nanogpt-5.4-nanoText and reasoningOur cheapest GPT-5.4-class model for simple high-volume tasks
GPT-5.4 Progpt-5.4-proText and reasoningVersion of GPT-5.4 that produces smarter and more precise responses.
GPT-5.5gpt-5.5Text and reasoningA new class of intelligence for coding and professional work.
GPT-5.5 Progpt-5.5-proText and reasoningVersion of GPT-5.5 that produces smarter and more precise responses.
GPT-5.6 Lunagpt-5.6-lunaText and reasoningGPT-5.6 model optimized for cost-sensitive workloads
GPT-5.6 Terragpt-5.6-terraText and reasoningGPT-5.6 model that balances intelligence and cost
GPT-5.6 Solgpt-5.6-solText and reasoningFrontier model for complex professional work
gpt-audiogpt-audioText and reasoningFor audio inputs and outputs with Chat Completions API
gpt-audio-1.5gpt-audio-1.5Text and reasoningThe best voice model for audio in, audio out with Chat Completions.
gpt-audio-minigpt-audio-miniText and reasoningA cost-efficient version of GPT Audio
o1o1Text and reasoningPrevious full o-series reasoning model
o1-proo1-proText and reasoningVersion of o1 with more compute for better responses
o3o3Text and reasoningReasoning model for complex tasks, succeeded by GPT-5
o3-minio3-miniText and reasoningA small model alternative to o3
o3-proo3-proText and reasoningVersion of o3 with more compute for better responses
o4-minio4-miniText and reasoningFast, cost-efficient reasoning model, succeeded by GPT-5 mini
GPT-6 Astragpt-6-astraText and reasoningAdvanced model for complex reasoning, coding, research, and document workflows.
GPT-6 Lunagpt-6-lunaText and reasoningEfficient reasoning model for focused, high-volume tasks.
gpt-3.5-turbo-0125gpt-3.5-turbo-0125Text and reasoningDocumented API model: gpt-3.5-turbo-0125. Alias or version of gpt-3.5-turbo.
gpt-3.5-turbo-1106gpt-3.5-turbo-1106Text and reasoningDocumented API model: gpt-3.5-turbo-1106. Alias or version of gpt-3.5-turbo.
gpt-4-turbo-2024-04-09gpt-4-turbo-2024-04-09Text and reasoningDocumented API model: gpt-4-turbo-2024-04-09. Alias or version of gpt-4-turbo.
gpt-4.1-mini-2025-04-14gpt-4.1-mini-2025-04-14Text and reasoningDocumented API model: gpt-4.1-mini-2025-04-14. Alias or version of gpt-4.1-mini.
gpt-4.1-nano-2025-04-14gpt-4.1-nano-2025-04-14Text and reasoningDocumented API model: gpt-4.1-nano-2025-04-14. Alias or version of gpt-4.1-nano.
gpt-4.1-2025-04-14gpt-4.1-2025-04-14Text and reasoningDocumented API model: gpt-4.1-2025-04-14. Alias or version of gpt-4.1.
gpt-4-0613gpt-4-0613Text and reasoningDocumented API model: gpt-4-0613. Alias or version of gpt-4.
gpt-4o-mini-2024-07-18gpt-4o-mini-2024-07-18Text and reasoningDocumented API model: gpt-4o-mini-2024-07-18. Alias or version of gpt-4o-mini.
gpt-4o-2024-11-20gpt-4o-2024-11-20Text and reasoningDocumented API model: gpt-4o-2024-11-20. Alias or version of gpt-4o.
gpt-4o-2024-08-06gpt-4o-2024-08-06Text and reasoningDocumented API model: gpt-4o-2024-08-06. Alias or version of gpt-4o.
gpt-4o-2024-05-13gpt-4o-2024-05-13Text and reasoningDocumented API model: gpt-4o-2024-05-13. Alias or version of gpt-4o.
gpt-5-mini-2025-08-07gpt-5-mini-2025-08-07Text and reasoningDocumented API model: gpt-5-mini-2025-08-07. Alias or version of gpt-5-mini.
gpt-5-nano-2025-08-07gpt-5-nano-2025-08-07Text and reasoningDocumented API model: gpt-5-nano-2025-08-07. Alias or version of gpt-5-nano.
gpt-5-pro-2025-10-06gpt-5-pro-2025-10-06Text and reasoningDocumented API model: gpt-5-pro-2025-10-06. Alias or version of gpt-5-pro.
gpt-5.1-2025-11-13gpt-5.1-2025-11-13Text and reasoningDocumented API model: gpt-5.1-2025-11-13. Alias or version of gpt-5.1.
gpt-5.2-pro-2025-12-11gpt-5.2-pro-2025-12-11Text and reasoningDocumented API model: gpt-5.2-pro-2025-12-11. Alias or version of gpt-5.2-pro.
gpt-5.2-2025-12-11gpt-5.2-2025-12-11Text and reasoningDocumented API model: gpt-5.2-2025-12-11. Alias or version of gpt-5.2.
gpt-5.4-mini-2026-03-17gpt-5.4-mini-2026-03-17Text and reasoningDocumented API model: gpt-5.4-mini-2026-03-17. Alias or version of gpt-5.4-mini.
gpt-5.4-nano-2026-03-17gpt-5.4-nano-2026-03-17Text and reasoningDocumented API model: gpt-5.4-nano-2026-03-17. Alias or version of gpt-5.4-nano.
gpt-5.4-pro-2026-03-05gpt-5.4-pro-2026-03-05Text and reasoningDocumented API model: gpt-5.4-pro-2026-03-05. Alias or version of gpt-5.4-pro.
gpt-5.4-2026-03-05gpt-5.4-2026-03-05Text and reasoningDocumented API model: gpt-5.4-2026-03-05. Alias or version of gpt-5.4.
gpt-5.5-pro-2026-04-23gpt-5.5-pro-2026-04-23Text and reasoningDocumented API model: gpt-5.5-pro-2026-04-23. Alias or version of gpt-5.5-pro.
gpt-5.5-2026-04-23gpt-5.5-2026-04-23Text and reasoningDocumented API model: gpt-5.5-2026-04-23. Alias or version of gpt-5.5.
gpt-5.6-cybergpt-5.6-cyberText and reasoningRestricted API access: requires separate Daybreak approval and provisioning.
gpt-5-2025-08-07gpt-5-2025-08-07Text and reasoningDocumented API model: gpt-5-2025-08-07. Alias or version of gpt-5.
gpt-audio-mini-2025-12-15gpt-audio-mini-2025-12-15Text and reasoningDocumented API model: gpt-audio-mini-2025-12-15. Alias or version of gpt-audio-mini.
gpt-audio-2025-08-28gpt-audio-2025-08-28Text and reasoningDocumented API model: gpt-audio-2025-08-28. Alias or version of gpt-audio.
gpt-daybreak-blue-latestgpt-daybreak-blue-latestText and reasoningRestricted API access: requires separate Daybreak approval and provisioning.
gpt-daybreak-red-latestgpt-daybreak-red-latestText and reasoningRestricted API access: requires separate Daybreak approval and provisioning.
o1-pro-2025-03-19o1-pro-2025-03-19Text and reasoningDocumented API model: o1-pro-2025-03-19. Alias or version of o1-pro.
o1-2024-12-17o1-2024-12-17Text and reasoningDocumented API model: o1-2024-12-17. Alias or version of o1.
o3-mini-2025-01-31o3-mini-2025-01-31Text and reasoningDocumented API model: o3-mini-2025-01-31. Alias or version of o3-mini.
o3-pro-2025-06-10o3-pro-2025-06-10Text and reasoningDocumented API model: o3-pro-2025-06-10. Alias or version of o3-pro.
o3-2025-04-16o3-2025-04-16Text and reasoningDocumented API model: o3-2025-04-16. Alias or version of o3.
o4-mini-2025-04-16o4-mini-2025-04-16Text and reasoningDocumented API model: o4-mini-2025-04-16. Alias or version of o4-mini.
gpt-5.6gpt-5.6Text and reasoningDocumented API model: gpt-5.6. Alias or version of gpt-5.6-sol.
gpt-3.5-turbo-completionsgpt-3.5-turbo-completionsText and reasoningLegacy Completions API alias. Scheduled shutdown: 2026-10-23.
gpt-4-0613-completionsgpt-4-0613-completionsText and reasoningLegacy Completions API alias. Scheduled shutdown: 2026-10-23.
gpt-4-completionsgpt-4-completionsText and reasoningLegacy Completions API alias. Scheduled shutdown: 2026-10-23.
gpt-4-turbo-completionsgpt-4-turbo-completionsText and reasoningLegacy Completions API alias. Scheduled shutdown: 2026-10-23.
gpt-5-search-apigpt-5-search-apiText and reasoningSearch-specialized Chat Completions API model.
gpt-5.5-cybergpt-5.5-cyberText and reasoningRestricted Daybreak API model; account provisioning required.
gpt-5.4-cybergpt-5.4-cyberText and reasoningRestricted Daybreak API model; account provisioning required.
gpt-4-1106-previewgpt-4-1106-previewText and reasoningLegacy API snapshot. Latest lifecycle notice lists 2026-10-23 shutdown; an older notice conflicts, so account availability needs verification.
chatgpt-image-latestchatgpt-image-latestImagesPrevious image model used in ChatGPT.
GPT Image 1gpt-image-1ImagesOur previous image generation model
gpt-image-1-minigpt-image-1-miniImagesA cost-efficient version of GPT Image 1
GPT Image 1.5gpt-image-1.5ImagesOur previous image generation model
GPT Image 2.5 Flaregpt-image-2.5-flareImagesFast image generation and editing with GPT Image 2.5
GPT Image 2.5 Sunburstgpt-image-2.5-sunburstImagesImage generation and precise editing with GPT Image 2.5
GPT Image 2gpt-image-2ImagesState-of-the-art image generation model
gpt-image-1.5-2025-12-16gpt-image-1.5-2025-12-16ImagesDocumented API model: gpt-image-1.5-2025-12-16. Alias or version of gpt-image-1.5.
gpt-image-2-2026-04-21gpt-image-2-2026-04-21ImagesDocumented API model: gpt-image-2-2026-04-21. Alias or version of gpt-image-2.
sora-2-prosora-2-proVideoDocumented API model: sora-2-pro.
sora-2-pro-2025-10-06sora-2-pro-2025-10-06VideoDocumented API model: sora-2-pro-2025-10-06. Alias or version of sora-2-pro.
sora-2sora-2VideoDocumented API model: sora-2.
sora-2-2025-12-08sora-2-2025-12-08VideoDocumented API model: sora-2-2025-12-08. Alias or version of sora-2.
sora-2-2025-10-06sora-2-2025-10-06VideoDocumented API model: sora-2-2025-10-06. Alias or version of sora-2.
GPT-Realtimegpt-realtimeRealtimeModel capable of realtime text and audio inputs and outputs
GPT-Realtime-1.5gpt-realtime-1.5RealtimeThe best voice model for audio in, audio out
GPT-Realtime-2gpt-realtime-2RealtimeReasoning model with tool use
GPT-Realtime-2.1gpt-realtime-2.1RealtimeReasoning model with tool use
GPT-Realtime-2.1 minigpt-realtime-2.1-miniRealtimeReasoning model with tool use
GPT-Realtime minigpt-realtime-miniRealtimeA cost-efficient version of GPT-Realtime
GPT-Realtime-Translategpt-realtime-translateRealtimeStreaming speech-to-speech translation model
gpt-realtime-mini-2025-12-15gpt-realtime-mini-2025-12-15RealtimeDocumented API model: gpt-realtime-mini-2025-12-15. Alias or version of gpt-realtime-mini.
gpt-realtime-2025-08-28gpt-realtime-2025-08-28RealtimeDocumented API model: gpt-realtime-2025-08-28. Alias or version of gpt-realtime.
text-embedding-3-largetext-embedding-3-largeEmbeddingsMost capable embedding model
text-embedding-3-smalltext-embedding-3-smallEmbeddingsSmall embedding model
text-embedding-ada-002text-embedding-ada-002EmbeddingsOlder embedding model
GPT-4o mini TTSgpt-4o-mini-ttsSpeech generationText-to-speech model powered by GPT-4o mini
TTS-1tts-1Speech generationText-to-speech model optimized for speed
TTS-1 HDtts-1-hdSpeech generationText-to-speech model optimized for quality
gpt-4o-mini-tts-2025-03-20gpt-4o-mini-tts-2025-03-20Speech generationDocumented API model: gpt-4o-mini-tts-2025-03-20. Alias or version of gpt-4o-mini-tts.
gpt-4o-mini-tts-2025-12-15gpt-4o-mini-tts-2025-12-15Speech generationDocumented API model: gpt-4o-mini-tts-2025-12-15. Alias or version of gpt-4o-mini-tts.
GPT-4o mini Transcribegpt-4o-mini-transcribeSpeech recognitionSpeech-to-text model powered by GPT-4o mini
GPT-4o Transcribegpt-4o-transcribeSpeech recognitionSpeech-to-text model powered by GPT-4o
GPT-4o Transcribe Diarizegpt-4o-transcribe-diarizeSpeech recognitionTranscription model that identifies who's speaking when
GPT Live Transcribegpt-live-transcribeSpeech recognitionLow-latency streaming speech-to-text model for live audio.
GPT-Realtime-Whispergpt-realtime-whisperSpeech recognitionStreaming speech-to-text model for realtime transcription
GPT Transcribegpt-transcribeSpeech recognitionHigh-accuracy speech-to-text model for files and realtime input transcription.
Whisperwhisper-1Speech recognitionGeneral-purpose speech recognition model
gpt-4o-mini-transcribe-2025-03-20gpt-4o-mini-transcribe-2025-03-20Speech recognitionDocumented API model: gpt-4o-mini-transcribe-2025-03-20. Alias or version of gpt-4o-mini-transcribe.
gpt-4o-mini-transcribe-2025-12-15gpt-4o-mini-transcribe-2025-12-15Speech recognitionDocumented API model: gpt-4o-mini-transcribe-2025-12-15. Alias or version of gpt-4o-mini-transcribe.
omni-moderationomni-moderation-latestModerationIdentify potentially harmful content in text and images
omni-moderation-2024-09-26omni-moderation-2024-09-26ModerationDocumented API model: omni-moderation-2024-09-26. Alias or version of omni-moderation-latest.