Analysis Summary
GPT Audio Mini is an audio-capable model supporting text and audio input and output, with tool use, function calling, and a 128K context window. Pricing is $0.60 per million input tokens and $2.40 per million output tokens. Its modality is the clearest reason to consider it, especially for voice-first interfaces and conversational media tasks.
Suitable workloads include audio transcription exchanges, voice assistants, spoken content workflows, and tool-connected customer service prototypes. The supplied data contains no reasoning, coding, content, or agentic benchmark results, so capability outside audio interaction is not established. Use it where bidirectional audio is essential, with transcript and response quality checks before deploying it in customer-facing workflows.
Assessed September 7, 2026
Editorial notes
GPT Audio Mini supports text and audio exchange, function calling, and a 128K context window, making it useful for voice workflows, but no benchmark data is supplied for broader capability assessment.
Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?
DFO Verdict
GPT Audio Mini supports text and audio exchange, function calling, and a 128K context window, making it useful for voice workflows, but no benchmark data is supplied for broader capability assessment.
How OpenAI: GPT Audio Mini compares
Its 128K-token context window is larger than 35% of the models we list. At $0.60 per million input tokens it is cheaper than 36% of comparable models.
Dark bar = input · light bar = output, scaled to the priciest peer.
1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.
Strongest on business fit. The pulled-in technical corner is the trade-off, and if the shape matters more than the price, this is your model.
Compare shapes side-by-side →Pricing
| Token Type | Cost per 1M tokens | Cost per 1K tokens |
|---|---|---|
| Input | $0.60 | $0.000600 |
| Output | $2.40 | $0.002400 |
What would OpenAI: GPT Audio Mini cost your business?
Pick the job that looks most like yours, then fine-tune with the sliders. Estimates update live.
A website chatbot handling around 100 customer conversations a day, a few short messages each.
Full calculator with 756 models → Price Calculator
These numbers get smaller with the right architecture.
We route routine calls to cheap models and save OpenAI: GPT Audio Mini for the hard ones. Most clients cut their estimate by 60-80%.
Talk to our teamAbout OpenAI: GPT Audio Mini
A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million..
Explore Related Models
Frequently asked questions about OpenAI: GPT Audio Mini
How much does OpenAI: GPT Audio Mini cost?
OpenAI: GPT Audio Mini costs $0.60 per million input tokens and $2.40 per million output tokens.
What is the context window of OpenAI: GPT Audio Mini?
OpenAI: GPT Audio Mini has a context window of 128,000 tokens (128K).
What can OpenAI: GPT Audio Mini do?
OpenAI: GPT Audio Mini supports tool use and function calling.
Who created OpenAI: GPT Audio Mini?
OpenAI: GPT Audio Mini is developed by OpenAI and was released on January 19, 2026.
Data sourced from the OpenRouter API, Artificial Analysis, the Hugging Face Open LLM Leaderboard and our own internal testing. Scores are editorially curated by our team.
Last updated: September 7, 2026 8:38 pm