Model Terminal
GPT Audio
GPT Audio is OpenAI's natively multimodal audio model for the Chat Completions API, accepting audio and text as input and producing audio and text as output. It is designed for asynchronous spoken interaction use cases such as voice summarization, audio sentiment analysis, and turn-based audio conversation. Source: supabase.
Value score
Not scored
Context
—
tokens
Max output
—
tokens
Price
—
Capability radar
Peer value bars
Identity
- Developer
- Udio
- Openness
- —
- Modalities
- —
- Release
- —
- Knowledge cutoff
- —
- Deprecation
- —
- API docs
- —
Benchmarks
No benchmark scores yet.
Compare nearby
Pricing
- Input / 1M
- —
- Output / 1M
- —
- Speed
- —