Loading...
Loading...
The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...
Best for
Local run command
# Refer to provider docs:
Presumptive Specs:
Typical GPU
12–24 GB VRAM
System RAM
16+ GB
Disk
20+ GB free
Runtime
Docker / local runtime
Community reliability estimate · not official
About this score: Community-estimated based on user reports and publicly available benchmark data (e.g. TruthfulQA). This is not an official score from the model provider. Scores may be inaccurate — always verify with the official leaderboard before making production decisions.
Not enough historical data yet. Check back after the next pricing sync.
Provider
Learn how to use OpenAI: GPT Audio
💡 Tip: Start with the sample prompts above to see how OpenAI: GPT Audio works best.
Proven prompts shared by the community for this model