gpt-3.5-turbo
openai
OpenAI's lightweight and fast GPT-3.5 Turbo — cost-effective for most applications.
- Context window
- 16,385 tokens
- Input cost
- —
- Output cost
- —
- Latency (p50)
- 787 ms
Loading...
Run side-by-side checks for pricing, context window, and latency.
Price (lower is better), Speed/Latency (lower is better), Performance (higher is better)
💰gpt-3.5-turbo is 100% cheaper than OpenAI: GPT-4 (older v0314)
⚡OpenAI: GPT-4 (older v0314) is 787ms faster than gpt-3.5-turbo
📈OpenAI: GPT-4 (older v0314) scores 1270.0 pts higher on benchmarks
madras1
Open-source Yara-50M-PTBR-QA model from madras1 — available for download and self-hosting on Hugging Face.
manuel121357
Open-source Dolphin3.0-Llama3.1-8B-q4f16_1-MLC model from manuel121357 — available for download and self-hosting on Hugging Face.
taniota
Open-source Warlock-1.7B model from taniota — available for download and self-hosting on Hugging Face.
Let our experts help you evaluate models based on your actual use case and budget. We'll provide a free analysis and actionable recommendations.
openai
OpenAI's lightweight and fast GPT-3.5 Turbo — cost-effective for most applications.
openai
GPT-4-0314 is the first version of GPT-4 released, with a context length of 8,192 tokens, and was supported until June 14. Training data: up to Sep 2021.