gpt-3.5-turbo
openai
OpenAI's lightweight and fast GPT-3.5 Turbo — cost-effective for most applications.
- Context window
- 16,385 tokens
- Input cost
- —
- Output cost
- —
- Latency (p50)
- 787 ms
Loading...
Run side-by-side checks for pricing, context window, and latency.
Price (lower is better), Speed/Latency (lower is better), Performance (higher is better)
💰gpt-3.5-turbo is 100% cheaper than OpenAI: GPT-4 Turbo (older v1106)
⚡OpenAI: GPT-4 Turbo (older v1106) is 787ms faster than gpt-3.5-turbo
madras1
Open-source Yara-50M-PTBR-QA model from madras1 — available for download and self-hosting on Hugging Face.
manuel121357
Open-source Dolphin3.0-Llama3.1-8B-q4f16_1-MLC model from manuel121357 — available for download and self-hosting on Hugging Face.
taniota
Open-source Warlock-1.7B model from taniota — available for download and self-hosting on Hugging Face.
Let our experts help you evaluate models based on your actual use case and budget. We'll provide a free analysis and actionable recommendations.
openai
OpenAI's lightweight and fast GPT-3.5 Turbo — cost-effective for most applications.
openai
The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to April 2023.