Qwen: Qwen-Plus
qwen
Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.
- Context window
- 1,000,000 tokens
- Input cost
- —
- Output cost
- —
- Latency (p50)
- 20 ms
Loading...
Run side-by-side checks for pricing, context window, and latency.
Price (lower is better), Speed/Latency (lower is better), Performance (higher is better)
💰Qwen: Qwen-Plus is 100% cheaper than Qwen: Qwen2.5 Coder 7B Instruct
⚡Qwen: Qwen2.5 Coder 7B Instruct is 20ms faster than Qwen: Qwen-Plus
ilsp
Open-source CoRM model from ilsp — available for download and self-hosting on Hugging Face.
dr-housemd
Open-source GLM-4.7-Flash-exl3-3bpw-H4 model from dr-housemd — available for download and self-hosting on Hugging Face.
redhatai
Open-source gpt-oss-120b-essential model from redhatai — available for download and self-hosting on Hugging Face.
Let our experts help you evaluate models based on your actual use case and budget. We'll provide a free analysis and actionable recommendations.
qwen
Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.
qwen
Qwen2.5-Coder-7B-Instruct is a 7B parameter instruction-tuned language model optimized for code-related tasks such as code generation, reasoning, and bug fixing. Based on the Qwen2.5 architecture, it incorporates enhancements like RoPE,...