Qwen3.5-9B
qwen
Open-source Qwen3.5-9B model from qwen — available for download and self-hosting on Hugging Face.
- Context window
- 262,144 tokens
- Input cost
- $0.10 / 1M
- Output cost
- $0.15 / 1M
- Latency (p50)
- —
Loading...
Run side-by-side checks for pricing, context window, and latency.
Price (lower is better), Speed/Latency (lower is better), Performance (higher is better)
Qwen: Qwen2.5 Coder 7B Instruct is 233% cheaper
💰Qwen: Qwen2.5 Coder 7B Instruct is 70% cheaper than Qwen3.5-9B
madras1
Open-source Yara-50M-PTBR-QA model from madras1 — available for download and self-hosting on Hugging Face.
manuel121357
Open-source Dolphin3.0-Llama3.1-8B-q4f16_1-MLC model from manuel121357 — available for download and self-hosting on Hugging Face.
taniota
Open-source Warlock-1.7B model from taniota — available for download and self-hosting on Hugging Face.
Let our experts help you evaluate models based on your actual use case and budget. We'll provide a free analysis and actionable recommendations.
qwen
Open-source Qwen3.5-9B model from qwen — available for download and self-hosting on Hugging Face.
qwen
Qwen2.5-Coder-7B-Instruct is a 7B parameter instruction-tuned language model optimized for code-related tasks such as code generation, reasoning, and bug fixing. Based on the Qwen2.5 architecture, it incorporates enhancements like RoPE,...