Qwen/Qwen2.5-72B-Instruct
deepinfra
Alibaba Qwen2.5 72B Instruct on DeepInfra — top open-source model with multilingual and coding strengths.
- Context window
- — tokens
- Input cost
- —
- Output cost
- —
- Latency (p50)
- 99 ms
Loading...
Run side-by-side checks for pricing, context window, and latency.
Price (lower is better), Speed/Latency (lower is better), Performance (higher is better)
💰Qwen/Qwen2.5-72B-Instruct is 100% cheaper than Google: Gemma 2 9B
⚡Google: Gemma 2 9B is 99ms faster than Qwen/Qwen2.5-72B-Instruct
ilsp
Open-source CoRM model from ilsp — available for download and self-hosting on Hugging Face.
dr-housemd
Open-source GLM-4.7-Flash-exl3-3bpw-H4 model from dr-housemd — available for download and self-hosting on Hugging Face.
redhatai
Open-source gpt-oss-120b-essential model from redhatai — available for download and self-hosting on Hugging Face.
Let our experts help you evaluate models based on your actual use case and budget. We'll provide a free analysis and actionable recommendations.
deepinfra
Alibaba Qwen2.5 72B Instruct on DeepInfra — top open-source model with multilingual and coding strengths.
Gemma 2 9B by Google is an advanced, open-source language model that sets a new standard for efficiency and performance in its size class. Designed for a wide variety of...