microsoft/phi-4
deepinfra
Microsoft Phi-4 14B — small language model achieving state-of-the-art results on reasoning tasks.
- Context window
- 16,384 tokens
- Input cost
- —
- Output cost
- —
- Latency (p50)
- 70 ms
Loading...
Run side-by-side checks for pricing, context window, and latency.
Price (lower is better), Speed/Latency (lower is better), Performance (higher is better)
💰microsoft/phi-4 is 100% cheaper than Google: Gemma 2 9B
⚡Google: Gemma 2 9B is 70ms faster than microsoft/phi-4
giannisan
Open-source Laguna-S-2.1-Q2K-pulsar model from giannisan — available for download and self-hosting on Hugging Face.
vontra
Open-source Solar-Open2-250B-MLX-6bit model from vontra — available for download and self-hosting on Hugging Face.
atomicchat
Open-source ornith-35b-MLX-6bit model from atomicchat — available for download and self-hosting on Hugging Face.
Let our experts help you evaluate models based on your actual use case and budget. We'll provide a free analysis and actionable recommendations.
deepinfra
Microsoft Phi-4 14B — small language model achieving state-of-the-art results on reasoning tasks.
Gemma 2 9B by Google is an advanced, open-source language model that sets a new standard for efficiency and performance in its size class. Designed for a wide variety of...