Loading...
Loading...
Side-by-side comparison of API pricing, context window, latency, and benchmark performance. Data refreshed daily from provider APIs and ModelStop's own latency probes.
💰Meta: Llama 3.3 70B Instruct (free) is 100% cheaper than Amazon Nova Micro
⚡Amazon Nova Micro is 37ms faster than Meta: Llama 3.3 70B Instruct (free)
Price (lower is better), Speed/Latency (lower is better), Performance (higher is better)
| Spec | Amazon Nova Micro | Meta: Llama 3.3 70B Instruct (free) |
|---|---|---|
| Provider | amazon | meta-llama |
| Input price / 1M tokens | $0.04 | Free |
| Output price / 1M tokens | $0.14 | — |
| Context window | 128k tokens | 131k tokens |
| Latency (p50) | — | 37ms |
| Best benchmark score | — | — |
| Open source | No | No |
Amazon Nova Micro is the fastest and most cost-effective text-only model in the Nova family, optimized for speed and low latency. Ideal for customer service, summarization, and translation at scale.
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...
Want to swap in a different model or estimate monthly costs?