Loading...
Loading...
4,695+ models, every major provider, live pricing — one search.
4,695
models tracked daily
260
providers indexed
Price pulse
Sponsored · Get featured
The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored writing to improve relevance & readability. It’s also better at working with uploaded...
Deploy Now →openai
OpenAI: GPT-5.4 ProGPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
Deploy Now →Jump straight to the models that match your project.
Recently added and trending across providers.
Open-source Laguna-S-2.1-Q2K-pulsar model from giannisan — available for download and self-hosting on Hugging Face.
Open-source Solar-Open2-250B-MLX-6bit model from vontra — available for download and self-hosting on Hugging Face.
Other sites list models. We help you choose the right one — with real data, not marketing copy.
Live pricing — always accurate
We sync directly from provider APIs every day. No stale rates. No manual updates. The price you see is the price you pay.
Every provider. One search.
OpenAI, Anthropic, Mistral, Groq, Replicate, Hugging Face, Ollama and more — all indexed in one place so you stop switching tabs.
Real benchmark scores
LiveBench, LMSYS Chatbot Arena, MMLU — surfaced per model so you can make evidence-based decisions, not marketing-driven ones.
Price alerts & Pulse
Get notified when a model you care about drops in price, releases a new version, or hits the news — before your competitors do.
Side-by-side comparison
Compare any two models on price, context window, benchmarks and features with one click. Decide faster. Switch smarter.
Price changes, launches, and model news — in real time.
Cursor's agent swarm suggests cheaper models can handle most coding when frontier models plan the work
7/26/2026
Anthropic's Opus 5 blows past Fable 5 and GPT-5.6 Sol on the benchmark designed to measure real intelligence
7/26/2026
bge-large-en-v1.5 pricing increased
7/26/2026
qwq-32b pricing increased
7/26/2026
Search, filter, compare and calculate costs — all free, no sign-up required.
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
Long Context
100K+ token windows
Pro Tools
Enter your token volumes and see projected monthly costs per model. Know what you'll spend before you commit.