Loading...
Browse 6,951 models across providers, modalities, and use cases.
🌐 All Models
6,951 models · Page 176 of 194
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
Mixtral 8×7B Instruct on DeepInfra — popular MoE model with 32K context and strong multilingual performance.