AI Model Catalogue
Browse 4,151 models across providers, modalities, and use cases.
🌐 All Models
4,151 models · Page 116 of 116
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
textreasoningagents
Run locally
Input$0.0390/1M
Output$0.1900/1M
Explore specs and pricing
NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...
textreasoningfree
Run locally
Explore specs and pricing
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
textcodereasoning
Run locally
Input$0.0900/1M
Output$1.1000/1M
Explore specs and pricing
NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...
textvisionmultimodal
Run locally
Input$0.2000/1M
Output$0.6000/1M
Explore specs and pricing
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
textagentsfree
Run locally
Input$0.0500/1M
Output$0.2000/1M
Explore specs and pricing
The simplest way to get free inference. openrouter/free is a router that selects free models at random from the models available on OpenRouter. The router smartly filters for models that...
textvisionmultimodal
Run locally
Explore specs and pricing
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...
textagentsfree
Run locally
Input$0.1000/1M
Output$0.5000/1M
Explore specs and pricing
30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate...
textvisionimage
Run locally
InputFree
Output$0.0000/1M
Explore specs and pricing
Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...
textvisionimage
Run locally
InputFree
Output$0.0000/1M
Explore specs and pricing
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
textvisionmultimodal
Run locally
Input$0.1400/1M
Output$0.4000/1M
Explore specs and pricing
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
textvisionmultimodal
Run locally
Input$0.1200/1M
Output$0.4000/1M
Explore specs and pricing