AI Model Catalogue
Browse 90 models across providers, modalities, and use cases.
Amazon
🆓 Free & Open
90 models · Page 3 of 3
Upscale images 2x or 4x times
visionfree
Run locally
Explore specs and pricing
Google’s hybrid “thinking” AI model optimized for speed and cost-efficiency
textreasoningfree
Run locally
Explore specs and pricing
Google's fast image generation model with conversational editing, multi-image fusion, and character consistency
visionimagefree
Explore specs and pricing
Google's most intelligent model, with improved reasoning and a new medium thinking level
textreasoningfree
Run locally
Explore specs and pricing
Google's state of the art image generation and editing model 🍌🍌
visionimagefree
Run locally
Explore specs and pricing
New and improved version of Veo 3, with higher-fidelity video, context-aware audio, reference image and last frame support
visionaudiofree
Explore specs and pricing
New and improved version of Veo 3 Fast, with higher-fidelity video, context-aware audio and last frame support
audiofree
Explore specs and pricing
Generate full-length songs up to 3 minutes from text prompts or images with Lyria 3 Pro, Google's most capable music generation model
visionimagefree
Explore specs and pricing
Google's cost-efficient video generation model with native audio, optimized for high-volume applications
audiofree
Explore specs and pricing
Use this ultra version of Imagen 4 when quality matters more than speed and cost
visionfree
Explore specs and pricing
Generate 30-second music clips from text prompts or images with Lyria 3, Google's music generation model
visionimagefree
Explore specs and pricing
Google's fast, expressive text-to-speech model with 30 voices and 70+ language support
textfree
Run locally
Input$5.0000/1M
Output$15.0000/1M
Explore specs and pricing
Open-source gemma-3-1b-it model from google — available for download and self-hosting on Hugging Face.
textfree
Explore specs and pricing
Open-source t5gemma-s-s-prefixlm model from google — available for download and self-hosting on Hugging Face.
textfree
Run locally
InputFree
Output$0.0000/1M
Explore specs and pricing
30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate...
textvisionimage
Run locally
InputFree
Output$0.0000/1M
Explore specs and pricing
Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...
textvisionimage
Run locally
InputFree
Output$0.0000/1M
Explore specs and pricing
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
textvisionmultimodal
Run locally
Input$0.1400/1M
Output$0.4000/1M
Explore specs and pricing
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
textvisionmultimodal
Run locally
Input$0.1200/1M
Output$0.4000/1M
Explore specs and pricing