AI Model Catalogue
Browse 17 models across providers, modalities, and use cases.
☁️
Amazon
🧠 Reasoning
17 models · Page 1 of 1
Gemma 3 models are well-suited for a variety of text generation and image understanding tasks, including question answering, summarization, and reasoning. Gemma 3 models are multimodal, handling text and image input and generating text output, with a large, 128K context window, multilingual support in over 140 languages, and is available in more sizes than previous versions.
textvisionreasoning
Input$0.3500/1M
Output$0.5600/1M
Explore specs and pricing
Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
textvisionreasoning
Run locally
Explore specs and pricing
Google’s hybrid “thinking” AI model optimized for speed and cost-efficiency
textreasoningfree
Run locally
Explore specs and pricing
Google's most intelligent model, with improved reasoning and a new medium thinking level
textreasoningfree
Run locally
Explore specs and pricing
Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
textvisionmultimodal
Run locally
Input$0.0800/1M
Output$0.1600/1M
Explore specs and pricing
Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
textvisionmultimodal
Run locally
Input$0.0400/1M
Output$0.1300/1M
Explore specs and pricing
Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
textvisionmultimodal
Run locally
Input$0.0400/1M
Output$0.0800/1M
Explore specs and pricing
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
textvisionmultimodal
Run locally
Explore specs and pricing
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
textvisionmultimodal
Run locally
Explore specs and pricing
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
textvisionmultimodal
Run locally
Explore specs and pricing
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
textvisionmultimodal
Run locally
Explore specs and pricing
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
textvisionimage
Input$0.1000/1M
Output$0.4000/1M
Explore specs and pricing
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
textvisionimage
Run locally
Input$0.1000/1M
Output$0.4000/1M
Explore specs and pricing
Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and...
textvisionimage
Run locally
Input$2.0000/1M
Output$12.0000/1M
Explore specs and pricing
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
textvisionmultimodal
Run locally
Input$0.5000/1M
Output$3.0000/1M
Explore specs and pricing
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
textvisionmultimodal
Run locally
Explore specs and pricing
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
textvisionmultimodal
Run locally
Input$0.1400/1M
Output$0.4000/1M
Explore specs and pricing