Loading...
Browse 464 models across providers, modalities, and use cases.
📄 Long Context
464 models · Page 1 of 13
Meta's 70B parameter model balancing quality and speed.
Anthropic's standard model — optimal balance of cost and capability for most workloads.
Open-source Laguna-XS-2.1 model from poolside — available for download and self-hosting on Hugging Face.
Building upon Mistral Small 3 (2501), Mistral Small 3.1 (2503) adds state-of-the-art vision understanding and enhances long context capabilities up to 128k tokens without compromising text performance. With 24 billion parameters, this model achieves top-tier capabilities in both text and vision tasks.
The Llama 3.2-Vision instruction-tuned models are optimized for visual recognition, image reasoning, captioning, and answering general questions about an image.
BAAI general embedding (Base) model that transforms any given text into a 768-dimensional vector
SEA-LION stands for Southeast Asian Languages In One Network, which is a collection of Large Language Models (LLMs) which have been pretrained and instruct-tuned for the Southeast Asia (SEA) region.
Gemma 4 is Google's most intelligent family of open models, built from Gemini 3 research to maximize intelligence-per-parameter.
OpenAI’s open-weight models designed for powerful reasoning, agentic tasks, and versatile developer use cases – gpt-oss-20b is for lower latency, and local or specialized use-cases.
Meta's Llama 4 Scout is a 17 billion parameter model with 16 experts that is natively multimodal. These models leverage a mixture-of-experts architecture to offer industry-leading performance in text and image understanding.
Kimi K2.5 is a frontier-scale open-source model with a 256k context window, multi-turn tool calling, vision inputs, and structured outputs for agentic workloads.
Llama Guard 3 is a Llama-3.1-8B pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM inputs (prompt classification) and in LLM responses (response classification). It acts as an LLM – it generates text in its output that indicates whether a given prompt or response is safe or unsafe, and if unsafe, it also lists the content categories violated.
GLM-4.7-Flash is a fast and efficient multilingual text generation model with a 131,072 token context window. Optimized for dialogue, instruction-following, and multi-turn tool calling across 100+ languages.
Granite 4.0 instruct models deliver strong performance across benchmarks, achieving industry-leading results in key agentic tasks like instruction following and function calling. These efficiencies make the models well-suited for a wide range of use cases like retrieval-augmented generation (RAG), multi-agent workflows, and edge deployments.
NVIDIA Nemotron 3 Super is a hybrid MoE model with leading accuracy for multi-agent applications and specialized agentic AI systems.
Kimi K2.6 is a frontier-scale open-source 1T parameter model with a 262.1k context window, multi-turn tool calling, vision inputs, and structured outputs for agentic workloads.
OpenAI’s open-weight models designed for powerful reasoning, agentic tasks, and versatile developer use cases – gpt-oss-120b is for production, general purpose, high reasoning use-cases.
Gemma 3 4B IT — available via AWS Bedrock (us-east-1).
Gemma 3 27B PT — available via AWS Bedrock (us-east-1).
GPT OSS Safeguard 20B — available via AWS Bedrock (us-east-1).
Open-source Qwen3-VL-30B-A3B-Instruct model from qwen — available for download and self-hosting on Hugging Face.
Open-source Qwen3.5-9B model from qwen — available for download and self-hosting on Hugging Face.
Open-source Qwen3-VL-8B-Instruct model from qwen — available for download and self-hosting on Hugging Face.
Open-source Qwen3.5-35B-A3B model from qwen — available for download and self-hosting on Hugging Face.
Open-source Qwen3-VL-32B-Instruct model from qwen — available for download and self-hosting on Hugging Face.
Open-source Qwen3.5-27B model from qwen — available for download and self-hosting on Hugging Face.
Open-source Llama-3.2-11B-Vision-Instruct model from meta-llama — available for download and self-hosting on Hugging Face.
Open-source Llama-Guard-3-8B model from meta-llama — available for download and self-hosting on Hugging Face.
Open-source Llama-Guard-4-12B model from meta-llama — available for download and self-hosting on Hugging Face.
Open-source Qwen3-Coder-Next model from qwen — available for download and self-hosting on Hugging Face.
Open-source Llama-3.3-70B-Instruct model from meta-llama — available for download and self-hosting on Hugging Face.
Open-source Qwen3-235B-A22B model from qwen — available for download and self-hosting on Hugging Face.