Loading...
Browse 279 models across providers, modalities, and use cases.
🌐 All Models
279 models · Page 1 of 8
Anthropic's standard model — optimal balance of cost and capability for most workloads.
Open-source Krea2-Loras model from deverstyle — available for download and self-hosting on Hugging Face.
Open-source anima-mlx model from xocialize — available for download and self-hosting on Hugging Face.
Open-source AZ model from az2671 — available for download and self-hosting on Hugging Face.
Open-source 100m_image_new model from eliochampaney — available for download and self-hosting on Hugging Face.
Open-source Zeta-Chroma model from lodestones — available for download and self-hosting on Hugging Face.
UForm-Gen is a small generative vision-language model primarily designed for Image Captioning and Visual Question Answering. The model was pre-trained on the internal image captioning dataset and fine-tuned on public instructions datasets: SVIT, LVIS, VQAs datasets.
FLUX.2 [dev] is an image model from Black Forest Labs where you can generate highly realistic and detailed images, with multi-reference support.
Stable Diffusion is a latent text-to-image diffusion model capable of generating photo-realistic images. Img2img generate a new image from an input image with Stable Diffusion.
Lucid Origin from Leonardo.AI is their most adaptable and prompt-responsive model to date. Whether you're generating images with sharp graphic design, stunning full-HD renders, or highly specific creative direction, it adheres closely to your prompts, renders text with accuracy, and supports a wide array of visual styles and aesthetics – from stylized concept art to crisp product mockups.
FLUX.2 [klein] 9B is a 9 billion parameter model that can generate images from text descriptions and supports multi-reference editing capabilities.
LLaVA is an open-source chatbot trained by fine-tuning LLaMA/Vicuna on GPT-generated multimodal instruction-following data. It is an auto-regressive language model, based on the transformer architecture.
FLUX.1 [schnell] is a 12 billion parameter rectified flow transformer capable of generating images from text descriptions.
SDXL-Lightning is a lightning-fast text-to-image generation model. It can generate high-quality 1024px images in a few steps.
Stable Diffusion model that has been fine-tuned to be better at photorealism without sacrificing range.
Phoenix 1.0 is a model by Leonardo.Ai that generates images with exceptional prompt adherence and coherent text.
Diffusion-based text-to-image generative model by Stability AI. Generates and modify images based on text prompts.
FLUX.2 [klein] is an ultra-fast, distilled image model. It unifies image generation and editing in a single model, delivering state-of-the-art quality enabling interactive workflows, real-time previews, and latency-critical applications.
Stable Diffusion Inpainting is a latent text-to-image diffusion model capable of generating photo-realistic images given any text input, with the extra capability of inpainting the pictures by using a mask.
Claude 3.7 Sonnet — available via AWS Bedrock (us-east-1).
Claude Sonnet 4.5 — available via AWS Bedrock (us-east-1).
Nova Canvas — available via AWS Bedrock (us-east-1).
Nova Reel — available via AWS Bedrock (us-east-1).
Gemma 3 4B IT — available via AWS Bedrock (us-east-1).
Claude 3 Haiku — available via AWS Bedrock (us-east-1).
Titan Multimodal Embeddings G1 — available via AWS Bedrock (us-east-1).
Stable Image Style Transfer — available via AWS Bedrock (us-east-1).
Nova Premier — available via AWS Bedrock (us-east-1).
Claude 3 Sonnet — available via AWS Bedrock (us-east-1).
Claude Opus 4 — available via AWS Bedrock (us-east-1).
Ministral 3 8B — available via AWS Bedrock (us-east-1).
Gemma 3 27B PT — available via AWS Bedrock (us-east-1).
Qwen3 VL 235B A22B — available via AWS Bedrock (us-east-1).