Loading...
Browse 4,151 models across providers, modalities, and use cases.
🌐 All Models
4,151 models · Page 104 of 116
SAM-Audio is a foundation model for isolating any sound in audio using text
Agentic image model optimized for robust, high-precision generations supporting font control
FireRed-Image-Edit is a general-purpose image editing model that delivers high-fidelity and consistent editing across a wide range of scenarios.
Bria AI's remove background model
Monocular relative depth estimation
Monocular metric depth estimation
NghiTTS API for Vietnamese
Anthropic's most intelligent model with state-of-the-art coding, reasoning, and agentic capabilities
Image-to-video generation with optional audio, multi-shot narrative support, and faster inference
A faster and cheaper Imagen 3 model, for when price or speed are more important than final image quality
SOTA image model from xAI
Fast LoRA trainer for p-image, a super fast text-to-image model developed by Pruna AI. Use LoRAs here: https://replicate.com/prunaai/p-image-lora. Find or contribute LoRAs here: https://huggingface.co/collections/PrunaAI/p-image
Use trained LoRAs from the https://replicate.com/prunaai/p-image-edit-trainer. Find or contribute LoRAs here: https://huggingface.co/collections/PrunaAI/p-image-edit-loras.
Anime Illust model
SOTA Open source model trained on licensed data, transforming intent into structured control for precise, high-quality AI image generation in enterprise and agentic workflows.
Commercial-ready, trained entirely on licensed data, text-to-image model. With only 4B parameters provides exceptional aesthetics and text rendering. Evaluated to be on par to other leading models in the market
A foundation model for isolating any sound in audio using text, visual, or temporal prompts
Modify an existing video through natural-language commands, changing subjects, environments, and visual style while preserving the original motion and timing.
Google's Imagen 4 flagship model
Google's highest quality text-to-image model, capable of generating images with detail, rich lighting and beauty
Automatically remove backgrounds from videos -perfect for creating clean, professional content without a green screen.
Upscale videos up to 8K output resolution. Trained on fully licensed and commercially safe data.
3D models with texture fidelity and geometry precision
Realistic camera simulation and photo post-processing engine.
Animate any character, humans, cartoons, animals, even non-humans, from a single image + driving video
A high-fidelity capability for erasing unwanted objects, people, or visual elements from videos while maintaining aesthetic quality and temporal consistency
I WILL ADD DESCRIPTION SOON