Loading...
Browse 4,043 models across providers, modalities, and use cases.
💬 Text Generation
4,043 models · Page 106 of 113
The smartest diffusion language model up to ~800+ tps
This is a version of the MoE Qwen 3.5 35B optimised by Pruna AI.
Ultra-fast, cost-efficient text-to-speech with ~120ms latency and 15-language support
Temporary publish target for bgogo deployment validation.
Take a flat graphic, remove text, and get structured text layers back for editing and recomposing
A 4B-parameter safety and content moderation model that classifies user prompts and assistant responses as Safe, Unsafe, or Controversial with fine-grained category labels and refusal detection. Supports 119 languages.
This is a version of Qwen 3.5 27B optimised by Pruna AI.
Ultralytics YOLOv8s worldv2 Real-Time Open-Vocabulary Object Detection model with 12.7M parameters. Achieves 37.7 mAP50-95 on COCO dataset. Optimized for real-time inference
This is Platmoji 2, trained more to mimic emojis in an extremely similar way. (Realism in emojis go to Platmoji 1)
ML model that detects acne, dark circles, wrinkles and oily skin in one go.
Beta/RFC version of https://replicate.com/nelsonjchen/op-replay-clipper
Universal Feed-Forward Metric 3D Reconstruction
Ultralytics YOLOE-L Real-Time Seeing Anything model with 26.2M parameters. Achieves 52.6 mAP50-95 on COCO dataset. Optimized for real-time inference with 6.2 ms speed on T4 GPU..
The fastest diffusion language model with up to ~1000+ tps
Arise V2
A visual style focused on modern enterprise architecture—reflective blue glass skyscrapers captured from upward perspectives, emphasizing symmetry, scale, and a clean, premium corporate aesthetic.
This chatbot is designed to answer medical-related questions and has been fine-tuned on a large dataset. It is built on the Qwen 2.5 3B Instruct base model.
This is a version of the MoE Gemma 4 26B optimised by Pruna AI.
Open-source Phi-3-mini-4k-instruct model from microsoft — available for download and self-hosting on Hugging Face.
Inspired by 90s streetwear, graffiti culture, and macabre aesthetics
Ultralytics YOLO11n object detection model with 2.6M parameters. Achieves 39.5 mAP50-95 on COCO dataset. Optimized for real-time inference with 1.55 ms speed on T4 GPU..
This is an optimised version of the hidream-l1 model using the pruna ai optimisation toolkit!
Object Recognition model
A unified Text-to-Speech demo featuring three powerful modes: Voice, Clone and Design
Google's fast, expressive text-to-speech model with 30 voices and 70+ language support