Loading...
Browse 4,143 models across providers, modalities, and use cases.
🌐 All Models
4,143 models · Page 109 of 116
Temporary publish target for bgogo deployment validation.
Take a flat graphic, remove text, and get structured text layers back for editing and recomposing
A 4B-parameter safety and content moderation model that classifies user prompts and assistant responses as Safe, Unsafe, or Controversial with fine-grained category labels and refusal detection. Supports 119 languages.
This is a version of Qwen 3.5 27B optimised by Pruna AI.
Ultralytics YOLOv8s worldv2 Real-Time Open-Vocabulary Object Detection model with 12.7M parameters. Achieves 37.7 mAP50-95 on COCO dataset. Optimized for real-time inference
This is Platmoji 2, trained more to mimic emojis in an extremely similar way. (Realism in emojis go to Platmoji 1)
ML model that detects acne, dark circles, wrinkles and oily skin in one go.
Beta/RFC version of https://replicate.com/nelsonjchen/op-replay-clipper
Universal Feed-Forward Metric 3D Reconstruction
Ultralytics YOLOE-L Real-Time Seeing Anything model with 26.2M parameters. Achieves 52.6 mAP50-95 on COCO dataset. Optimized for real-time inference with 6.2 ms speed on T4 GPU..
The fastest diffusion language model with up to ~1000+ tps
Arise V2
A visual style focused on modern enterprise architecture—reflective blue glass skyscrapers captured from upward perspectives, emphasizing symmetry, scale, and a clean, premium corporate aesthetic.
This chatbot is designed to answer medical-related questions and has been fine-tuned on a large dataset. It is built on the Qwen 2.5 3B Instruct base model.
This is a version of the MoE Gemma 4 26B optimised by Pruna AI.
Open-source Phi-3-mini-4k-instruct model from microsoft — available for download and self-hosting on Hugging Face.
Inspired by 90s streetwear, graffiti culture, and macabre aesthetics
Ultralytics YOLO11n object detection model with 2.6M parameters. Achieves 39.5 mAP50-95 on COCO dataset. Optimized for real-time inference with 1.55 ms speed on T4 GPU..
This is an optimised version of the hidream-l1 model using the pruna ai optimisation toolkit!
Object Recognition model
A unified Text-to-Speech demo featuring three powerful modes: Voice, Clone and Design
Google's fast, expressive text-to-speech model with 30 voices and 70+ language support
Open-source OTel-LLM-1B-IT model from farbodtavakkoli — available for download and self-hosting on Hugging Face.
Open-source OTel-LLM-0.6B-IT model from farbodtavakkoli — available for download and self-hosting on Hugging Face.
minimax-m2 — available to run locally via Ollama on CPU and GPU hardware.
Open-source tiny-gpt2 model from sshleifer — available for download and self-hosting on Hugging Face.
Open-source Llama-3.2-1B-Instruct-FP8 model from redhatai — available for download and self-hosting on Hugging Face.
Open-source DeepSeek-V2-Lite-Chat model from deepseek-ai — available for download and self-hosting on Hugging Face.
Open-source Llama-3.2-1B-Instruct-Q8_0-GGUF model from hugging-quants — available for download and self-hosting on Hugging Face.
Open-source Qwen3-Coder-Next-FP8 model from qwen — available for download and self-hosting on Hugging Face.