Bria/expand
deepinfra
No description available.
- Context window
- — tokens
- Input cost
- —
- Output cost
- —
- Latency (p50)
- 79 ms
Loading...
Run side-by-side checks for pricing, context window, and latency.
Price (lower is better), Speed/Latency (lower is better), Performance (higher is better)
💰Bria/expand is 100% cheaper than Meituan: LongCat Flash Chat
⚡Meituan: LongCat Flash Chat is 79ms faster than Bria/expand
formlessai
Open-source Qwen2.5-7B-Instruct-Code-Unsloth model from formlessai — available for download and self-hosting on Hugging Face.
bmax16634
Open-source solollm-1.0-152m-base model from bmax16634 — available for download and self-hosting on Hugging Face.
joyboy5656n
Open-source coding-assistant-llm model from joyboy5656n — available for download and self-hosting on Hugging Face.
Let our experts help you evaluate models based on your actual use case and budget. We'll provide a free analysis and actionable recommendations.
deepinfra
No description available.
meituan
LongCat-Flash-Chat is a large-scale Mixture-of-Experts (MoE) model with 560B total parameters, of which 18.6B–31.3B (≈27B on average) are dynamically activated per input. It introduces a shortcut-connected MoE design to reduce...