MythoMax 13B
gryphe
One of the highest performing and most popular fine-tunes of Llama 2 13B, with rich descriptions and roleplay. #merge
- Context window
- 4,096 tokens
- Input cost
- $0.06 / 1M
- Output cost
- $0.06 / 1M
- Latency (p50)
- 16 ms
Loading...
Run side-by-side checks for pricing, context window, and latency.
Price (lower is better), Speed/Latency (lower is better), Performance (higher is better)
Qwen: Qwen2.5 Coder 7B Instruct is 100% cheaper
💰Qwen: Qwen2.5 Coder 7B Instruct is 50% cheaper than MythoMax 13B
⚡Qwen: Qwen2.5 Coder 7B Instruct is 16ms faster than MythoMax 13B
giannisan
Open-source Laguna-S-2.1-Q2K-pulsar model from giannisan — available for download and self-hosting on Hugging Face.
vontra
Open-source Solar-Open2-250B-MLX-6bit model from vontra — available for download and self-hosting on Hugging Face.
atomicchat
Open-source ornith-35b-MLX-6bit model from atomicchat — available for download and self-hosting on Hugging Face.
Let our experts help you evaluate models based on your actual use case and budget. We'll provide a free analysis and actionable recommendations.
gryphe
One of the highest performing and most popular fine-tunes of Llama 2 13B, with rich descriptions and roleplay. #merge
qwen
Qwen2.5-Coder-7B-Instruct is a 7B parameter instruction-tuned language model optimized for code-related tasks such as code generation, reasoning, and bug fixing. Based on the Qwen2.5 architecture, it incorporates enhancements like RoPE,...