Loading...
Loading...
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
Best for
Price history
Local run command
# Refer to provider docs:
Presumptive Specs:
Typical GPU
12–24 GB VRAM
System RAM
16+ GB
Disk
20+ GB free
Runtime
Docker / local runtime
Community reliability estimate · not official
About this score: Community-estimated based on user reports and publicly available benchmark data (e.g. TruthfulQA). This is not an official score from the model provider. Scores may be inaccurate — always verify with the official leaderboard before making production decisions.
Input & output cost per 1M tokens over time
Price history is a Pro feature
Track pricing trends and catch price drops early.
Upgrade to Pro — $19/mo →Provider
Learn how to use Inception: Mercury 2
💡 Tip: Start with the sample prompts above to see how Inception: Mercury 2 works best.
Proven prompts shared by the community for this model