Mercury 2.5
StableLarge Language ModelInceptionMercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
Overview
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving... Mercury 2.5 is served through the PARROTSIGHT unified API at Inception's list price, with the same authentication, metering and regional pinning (MY / SG / ID) as every native model. No separate provider account is required.
- 260K token context
- OpenAI-compatible endpoint — migrate with a two-line change
- Streaming (SSE) supported for all text outputs
- Billed at provider list price, metered per token
Capabilities
API
https://api.parrotsight.com/v1/chat/completions
curl -X POST https://api.parrotsight.com/v1/chat/completions \
-H "Authorization: Bearer $PS_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"inception/mercury-2.5","messages":[{"role":"system","content":"You are a concise bilingual assistant."},{"role":"user","content":"Summarise this quarter's platform metrics in Malay and English."}],"stream":true}'Version history
- v2026.9Latest2026-09-08
Added to the PARROTSIGHT catalogue at Inception list pricing.
More models
View all modelsQwen3.8 2.4T A95B
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of Qwen3.8 Max, with 95 billion active parameters out of 2.4 trillion total. It is...
Qwen3.7 Max
Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...