| |
Mercury 2.5, a new diffusion-based language model, delivers a 40% intelligence improvement over Mercury 2 while maintaining low latency (1,107 tokens per second) and cost ($0.20 per million input tokens). The model is designed for production workloads in search, voice, and coding applications, with an 80% launch discount and capabilities including tunable reasoning and parallel tool calls.
Read Full Article →
← More Tech news