Back to News
Inception LabsFebruary 27, 2026

Mercury 2

Paid
LLMs

Explore Mercury 2

Visit the official website to learn more and get started

### TL;DR

Mercury 2 is a cutting-edge diffusion-based large language model (dLLM) developed by Inception Labs, designed to deliver ultra-fast and efficient AI-driven text generation. By leveraging diffusion technology, Mercury 2 generates multiple tokens simultaneously, achieving speeds over 1,000 tokens per second on NVIDIA Blackwell GPUs, significantly outperforming traditional autoregressive models.

Key Insights & Metrics

Pricing
$0.25 per million input tokens; $0.75 per million output tokens
Cost structure
Version
2.0
Current release version
Hardware
NVIDIA Blackwell GPUs
Compute requirements
Category
Paid
Licensing model
Region
United States
Primary region

Key Features

  • Diffusion-based architecture for parallel token generation
  • Over 1,000 tokens per second processing speed on NVIDIA Blackwell GPUs
  • Competitive performance with leading speed-optimized models
  • Tunable reasoning capabilities
  • 128K context window support
  • Native tool usage
  • Schema-aligned JSON output

Related Releases

Protenix

Protenix is an open-source, trainable PyTorch implementation of AlphaFold 3, designed for high-accuracy biomolecular structure prediction. It aims to advance accessible and extensible research tools for the computational biology community. ([github.com](https://github.com/bytedance/Protenix?utm_source=openai))

ByteDanceNov 5
Open

Step-Audio-R1

Step-Audio-R1 is an advanced audio language model developed by StepFun AI, designed to enhance audio reasoning capabilities by grounding its reasoning in acoustic features. It introduces Modality-Grounded Reasoning Distillation (MGRD), an iterative training framework that shifts the model's reasoning from textual abstractions to acoustic properties, effectively addressing the 'inverted scaling' problem where performance degrades with longer reasoning. This model has demonstrated superior performance across various audio understanding and reasoning benchmarks, surpassing models like Gemini 2.5 Pro and achieving results comparable to Gemini 3 Pro.

StepFun AINov 29
Open

SINQ

SINQ (Sinkhorn-Normalized Quantization) is a novel, fast, and high-quality quantization method designed to make any Large Language Model (LLM) smaller while preserving accuracy. It offers a plug-and-play, model-agnostic technique that delivers state-of-the-art performance for LLMs without sacrificing accuracy.

HuaweiNov 14
Open

AIBuildAI

AIBuildAI is an AI agent that autonomously constructs AI models. Given a specific task, it initiates an agent loop to analyze the problem, design models, and execute training processes, all without human intervention.

AIBuildAI
Open

Discussion

0
Upvotes
0
Downvotes
0 reviews

Sign in to leave a review

Reviews

No reviews yet. Be the first to review!

🚀 Join the AI dev community — follow us everywhere

© 2026 MARKTECHPOST AI MEDIA INC. All rights reserved.Terms & ConditionsPrivacy Policy
Beta Mode