跳到正文
原文
Inception Labs:Blog·· 2026-07-14AI 評分69

Inception Labs 發佈 Mercury 2 推理模型,解碼速度達 1,000+ tokens/second

Mercury 2: the first reasoning model fast enough to pick up the phone

AI 導讀

Mercury 2 是 Inception Labs 推出的擴散式推理語言模型,官方稱其在標準 NVIDIA GPU 上解碼速度達 1,000+ tokens/second,300-token 推理鏈完成時間不到 300ms。

來源:Inception Labs:Blog · inceptionlabs.ai