Inception Labs:Blog·· 2026-07-14AI 評分69
Inception Labs 發佈 Mercury 2 推理模型,解碼速度達 1,000+ tokens/second
Mercury 2: the first reasoning model fast enough to pick up the phone
AI 導讀
Mercury 2 是 Inception Labs 推出的擴散式推理語言模型,官方稱其在標準 NVIDIA GPU 上解碼速度達 1,000+ tokens/second,300-token 推理鏈完成時間不到 300ms。
來源:Inception Labs:Blog · inceptionlabs.ai