跳到主内容
精选75Chubby♨️行业动态多源精选 ×2

英伟达 Groq 3 LPX 全面投产,Gemma 4 31B 输出达每秒

Inference on steroid: NVIDIA’s Groq 3 LPX is now in full production, adding a de…

原文

Inference on steroid: NVIDIA’s Groq 3 LPX is now in full production, adding a dedicated token-generation accelerator to the Vera Rubin platform.

NVIDIA says it reached 3,400 output tokens per second running Gemma 4 31B with a 100,000-token context in Artificial Analysis benchmarking, the fastest recorded result for that model.

The company also claims 4x faster responsiveness than the nearest alternative platform for agents and latency-sensitive workloads.

Nebius will be the first AI cloud to deploy Groq 3 LPX through its Token Factory, followed by Groq itself.

Intelligence too fast to meter.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源

相似阅读

另一事件,读法相近