精选75Chubby♨️行业动态多源精选 ×2
英伟达 Groq 3 LPX 全面投产,Gemma 4 31B 输出达每秒
Inference on steroid: NVIDIA’s Groq 3 LPX is now in full production, adding a de…
Inference on steroid: NVIDIA’s Groq 3 LPX is now in full production, adding a dedicated token-generation accelerator to the Vera Rubin platform.
NVIDIA says it reached 3,400 output tokens per second running Gemma 4 31B with a 100,000-token context in Artificial Analysis benchmarking, the fastest recorded result for that model.
The company also claims 4x faster responsiveness than the nearest alternative platform for agents and latency-sensitive workloads.
Nebius will be the first AI cloud to deploy Groq 3 LPX through its Token Factory, followed by Groq itself.
Intelligence too fast to meter.
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力