跳到主内容
@wquguru
精选75Rohan Paul模型发布/更新多源精选 ×4

英伟达发布Nemotron 3.5 Lightning,面向长时Agent执行

NVIDIA released Nemotron 3.5 Lightning for long-running agent execution.

原文
发到 X

NVIDIA released Nemotron 3.5 Lightning for long-running agent execution.

  • 30B total / 3B active parameters, with up to 1M tokens of context. Ready for commercial use.
  • Local deployment: NVIDIA ships BF16 and much smaller NVFP4 checkpoints, lists 1× DGX Spark or 1× H100 for single-GPU deployment, and also lists RTX 5090 among supported hardware.
  • NVIDIA then adds multi-token prediction, DSpark and DFlash speculative decoding, plus NVFP4 quantization to push generation speed further.
  • NVIDIA claims up to 4× output speed; on PinchBench, 86% accuracy and 30% faster completion than Qwen3.6 35B at similar accuracy.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源

相似阅读

另一事件,读法相近