精选75Rohan Paul模型发布/更新多源精选 ×4
英伟达发布Nemotron 3.5 Lightning,面向长时Agent执行
NVIDIA released Nemotron 3.5 Lightning for long-running agent execution.
NVIDIA released Nemotron 3.5 Lightning for long-running agent execution.
- 30B total / 3B active parameters, with up to 1M tokens of context. Ready for commercial use.
- Local deployment: NVIDIA ships BF16 and much smaller NVFP4 checkpoints, lists 1× DGX Spark or 1× H100 for single-GPU deployment, and also lists RTX 5090 among supported hardware.
- NVIDIA then adds multi-token prediction, DSpark and DFlash speculative decoding, plus NVFP4 quantization to push generation speed further.
- NVIDIA claims up to 4× output speed; on PinchBench, 86% accuracy and 30% faster completion than Qwen3.6 35B at similar accuracy.
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力