精选70Ollama(GitHub Releases)AI 编程与模型
Ollama v0.32.4:支持Apple GPU的Laguna模型,修复Qwen3 MoE解码
v0.32.4
What's Changed
- Support Laguna on Apple GPUs via the MLX engine
- Quantize draft-model output heads at the requested type when creating speculative-decoding drafts.
- Fixed Qwen3 MoE decoding for differently-quantized experts, plus faster packed gate/up projection (~4–9% on M5 Max).
Full Changelog: v0.32.3...v0.32.4
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力