精选70Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)模型发布/更新
模型增益来自不同失败,而非堆叠模型
«Gains come from models failing on different questions, not from adding more mod…
«Gains come from models failing on different questions, not from adding more models.»
Haven't read it yet but sounds right. Mixture of Models, where all models are general-purpose competing LLMs, is a cope. Just train experts, do MOPD, and then do single-model test time scaling.
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力