精选70Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)论文研究
Zyphra拟合持续训练LLM可塑性损失缩放定律
Zyphra fits a scaling law for plasticity loss in continuously trained LLMs. What…
Zyphra fits a scaling law for plasticity loss in continuously trained LLMs. What can we do to push the point of rigidity onset towards infinity? I recall Sutton's team could only come up with continual backpropagation (random reinitialization of some units)… suboptimal.
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力