Mira Murati 实验室发布开源模型 Inkling,LMCache 加速推理
Today’s edition of my newsletter just went out.
Today’s edition of my newsletter just went out.
🔗 https://www.rohan-paul.com/p/mira-muratis-thinking-machines-lab
🗞️ Mira Murati’s Thinking Machines Lab drops massive open-weight AI model, Inkling, Apache 2.0 license, no restrictions.
🗞️ Upto 10.7X speedup on LLM inference with LMCache - Reuses the Most Expensive Part of Long Prompts, 10.6K GitHub Stars
🗞️ Meta is so back in the AI coding race with Muse Spark 1.1, using cut-rate pricing to pressure OpenAI and Anthropic in agentic coding.
🗞️ Apple sued OpenAI, alleging stolen hardware secrets now underpin its $6.5B device push.
🗞️ For Perplexity Computer, Grok 4.5 became the strongest orchestrator, scoring 0.328 (WANDR benchmark score) at $4.76 per trial.
🗞️ OpenAI Releases Prompt Behind GPT-5.6 Sol Ultra Math Proof
🗞️ GPT 5.6 Terra is behind across the entire intelligence-cost curve.
🗞️ Goldman Sachs: “Token use by AI agents is expected to multiply 24 times by 2030”
🗞️ Surprising and such a good news for open source coding model, and also that there are lots of hidden chances to reduce cost while improving quality.
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力