1. OpenAI publishes 722 math papers from an unreleased model
Rumoured on Monday as "about 400 proofs", the drop came on Tuesday: 722 manuscripts in 372 families of results, among them a zero-free strip for the Riemann zeta function and a matrix multiplication exponent of at most 2.25 (the record was about 2.37), many with Lean proofs. By Wednesday a mathematicians' group was urging a boycott that Terence Tao reposted, and by Friday Fields medalist Hugo Duminil-Copin said every problem he used to cite had been solved.
What stuck: First awe, with @AlexKontorovich on the zeta result: "If a human did this, it would be an instant Fields Medal, no questions asked." Then the human cost: NYU's Tristan Buckmaster said the drop "destroyed the careers of early career mathematicians", and r/accelerate's top reply: "They should just do what coders do and become vibe mathers".
Heat: 42.1M views, 293 posts · 6 days in the news · 1,390 pts on r/accelerate
2. Grok Bot now runs on Claude Opus 5.5
A day after Musk's burst of Grok Bot posts, he said on Wednesday that SpaceX will "use the best back end model for any given task, including Claude Opus 5.5", and @poteto, who works on Grok Bot, said "all bots will be powered by opus 5.5". The bot kept shipping all week (Shopify stores, its own email address), and Musk pushed its shopping through Stripe Link and says a fast Grok 4.8 will take the simple requests "when that comes out".
What stuck: @minchoi's "Best model wins" (7.6M) carried the cheer, and @AlexFinn put it as "The Claude Bot everyone has been waiting for has arrived and it's.... Grok Bot". The counter-narrative was a question: "is elon giving up on grok?"
Heat: 403.7M views, 1,499 posts · 5 days in the news · quiet on Reddit
3. OpenAI starts 28 days of Codex upgrades or resets
On Sunday night OpenAI's Tibo promised that each day for 28 days Codex would ship "a clear improvement" for most users or a full reset of usage limits. Six days in, users got about 50% faster Astra and Sol, free auto-review, GPT-6 in ChatGPT's Chat tab, instant steering, composer predictions (Pro only) and three resets: two came on top of features, and on Saturday no feature, just a global reset.
What stuck: @kimmonismus called it "an attempt to stop people from defecting to Claude", and users made clear what they wanted: "You guys could ship Astra 502.1 and we'd still vote reset".
The resets are meant to stop coders leaving for Claude. @cherry_mx_reds left anyway and still found a use for them: Opus 5.5 is the boss and hands the bug fixes to Codex (Opus keeps the design work), so the Codex resets make the Claude plan last longer. So Codex is now Claude's intern (on OpenAI's compute), and I think only a new model gets it the boss job back.
Heat: 31.7M views, 109 posts · 7 days in the news · 304 pts on r/accelerate
4. Nadella says to treat AI models as insider risks
On Saturday, a day after Axios reported AI executives privately gaming out a catastrophic AI event, Microsoft's CEO published "Models as Insider Risks in the Super Intelligence Era": treat closed and open models like insiders who can err or be compromised, keep the controls outside the model, and give an authorized person an "emergency brake" to stop a model mid-task. He calls for no pause and asks nothing of governments.
What stuck: The brake spread as a brake on AI itself (@WatcherGuru: Nadella "calls for "emergency brake" on advanced AI", 1.1M). @DavidSacks turned it on Anthropic: "Satya is right. The way to make SI safe is not to train it with a sense of self, its own moral philosophy, and permission to act as a conscientious objector."
Microsoft asked for this same brake in 2023, but written into the law, with a new government agency licensing the big models. Now the customer holds the brake and nobody needs a license, so nothing slows the labs down (and Microsoft gets to sell the brake).
Heat: 19.1M views, 39 posts · 1 day in the news · quiet on Reddit
Money
- TypeSafe AI raised an $870M Series A at a $7.5B valuation, led by a16z with Sequoia and DCVC, 24 days after leaving stealth at $200M. Jev, a model that returns decisions instead of text.
- DeepSeek is close to raising at least $12B at a valuation of at least about $75B, per Bloomberg, with Tencent and CATL the biggest backers ahead of a 2027 IPO.
- Manus raised over $500M led by Boyu Capital and IDG Capital, its first round since the split with Meta. General-purpose AI agents.
- Lambda is raising up to $4B at a $14.5B pre-money valuation, led by Coatue and Blackstone, ahead of a planned 2027 IPO. AI computing.
- Arena raised a $200M Series B at a $3.1B valuation, co-led by Lightspeed and Khosla. The model leaderboard.
- Reflection AI is in early talks to be acquired or acqui-hired by NVIDIA, or to take more of its money, per the FT. The open-weight lab behind Beam, last valued at $25B.
What to watch next week
- Anthropic's IPO: institutional investors question its executives on Wednesday 14 Oct, which could lead to formal marketing.
- Codex's 28 days continue, with 22 days left of an upgrade or a reset each day.
- Gemini 4 Argon is still hard to get, and Google is reportedly testing a newer checkpoint, Carbon, that matches Opus 5.5 in coding (rumoured).
- Grok 4.8: Musk says a fast version will handle most Grok Bot requests "when that comes out" (teased, no date).
- Qwen's "let's do both!" after "big or small?": followers expect a big and a small Qwen 4 open-weight release (teased, no models or date named).