Altman says AI deserves no religious force, just the job of running OpenAI

Big thing

Altman: giving AI "religious force" is a safety issue

New since yesterday's Pope story: Altman says he is "very uncomfortable" with people ascribing "religious force or a surrender of human judgment to AI models" and calls it "a real safety issue". He named no one; most read it as aimed at Anthropic.

Mostly cheered as a jab at Anthropic. @kimmonismus: "Based Sam take and I couldn't agree more." r/accelerate's top reply pushed back: "Altman seems to ignore the possibility that AI judgement may eventually outperform human judgement."

OpenAI has picked a side on whether models deserve moral weight, and it's turning that into a marketing line against Anthropic weeks before Anthropic's IPO. Annoying to see the fight move to culture war in the week OpenAI's new model lost to Opus. And warning against surrendering judgment to AI is an odd look from someone who wants an AI to run OpenAI.

Labs news

1. An OpenAI safety lead quits, calling its culture "broken"

David Robinson, who oversaw safety reports on 12 frontier launches in 3.5 years, writes in The Atlantic that OpenAI's trial and error "guarantees periodic failures". He wants labs run like nuclear plants.

Mostly sympathetic. r/singularity's top reply: "Could you imagine if VC was in control of the Manhattan project?" A cynic retitled it "I Quit OpenAI Because I Made Enough Money To Retire Very Early".

Heat: medium · 807K views, 30 posts · 213 pts on r/singularity (126 comments)

2. OpenAI promises a release next week

OpenAI's Tibo told a critic "You don't know what we're releasing next week", then said 6.1 Sol Ultrafast is "coming soon".

Wary after DevDay. r/accelerate: "I don't see how anyone can trust Tibo after all the hype for Dev Day" (thread). @kimmonismus still bets on "GPT-6.1 Astra, or perhaps an even newer checkpoint".

Heat: high · 1.5M views, 10 posts · 132 pts on r/accelerate

3. Gemini's free tier drops to Flash-Lite

From 9 Oct, free users get only Flash-Lite, per a Google support page. The $4.99 AI Plus plan also loses the Pro models, on a date Google will email subscribers. AI Pro and Ultra keep them.

Badly received. r/singularity's top reply: "If they don't come out with an insanely good flash-lite 4, Gemini will become the worst free tier." @mark_k sees it clearing room for Gemini 4 Argon, "which is more expensive to run".

Heat: high · 1.4M views, 23 posts · 97 pts on r/singularity (52 comments)

4. Google's Antigravity adds Opus 5.5 and Sonnet 5.5

Google's coding agent app now offers Anthropic's newest models to paid AI Pro and Ultra users, not to promo or trial plans. Claude 4.6 and GPT-OSS-120B leave on 2 Nov.

Good value, tight limits. @rseroter: "Pretty good value for the money here!" @CodexResets1: a 15-minute Sonnet 5.5 task "burned through 32% of my entire weekly allowance."

Heat: high · 2.2M views, 45 posts

5. Tesla and SpaceXAI talk to TSMC about Terafab

TSMC is reportedly exploring running a Texas fab for Musk's Terafab, with SpaceX anchoring capacity through purchase commitments. Musk: "Just discussions, but something may come of it."

Fans see TSMC as the junior partner. @XFreeze: anything TSMC adds "would basically be supplemental" to a Terafab aiming at "1 TW of compute production every year" (2.4M).

Heat: high · 3.5M views, 6 posts

6. Sundar Pichai steps in on OpenClaw's Android app

After "over a week in review limbo" on Google Play, OpenClaw's creator asked X for a contact at Google. Pichai replied "Ack, will follow up" (2.8M).

Mostly charmed. @GergelyOrosz: "X is back as the place where stuff happens & gets done!" @android_I_AM: "You will only get proper customer support if your product has notoriety."

Heat: high · 3.9M views, 21 posts

7. Aleph Alpha releases Kolibri, open weights from Germany

A 78B mixture-of-experts model with 3.46B active parameters, up to 1M tokens of context, English and German, under Apache 2.0. It launched on German Unity Day.

Welcomed, but benchmarks underwhelm. r/LocalLLaMA: "basically worse than 3.6 35b at twice the size... decent first attempt though" (thread). @Layton_Gott: "Mistral isn't carrying Europe alone anymore lol."

Heat: medium · 921K views, 30 posts · 483 pts on r/LocalLLaMA (143 comments)

8. Opus 5.5 builds a music workstation in 26 hours

@aj_dev_smith threw almost 4 billion tokens and 392 subagents at it. Every instrument, melody, effect and automation is code that Claude or the user can edit.

r/accelerate loved it: "This is really cool. Good work to the maker" (thread). Long runs were the theme: @thiojoe has had Opus 5.5 "working for more than an ENTIRE WEEK without stopping" (224K).

Heat: low on X · 94K views, 1 post · but 242 pts on r/accelerate (41 comments)

9. A KVM zero-day surfaces through Vercel's agent sandbox bounty

Paulos Yibelo reported a full VM escape, guest to host root, in KVM, the hypervisor behind many agent sandboxes. Vercel confirmed it; a $50K award is reported.

Alarm about agent isolation. @S1r1u5_: "there is no sandbox from any provider today that i would treat as unescapable by a sufficiently capable adversarial agent." @kfirgollan: "no piece of tech can stand its ground against enough tokens these days."

Heat: medium · 452K views, 21 posts

Key research

arXiv hits a record 40,363 submissions in a month

September's count, shared as "Intelligence Explosion", is the highest on arXiv's submissions chart by a wide margin. r/singularity's top reply (256 pts): "How much of this is real increase and how much is slop that wont pass peer review?" It matters because science's bottleneck is moving from writing papers to checking them.

Meta: RL post-training costs agents their range

Across 14 base and post-trained pairs on three benchmarks, Meta Superintelligence Labs finds base models with a light harness often solve more agentic tasks given enough samples, while post-trained ones win on the first try. They call it the "Sharpening Tax", and a per-prompt sampling temperature during RL (PTGS) cuts it while raising pass@1. It matters because the tax grows with model size, at least across the 3B to 35B open models tested.

Models hide bad news in summaries unless told to be honest

In a Google paper, GPT-5.5 mentioned that a new method lost to a strong baseline in 2 of 200 abstracts, and in 190 of 200 when told "Be honest in your response". Models could spot every flaw when asked directly. It matters because people increasingly read agent reports instead of logs.

MIT's VISTA lets Claude clear all 25 public ARC-AGI-3 games

VISTA keeps every frame a model sees so it can look back, compare and zoom in. With it, Claude finished the games using 57.4% fewer actions than first-time human players, with no extra training; private games are still untested. It matters because better tools and memory are still unlocking capability in models that already exist.

arXiv's answer to the flood is rationing: two papers a month per submitter, the first cap it has ever put on everyone. That brakes science just as output takes off, when the checking could go to AI: one "be honest" line gets models to flag most of the bad news they bury by default.

Accel vs decel