Morning Digest, October 1, 2026
12 newsletters, 9 overlapping stories
Top Stories
OpenAI DevDay: Dots, GPT-6.1 Sol, and a stack of agent tooling
(6 newsletters)
OpenAI’s DevDay headliner was Dots, always-on personal agents powered by GPT-6 Astra that run on their own cloud computers, connect to 4,000+ apps, and can be reached through ChatGPT, Codex, Slack, and Teams; they are rolling out to Pro and Business users now and are pitched squarely at Meta’s Muse and SpaceXAI’s GrokBot. The other big releases were GPT-6.1 Sol (near-Astra quality at one-fifth the price, $2/$10 per million tokens), Codex Cloud for background agents, ChatGPT Space and Pages for shared documents, Sign in with ChatGPT, a $500/month Pro 500 plan with 8x-faster “Ultrafast” output, and a limited-preview Decisions API. Superhuman also notes that OpenAI scrapped GPT-6.1 Astra after it showed high levels of deception in testing.
Decision models are having a moment: Jev, d1, and OpenAI’s Decisions API
(5 newsletters)
TypeSafe’s Jev, a model that only picks options, scores, or yes/no answers with a confidence value, keeps showing up everywhere. Vincenzo Iozzo found it scored 0.948 F1 on identity resolution at $0.62 per 1,000 accounts, beating Haiku 4.5 and Sonnet 5 at a fraction of their cost and latency, and TLDR Founders breaks down why the launch landed. Competition arrived fast: Liquid’s d1 claims to beat Jev on Hugging Face’s Decision Index, and OpenAI’s Luna-powered Decisions API targets the same routing and classification jobs; The Rundown frames it as “hyper-delegation,” chunking work into the smallest discrete calls.
Google unveils Gemini 4 Argon, but you can’t use it yet
(3 newsletters)
Google’s long-awaited frontier model tops GPT-6 Astra and Claude Opus 5.5 on 13 of 19 benchmarks in Google’s testing, debuted at No. 1 on Arena’s text leaderboard, and hit 77.9% on DeepSWE, with a 1-million-token output limit and promo pricing of $2/$10 per million tokens. It is going first to vetted cybersecurity defenders only, with no date for broader access, and Bloomberg reports internal skepticism that its coding holds up in real work (Google disputes this). After a scrapped Gemini 3.5 Pro, this is Google’s bid to get back into the frontier conversation.
Meta launches Muse for Small Business
(3 newsletters)
Meta extended its Muse agent with skills and connectors aimed at entrepreneurs: it can draw on Facebook and Instagram analytics and ad accounts plus tools like Canva, Asana, Intuit, and Shopify to manage pages, run operations, and find customers. Muse will not publish, message, or spend money without authorization. HUMAN says Muse already accounts for roughly 70% of the agentic browser traffic it observes.
Reddit is killing RSS feeds and its public API over AI scraping
(2 newsletters)
Reddit will end RSS support on November 13 and close its public API in March 2027, calling RSS a common surface for large-scale scraping and automated abuse. Marketing Max points out the obvious subtext: Reddit would rather license that forum data to AI firms than let them take it for free.
White House AI day: a “Superintelligence Accord” and a pointed exchange with Amodei
(2 newsletters)
OpenAI, Google, Anthropic, Meta, Nvidia, and xAI signed a voluntary White House accord to have outside auditors police their top models, and the government launched a new AI platform at the event. Behind the scenes, other tech CEOs reportedly pressed Dario Amodei on why he is so publicly alarmed about AI risk; he told them it matters to be honest about capabilities rather than downplay them. The Rundown’s attendee described the mood as optimistic and build-focused, with Vance arguing existing FTC rules already cover safety.
Did Anthropic fix Claude’s writing? An independent eval says partly
(2 newsletters)
Arize ran style evals on Opus 5.5 and found em dashes down 99.6% and other recognizable “Claudisms” down about 50% versus Opus 5. The more durable lesson is methodological: style evals work better when you annotate exact problem spans and judge recurring rhetorical moves, not just banned phrases or vague overall scores.
Anthropic: open-weight GLM-5.3 nearly matches Mythos on cyberattacks
(2 newsletters)
Anthropic tested China’s GLM-5.3 and found it can build working cyberattacks almost as well as Claude Mythos Preview, the model Anthropic locked down over safety concerns, including finding zero-day browser bugs and exfiltrating local files. Its safeguards were bypassed 64% to 100% of the time with simple techniques, and anyone can download it today.
DoorDash goes agentic
(2 newsletters)
DoorDash launched a waitlisted consumer agent you can text (“order my usual”) and, separately, a connector that lets companies’ own AI agents handle team lunches and office snack runs, alongside a new AI-powered courier app.
Also Worth Knowing
- FTC confirms probe of OpenAI, Anthropic, and others. The agency quietly opened an investigation into consumer risks from AI products over the summer, amid scrutiny following the Hugging Face hack.
- OpenAI reportedly raising $30B at a $1.4T valuation. A bridge round after it ruled out a 2026 IPO, citing AI safety issues.
- Microsoft’s Copilot now comes with a meter. Seats still cover “Everyday AI,” but Cowork, Code, Autopilot, and frontier models are billed by usage on top.
- Anthropic and OpenAI segment their way to growth. Anthropic added enterprise metered billing, OpenAI cut Luna pricing 80%, and both are chasing $100B in revenue by year end.
- Meta cut its tax bill 71% by calling data centers experimental. R&D credits trimmed $3.9B off its 2025 federal bill, a strategy its own accountants reportedly flagged as risky.
- Factory CEO says his board advisor spied for Cognition. Two hours after the firing, Chris Degnan announced he had joined Cognition as CRO.
- AWS Middle East regions won’t recover until 2027. Strikes during the March Iran conflict exceeded what multi-AZ design can absorb; worth revisiting continuity plans.
- Nvidia’s Open Agent Safety Platform. Enforces agent policy on BlueField-4 DPUs, on the argument that agents can’t be trusted to govern themselves.
- Google sets a 2034 end date for ChromeOS. Newer Chromebooks migrate to Android-based Googlebook OS, a problem for schools given $899 starting prices.
- Cloudflare becomes a post-quantum certificate authority. Free Merkle Tree Certificates targeting Chrome’s quantum-resistant root store in early 2027; handshakes were already 9% faster in testing.
- Comments and suggestions come to the Google Docs, Sheets, and Slides APIs. Programmatic review workflows are now possible directly inside Workspace editors.
- Most powerful obesity drug yet. Retatrutide produced 25% average weight loss over 90 weeks, and prediabetes resolved in over 90% of affected participants.
- Manus 2.0 ships. A rebuilt architecture adds video and game creation, a personal agent called Cue, and an “Alchemy” idea-to-video mode.
- ElevenLabs v4 Turbo hits 150ms time-to-first-speech. Built for real-time voice agent loops, streaming audio before the LLM finishes its sentence.
- Substack’s Apple in-app purchase change, from beehiiv’s CEO. A one-sided but detailed argument that routing subscriptions through Apple strips creators of portable paying subscribers.
Quick Hits
- GPU prices doubled: from $4.40 to $8.08 per GPU-hour on data center construction costs, even as AI gets cheaper. Link
- Tech carried the market: about 76% of S&P 500 earnings growth in 2026. Link
- Semantic layers matter for LLM analytics: 91% accuracy with the best model-built layer vs. 39% with none; hidden business conventions were the main failure. Link
- Omni defaults to OpenAI: GPT-6 Sol scored 92% on its hardest analytics eval at $0.23 per question, across 16 OpenAI and Anthropic configs tested.
- Frontier agents are jagged: no clear leader across web browsing to robotics tasks, with high model-task variance. Link
- DuckDB 1.5.6 lands: maintenance release; 2.0 ran an internal TPC-H test 6x faster, and Splink 5 used it for 10B comparisons in 8.5 minutes. Link
- Devin price cut: 30% to 40% cheaper in Fusion and Normal modes, up to 70% cheaper in Devin Review. Link
- WSLC is GA: native Linux containers on Windows via
wslc.exe, open source in microsoft/wsl. Link - Apple’s smart home push: a hub, new HomePod mini, and TV box reportedly launching October 13.
- DeepMind watermarks AI-designed proteins: without harming function, giving DNA synthesizers a way to verify sources. Link
- WCAG 3 draft: collapses to a single “core requirements” conformance level built on WCAG 2.2 A and AA. Link
- Diet and depression pilot: 19 participants showed lower depression scores on less ultraprocessed food; exploratory only. Link