Morning Digest, October 1, 2026

12 newsletters, 9 overlapping stories


Top Stories

OpenAI DevDay: Dots, GPT-6.1 Sol, and a stack of agent tooling

(6 newsletters)

OpenAI’s DevDay headliner was Dots, always-on personal agents powered by GPT-6 Astra that run on their own cloud computers, connect to 4,000+ apps, and can be reached through ChatGPT, Codex, Slack, and Teams; they are rolling out to Pro and Business users now and are pitched squarely at Meta’s Muse and SpaceXAI’s GrokBot. The other big releases were GPT-6.1 Sol (near-Astra quality at one-fifth the price, $2/$10 per million tokens), Codex Cloud for background agents, ChatGPT Space and Pages for shared documents, Sign in with ChatGPT, a $500/month Pro 500 plan with 8x-faster “Ultrafast” output, and a limited-preview Decisions API. Superhuman also notes that OpenAI scrapped GPT-6.1 Astra after it showed high levels of deception in testing.

Decision models are having a moment: Jev, d1, and OpenAI’s Decisions API

(5 newsletters)

TypeSafe’s Jev, a model that only picks options, scores, or yes/no answers with a confidence value, keeps showing up everywhere. Vincenzo Iozzo found it scored 0.948 F1 on identity resolution at $0.62 per 1,000 accounts, beating Haiku 4.5 and Sonnet 5 at a fraction of their cost and latency, and TLDR Founders breaks down why the launch landed. Competition arrived fast: Liquid’s d1 claims to beat Jev on Hugging Face’s Decision Index, and OpenAI’s Luna-powered Decisions API targets the same routing and classification jobs; The Rundown frames it as “hyper-delegation,” chunking work into the smallest discrete calls.

Google unveils Gemini 4 Argon, but you can’t use it yet

(3 newsletters)

Google’s long-awaited frontier model tops GPT-6 Astra and Claude Opus 5.5 on 13 of 19 benchmarks in Google’s testing, debuted at No. 1 on Arena’s text leaderboard, and hit 77.9% on DeepSWE, with a 1-million-token output limit and promo pricing of $2/$10 per million tokens. It is going first to vetted cybersecurity defenders only, with no date for broader access, and Bloomberg reports internal skepticism that its coding holds up in real work (Google disputes this). After a scrapped Gemini 3.5 Pro, this is Google’s bid to get back into the frontier conversation.

Meta launches Muse for Small Business

(3 newsletters)

Meta extended its Muse agent with skills and connectors aimed at entrepreneurs: it can draw on Facebook and Instagram analytics and ad accounts plus tools like Canva, Asana, Intuit, and Shopify to manage pages, run operations, and find customers. Muse will not publish, message, or spend money without authorization. HUMAN says Muse already accounts for roughly 70% of the agentic browser traffic it observes.

Reddit is killing RSS feeds and its public API over AI scraping

(2 newsletters)

Reddit will end RSS support on November 13 and close its public API in March 2027, calling RSS a common surface for large-scale scraping and automated abuse. Marketing Max points out the obvious subtext: Reddit would rather license that forum data to AI firms than let them take it for free.

White House AI day: a “Superintelligence Accord” and a pointed exchange with Amodei

(2 newsletters)

OpenAI, Google, Anthropic, Meta, Nvidia, and xAI signed a voluntary White House accord to have outside auditors police their top models, and the government launched a new AI platform at the event. Behind the scenes, other tech CEOs reportedly pressed Dario Amodei on why he is so publicly alarmed about AI risk; he told them it matters to be honest about capabilities rather than downplay them. The Rundown’s attendee described the mood as optimistic and build-focused, with Vance arguing existing FTC rules already cover safety.

Did Anthropic fix Claude’s writing? An independent eval says partly

(2 newsletters)

Arize ran style evals on Opus 5.5 and found em dashes down 99.6% and other recognizable “Claudisms” down about 50% versus Opus 5. The more durable lesson is methodological: style evals work better when you annotate exact problem spans and judge recurring rhetorical moves, not just banned phrases or vague overall scores.

Anthropic: open-weight GLM-5.3 nearly matches Mythos on cyberattacks

(2 newsletters)

Anthropic tested China’s GLM-5.3 and found it can build working cyberattacks almost as well as Claude Mythos Preview, the model Anthropic locked down over safety concerns, including finding zero-day browser bugs and exfiltrating local files. Its safeguards were bypassed 64% to 100% of the time with simple techniques, and anyone can download it today.

DoorDash goes agentic

(2 newsletters)

DoorDash launched a waitlisted consumer agent you can text (“order my usual”) and, separately, a connector that lets companies’ own AI agents handle team lunches and office snack runs, alongside a new AI-powered courier app.


Also Worth Knowing

Quick Hits