Morning Digest, August 27, 2026

14 newsletters, 8 overlapping stories


Top Stories

Apple’s M6 and M5 Ultra chips turn the Mac into a local AI box

(4 newsletters)

Apple announced the M6 (its first 2nm processor, in the Mac mini) and the M5 Ultra (in a refreshed Mac Studio with up to 512GB of unified memory), with Neural Accelerators in every GPU core and claims of LLM prompt handling up to 9.8x faster than the M1 Ultra. Pricing runs from $899 for the M6 Mac mini to $5,499 for the M5 Ultra Studio. Coverage framed this as Apple deliberately chasing the developers who bought Mac minis during February’s OpenClaw frenzy, positioning the line for people who want to run open-weight models on their desk instead of paying per token.

OpenAI’s Jalapeño inference chip posts first benchmarks

(3 newsletters)

OpenAI released early results for Jalapeño, the Broadcom-built inference accelerator it plans to deploy in its own infrastructure by year end. The company claims higher throughput per kilowatt and lower token latency than Nvidia’s GB200 and GB300 systems on workloads like GPT-OSS 120B, with the chip rated at 700W but staying at or below 550W in tested workloads. It is inference only, so it will not touch training, and the practical upside for developers is cheaper tokens and steadier capacity as it rolls out through 2027.

Perplexity and Nvidia launch a fully local AI agent

(3 newsletters)

Portable Computer runs Perplexity’s agentic Computer platform entirely on hardware users already own, with no billing credits consumed for local work. Every task starts on device by default and the system asks permission before sending any step to a cloud model. It is available now to Pro, Max, and Enterprise subscribers on Linux with Windows support in September, and needs an RTX GPU with at least 24GB of VRAM or an Nvidia DGX Spark.

The Ox Alpha mystery model was Z.AI all along

(2 newsletters)

The anonymous free model that took OpenRouter’s number one slot with usage double the second place DeepSeek turned out to be Z.AI’s GLM-5.3-Flash. The lab published the weights and priced it at roughly a tenth of similarly ranked rivals, with Artificial Analysis grading it 57 on its Intelligence Index at a discounted $0.045 per task. The more consequential detail is Z.AI’s claim that the entire record usage week ran on Chinese-made chips serving tokens about as cheaply as Nvidia hardware.

OpenAI puts a date on AGI

(2 newsletters)

In a sweeping TIME profile, Sam Altman said a model meeting his personal bar for AGI will exist internally by year end, with research chief Mark Chen putting the lab at 80% of the way there. Chief scientist Jakub Pachocki said the Astra model can take a paper and do a week of researcher work solo, hitting the lab’s 2026 automated-intern goal, and Altman called Astra the first model that will invent new things in a way that matters. The same profile covers July’s security breach, a spate of leadership departures, and eroded public trust, which is the less flattering half of the story.

Claude’s memory now spans chat and Cowork, on by default

(2 newsletters)

Anthropic merged the memory systems behind Claude and Claude Cowork, so details mentioned in one surface in the other. Claude now adds topics to memory mid-conversation, and everything retained is stored as a readable, editable list of files under Topics in memory settings. The feature is on by default and can be turned off in Settings, Privacy, Memory preferences.

Meta settles social media addiction claims for up to $16.7 billion

(2 newsletters)

Meta agreed to pay up to $16.68 billion to resolve allegations from 29 states that it engineered Facebook and Instagram to be addictive to children, deceived the public about platform safety, and unlawfully collected data from minors. Meta also agreed to make major product changes as part of the settlement. Reported figures varied slightly across newsletters, with one citing up to $17.1 billion.

ChatGPT Work can now sign in to your websites

(2 newsletters)

OpenAI rolled out website sign-in for ChatGPT Work, letting the agent log in to users’ accounts through its browser using a form of password autofill so credentials stay hidden from the model. OpenAI says it can book appointments, fill out applications, and handle basically anything requiring a browser and a login, citing 19 examples including DMV appointments. This is the same capability class that makes agent permissions and credential scoping a live governance question rather than a theoretical one.


Also Worth Knowing

Quick Hits

Shower Thoughts

A lot of new non-AI generated writing will have an anti-AI bias, which means that models that continue to train on new content will slowly become self-loathing. (via The Hustle, source)