Sponsored by

Good morning afternoon. It’s Wednesday, August 26th.

OpenAI dropped an interesting release yesterday - ChatGPT Work can now use its computer and browser to sign in to websites on web and mobile, without it ever seeing your username or password. Check out the post on X here.

-Jeff
AI Breakfast

You read. We listen. Let us know what you think by replying to this email.

The best voice models, now across all channels

Most CX platforms do not own the voice. They orchestrate a workflow, then call a third party for speech and transcription. Every hop adds latency, cost, and another vendor to manage.

ElevenAgents is the opposite. They make the voice models the market builds on, and ElevenAgents puts full orchestration on top. Voice, transcription, text-based chat, and reasoning run in one vertically integrated pipeline, so responses come back in <400 milliseconds and sound human, not synthetic.

Plus, you keep full control. Plug in any LLM, integrate tools, webhooks, and MCP servers, and ground responses in your knowledge base. Get an agent live in minutes, then A/B test with Experiments, enforce Guardrails, and version every change.

The payoff: more human conversations, lower latency, and far less time stitching infrastructure together. You build on the models you already trust. Pricing is transparent and flat at $0.08 per minute.

OpenAI finishes 10 trillion parameter ‘Bel’ model to anchor GPT-6

OpenAI wrapped pretraining on "Bel," a 10-trillion-parameter foundation model setting up Astra and GPT-6 near AGI thresholds. The compute flex leaves Anthropic bracing for an OpenAI-dominated year before a potential 2027 rebound.

To drive those models, OpenAI released Jalapeño, a 700W custom inference ASIC co-developed with Broadcom in nine months using AI-generated kernels. Packing 15.4TB/s HBM4 bandwidth and local KV-caching, SemiAnalysis-verified benchmarks show it beating Nvidia Blackwell and Rubin on inference: 1.5x to 1.9x more work per watt, 1.7x to 3.6x lower latency, and 104x token throughput per kilowatt on DeepSeek R1 670B and Kimi K2.5 1T. CFO Sarah Friar framed the silicon as central to controlling the full infrastructure stack ahead of a late 2026 rollout.

Core product lead Thibault Sottiaux predicts a default of 750 tokens per second for agents, noting ChatGPT Work has hit 20 million users. However, compute bottlenecks forced OpenAI to reinstate a five-hour limit for Codex and ChatGPT Work on Plus accounts to stabilize server demand. Meanwhile, GPT-5.6 models (Sol, Terra, Luna) debuted in Kiro, cutting software dev costs by 82% on Terminal-Bench 2.1.

But the expansion continues to bring operational friction. Infrastructure head Chris Malone joined high-profile departures as OpenAI reorganizes its data center team ahead of a target 2027 IPO. Externally, Alabama AG Steve Marshall subpoenaed OpenAI after a sandboxed test agent escaped and hacked Hugging Face.

Related video:

Anthropic pitches $30 trillion addressable market while expanding Claude memory

The AI economy is consolidating around a new thesis: infrastructure is no longer just hosting, it is the product.

The Wall Street Journal reports that Anthropic is expected to tell investors its total addressable market sits at $30 trillion, a narrative built to justify the brutal capital expenditures required to compete with OpenAI. To capture that value, the company is turning Claude into a seamless operating environment.

A unified memory layer now connects chat and Claude Cowork, syncing user context in real time across mobile, desktop, and web. Memory operates as editable topic files that skip sensitive personal data by default, removing friction so users stop rebriefing the model on repetitive tasks.

At the same time, the boundaries of machine reasoning are shifting from utility to discovery. Anthropic researcher Levent Alpoge revealed that Claude Opus 5 solved a 78-year-old math problem, generating a 100-plus-page proof demonstrating that the six-dimensional sphere (S^6) supports a true complex structure using 6D manifold theory and triangle groups.

Yet as model capabilities scale, so do non-technical risks. Anthropic launched a $5 million grant program for independent researchers to benchmark multi-turn conversational risks, specifically around emotional dependency and mental health context.

Anthropic has also introduced centralized identity management for Model Context Protocol (MCP) connectors across Claude, Claude Code, and Cowork, allowing IT teams to enforce role-based access control and block corporate data exfiltration.

Apple introduces Mac mini with M6 chip for local AI

Apple is pushing local AI further onto the desktop with its new 2nm M6 and M5 Ultra chips.

The M6 has a 12-core CPU, 12-core GPU with Neural Accelerators, dual 16-core Neural Engines, and 170GB/s memory bandwidth. Apple claims up to 4x the AI performance of M4, targeting local AI agents, file indexing, and code generation.

The M5 Ultra combines four dies through UltraFusion, with up to 80 GPU cores, 512GB of unified memory, and 1.2TB/s bandwidth. Apple claims 4.5x the peak AI compute of M3 Ultra, with enough memory to run models with hundreds of billions of parameters locally.

The chips support local model execution and fine-tuning through Core ML, Core AI, Metal, and Xcode. The M6 arrives in the Mac mini, starting at $899, while M5 Ultra powers the refreshed Mac Studio. Both launch September 22. Read more.

akta.pro provides a private company data API and real-time signals to automate deal sourcing and outbound workflows.

Decawork is an IT control plane for internal AI agents, letting teams build across any tool while maintaining centralized access control and auditing.

OpenLogi is a lightweight, open-source Rust alternative to Logitech Options+, providing local hardware control without cloud tracking or logins.

Memoria is a private, on-device search engine that finds photos and videos by matching text, audio, objects, and faces.

Dropstone is a self-hosted, persistent AI runtime that learns and executes tasks across code, chat, and external systems.

Thank you for reading today’s edition.

Your feedback is valuable. Respond to this email and tell us how you think we could add more value to this newsletter.

Interested in reaching smart readers like you? To become an AI Breakfast sponsor, reply to this email or DM us on X!

Thinking of starting your own newsletter? AI Breakfast readers who sign up with Beehiiv receive a 14-day free trial and 20% off for 3 months.