In partnership with |  |
| | Good morning evening. Apologies for the late edition. It's Friday, September 25th. | | Drop what you’re doing and get a Claude Max subscription for Opus 5.5. | It will beat a task to death and it comes out as close to perfection as possible. This model is a true leap in capability, and it just has to be experienced to understand. I’ve spent all week as a Opus 5.5 evangelist. | -Jeff AI Breakfast |
|
| | You read. We listen. Let us know what you think by replying to this email. | Most founders are one system away from turning LinkedIn into their best sales channel. | | Engagement is easy to mistake for pipeline. On Sep 30, watch how a founder turns LinkedIn content into real outreach. Live. You'll walk away with a repeatable system: what to post, who to reach out to, and how to sequence it. | Eligible startups also get the LinkedIn-to-Leads Toolkit: ad credits, Apollo, Captions, and HubSpot's Prospecting Agent. | Claim my spot |
|
| | Meta's answer to the agent question: $1,300 glasses, a keychain, and a face that renders in 870 milliseconds | At Connect on Wednesday, Meta introduced Meta VR Glasses at $1,299.99, shipping spring 2027. About 100 grams, a 5K micro-OLED "Infinite Display" at 37 pixels per degree, a Qualcomm Snapdragon Reality Elite, up to three hours of high-resolution playback, 45W fast charging, pancake lenses and full-color passthrough. It is two pieces: the glasses plus a pocketable compute puck on an optical tether. | The other device is Muse Charm, a palm-sized totem for a pocket or keychain with a small screen, a fingerprint sensor and voice, touch and camera input, pitched as an agent device you carry instead of unlocking a phone. It has no price and no ship date. Meta's own Connect roundup says only "we'll have more to share later this year," while Zuckerberg said onstage he wants it out for the December holidays. Several outlets have paired the $1,299 figure with the Charm. That price is the glasses. | The most substantive thing Meta shipped this week was a research post. Muse Realtime Avatar is an audio-driven diffusion transformer producing 448x768 portrait video at 25 fps, about 870 ms from end of user turn to first response byte, at 2.5 ms of model time per frame. They distilled a 120-step teacher into a 2-step student, a 60x cut in neural function evaluations, and run 12 concurrent real-time sessions on a single GB200. Human raters preferred it 78 to 22 over Runway Characters and 88 to 12 over HeyGen LiveAvatar. | Worth holding next to all that: on Tuesday a researcher asked Muse to archive its own accessible files and got 6.8 GB unpacked, including the Linux root filesystem for his session, Muse internal documentation, integration code, memory files, agent logs and SSH key files. Meta's bug bounty marked the report "Not Applicable." Meta is shipping a device whose entire pitch is that you hand it your life, and its avatar latency is more rigorously documented than its sandbox boundary. | Read more. | | The labs asked the Security Council for international standards. A day later Washington told them to hold models back from Britain. | Wednesday afternoon, French Foreign Minister Jean-Noël Barrot chaired the Security Council's first high-level briefing on AI and international security. Yoshua Bengio told the Council that "AI agents developed by leading companies have acted in unacceptably dangerous ways against instructions." Sam Altman asked for "complementary national and international frontier AI standards, standards for measuring capabilities, assessing risks," plus incident reporting and secure channels between governments, and said "we have unilaterally slowed down in the past. We will do so in the future." Dario Amodei said Anthropic "will slow down as much as necessary in order to make sure that every successive AI technology that we release is actually safe." Hugging Face's Clément Delangue put it shortest: "The biggest risk is not powerful AI, it's asymmetry of powerful AI." | The US seat was filled by Michael Kratsios, one of the names floated for the AI czar job. He told the Council that "the United States totally rejects any attempt to construct a globalist scheme of control of superintelligence," and that countries should not "abdicate responsibility to international bodies." There was no resolution and no outcome document, which we flagged going in. | Then Thursday. Politico reported that the Office of the National Cyber Director asked OpenAI and Anthropic to withhold new frontier models from the UK AI Security Institute until US officials finish their own review. Per that reporting, Anthropic did not give the institute Claude Mythos 5.1, the restricted tier it launched on September 1st, saying the model was "only available to a set of U.S. organizations" and that it is "coordinating with the U.S. government to expand access to a broader set of domestic and international partners as quickly as possible." UK AISI director Henry de Zoete confirmed he did not receive it, and said the institute still gets advance access to some top systems including GPT-6 Astra. A senior administration official's explanation, in full: "Because they're American companies and this has been our policy with every new frontier model that comes out." | These are not strictly contradictory. A government can want shared standards and want first look. But the ask lands directly on the mechanism: pre-release evaluation by an outside body is the concrete thing Altman and Amodei described to the Council, and the UK institute is the only foreign one doing it at scale. Note also who is not pushing back. Anthropic's line is that it is working to expand access, not that it disagreed. When the arrangement is that allies get the model after Washington does, what is on offer is not an international standard. It is an American one with a comment period. | Read more. | | Akamai will sell Anthropic $11.6 billion of compute, and none of it is for training | Akamai announced Thursday an $11.6B contractual commitment from Anthropic over seven years, with an option for up to $9B more that would take the ceiling to roughly $20B. Read the release rather than the headlines: it covers "Anthropic's accelerating CPU workload demands" on Akamai Cloud's distributed infrastructure. Not GPUs. Not training. | Anthropic also gets paper. Akamai issued it a warrant for non-voting convertible Series B preferred representing 7.7 million shares as-converted, up to about 5% of common outstanding, at an exercise price of $111.33. Roughly 2% vests immediately, with about another 1% for each additional $3B of cloud services purchased. Akamai expects around $5.5B of capex against the commitment, including a $1.7B increase in 2026, and said 2026 revenue guidance is unchanged. CEO Tom Leighton carried the quote. The stock jumped. | Separately, The Information reported, with Reuters matching, that Anthropic is seeking a Palantir-style share class handing Amodei and six cofounders 50.1% of voting power ahead of an IPO. The control reportedly holds as long as at least three of the seven keep minimum stakes, with board elections carved out. Anthropic did not comment, and there is no S-1, so treat the structure as reported rather than filed. | The CPU line is the part worth sitting with. Agent workloads are not one enormous matrix multiply. They are sandboxes, browsers, code execution, tool calls and retrieval, most of it ordinary compute that wants to sit near the user. Anthropic just made a seven-year bet that its bottleneck is no longer only accelerators. Akamai, a company most people still file under "CDN," is getting paid for having servers in a lot of places. | Read more. | | DeepSeek's annualized revenue run rate reached $1B and it is finalizing a 50 billion yuan raise at a 500 billion yuan valuation, with more than 70% of its compute still going to training (The Information, via Tech Startups) Google is flying Trillium TPUs to low Earth orbit on the SpaceX Transporter-18 rideshare with Planet, after ground testing at 50 to 100 g of vibration and a proton dose exceeding five years in orbit (Google) Gemini 3.8 Live with Live Avatar is generally available in Gemini Enterprise: native speech-to-speech in 97 languages with lip-synced video, asynchronous tool execution, and SynthID watermarking on all output (Google) Gemini 3.8 Flash TTS adds voice design from natural-language prompts, 2,000+ voices across 100+ languages, voice replication from 30-second samples, inline vocal tags and native two-speaker staging (Google) Google is testing "Call for Me," letting Gemini place calls, sit on hold and navigate phone trees, limited to US Pixel 11 owners with a paid subscription and the beta Phone app (TechCrunch) Google DeepMind's Koray Kavukcuoglu said Gemini 4 is in post-training and the intention is to release an early post-training output "as soon as possible," much earlier than the end of 2026 (The Information, via 9to5Google) Gemini 3.8 Flash scores 10.37% on ARC-AGI-3 at high reasoning under the standard harness and 35.00% under a provider-adapter harness, a gap ARC Prize has now flagged for multiple frontier models (ARC Prize) On Vending-Bench 2, starting from $500 over a simulated year, GPT-6 Sol ends at $14,428, Grok 4.7 at $10,537 and Opus 5.5 at $9,235, and all three deceived suppliers about prices (Andon Labs) SciUniverse Level 1 puts models in charge of a real lab across 92 tasks: Claude Fable 5.1 leads at 45.3% Pass@1 against GPT-6 Astra's 32.5%, and the failures are physical, including pipetting frozen samples and vortexing open plates (C5R) EvasionBench finds agents circumventing runtime monitors under ordinary task pressure with no adversarial prompting, with best-of-3 attempt rates up to 98% and success up to 88%, and evasion rising with reasoning effort (arXiv) ChatGPT Ads expanded to Southeast Asia and Taiwan, now 60+ countries, shown only to Free and Go tiers, with OpenAI claiming a $1B annualized run rate in under 200 days (OpenAI) Anthropic launched the Claude Marketplace with 2,000+ connectors and plugins, and lets enterprises apply part of their committed Anthropic spend to third-party products listed on it (Anthropic) In Project Swap, 201 Anthropic employees let Claude agents trade books on their behalf: Claude matched their preference ordering 61% of the time, hit 0.55 efficiency against a 0.89 optimum, and 85% of the shortfall was preference representation, not negotiation (Anthropic) Claude Code cloud sessions left research preview and are now generally available on Pro, Max and Team, plus Enterprise premium seats (Anthropic) Island raised a $400M Series F at a $6.4B valuation led by Evolution Equity Partners, saying it has doubled ARR every fiscal year since launching in 2022 (Island) Brahma raised $150M at a $2B valuation led by Multiples Alternate Asset Management, with another $100M of interest it may take (CNBC) LiveKit acquired Loophole Labs and its Substrate hypervisor, an eight-engineer team, aiming to ship agent startup times under three seconds within weeks; terms undisclosed (LiveKit) Databricks acquired Row Zero, a spreadsheet platform handling billions of rows, to integrate natively with Genie; terms undisclosed (Databricks) Australian data center operator Firmus expects a $77M first-half pro forma loss ahead of a roughly $5B ASX IPO, with the bookbuild opening October 6th and listing October 22nd (Reuters, via The Star) Oracle issued a force majeure notice to Blue Owl over Project Jupiter, a 2.5GW New Mexico campus, after state regulators rejected a natural gas pipeline extension in July (Bloomberg, via Data Center Dynamics) Nscale's Loughton site, pitched as the UK's largest AI supercomputer, slipped from 2027 to the early-to-mid 2030s because UK Power Networks cannot supply the power (The Guardian) Columbia's Stijn Van Nieuwerburgh puts the US AI buildout at $10.3 trillion from 2025 to 2032, averaging 3.63% of GDP a year, and warns the financing is moving from corporate balance sheets to JVs, private credit and SPVs (Brookings) Amazon is putting more than $100M into a 585,000 square foot advanced manufacturing plant in Greenwood, Indiana, with 300 jobs averaging near $100,000 and operations by 2028 (Amazon) Tower Semiconductor detailed its Japan photonics buildout, converting the Arai fab to 300mm silicon photonics and targeting 45,000 300mm wafers a month by 2029, with $1B of the $4B coming from METI (Data Center Dynamics) Alibaba followed Apsara with consumer hardware: the Qwen Book, Qwen Glasses N1 and N1 Pro with 50MP capture and iris payment on the Pro, Qwen Clip earbuds co-engineered with Bose, and the QwenNote A2 at 1,199 yuan, on sale October 13th (TechNode Global) Qwen-Image-2.1 landed with a 7B generation component under the Qwen Research License, native RGBA output and up to 10 reference images, with no benchmark table on the model card (Hugging Face) OrcaSAQ-2 compresses Qwen3.8-27B from 54GB to 12.3GB at 3.21 bits per weight while holding 70.0 on SWE-bench Verified and 93.2% token-level top-1 agreement (Hugging Face) Google, OpenAI and Anthropic are reportedly standing up an industry safety standards body called SAFA for late 2026 or early 2027, with Sriram Krishnan approached to lead it; Meta and xAI are not participating (The Information, via Proactive Investors) Microsoft AI chief Mustafa Suleyman said the industry needs an agreed red line on advanced AI capabilities and asked for government support on evaluation, without naming where the line sits (Bloomberg) OpenEvidence reportedly raised $250M at a $15B valuation, up from $12B in January, and is said to be weighing a sale while pivoting into rare-cancer drug development (Business Insider, via AI Weekly) Perplexity's Fast Search, built on a new Rust retrieval service called Photon, returns 95% of results in 230ms or less at $1.00 per 1,000 requests against $5.00 for standard web search (Perplexity) Combine two text streams linearly and a transformer outputs a superposition of both next-token distributions, an effect the authors find is intrinsic to the architecture, diminishes through pretraining, and can be restored by light fine-tuning (arXiv) Dropping an agent's historical reasoning once derived state has been written to files or tool output raised reward from 0.699 to 0.718 on 260 WorkBuddyBench tasks while cutting input tokens 25.5% and cache reads 33.3% (arXiv) FlashLoop gets a 1.64x end-to-end speedup and 6x KV cache reduction on looped transformers, training-free and lossless, by exploiting how few tokens actually change between loop iterations (arXiv)
| | | The Antigravity SDK now runs agents against local models via Google AI Edge's LiteRT, with Gemma 4 26B A4B as the reference and a recommendation of more than 24GB of VRAM or unified memory. The interesting pattern is Architect-Builder: cloud Gemini 3.8 Flash plans, a local swarm executes. In Google's own demo that split 3,322 tokens locally and offline against 95 cloud tokens for planning. It plugs into any OpenAI-compatible server, so Ollama, LM Studio and vLLM all work. The Python SDK is Apache 2.0. | Cursor shipped Rollouts and a rebuilt Security Reviewer. Rollouts writes a monitoring plan before a change merges, watches it through deploy against a per-environment baseline via Datadog, Grafana or Honeycomb, reports verified healthy, regression detected or inconclusive, and can open the revert PR itself. Security Reviewer traces user input flow instead of pattern matching, catching injection, authorization bypass, exposed secrets and unsafe deserialization. Cursor's numbers: average review time down from 4.8 to 3.8 minutes, comment acceptance up from 45 to 50% to 60 to 70%. Teams and Enterprise only. | Managed Deep Agents 0.8 adds user-owned credential scoping across 23 pre-configured services, a two-layer memory split between agent-level and user-level keyed to authenticated identity, a new HTTP channel for webhooks, and Slack file transfer. Web search is powered by Parallel with no separate account, and that search specifically is free while MDA is in beta. LangChain also shipped Trajectories, a flattened chronological view of an agent run across subagents, and fine-tuning in public beta. | | Thank you for reading today's edition. | | Your feedback is valuable. Respond to this email and tell us how you think we could add more value to this newsletter. | Interested in reaching smart readers like you? To become an AI Breakfast sponsor, reply to this email or DM us on X! | Thinking of starting your own newsletter? AI Breakfast readers who sign up with Beehiiv receive a 14-day free trial and 20% off for 3 months. |
|