In partnership with |  |
| | Good morning. It's Wednesday, September 16th. | | | | You read. We listen. Let us know what you think by replying to this email. | A free newsletter read by 117,000 marketers | | The best marketing ideas come from marketers who live it. | That’s what this newsletter delivers. | The Marketing Millennials is a look inside what’s working right now for other marketers. No theory. No fluff. Just real insights and ideas you can actually use—from marketers who’ve been there, done that, and are sharing the playbook. | Every newsletter is written by Daniel Murray, a marketer obsessed with what goes into great marketing. Expect fresh takes, hot topics, and the kind of stuff you’ll want to steal for your next campaign. | Because marketing shouldn’t feel like guesswork. And you shouldn’t have to dig for the good stuff. | Sign Up Free |
|
| | Apple shipped the new Siri. Google built a chunk of the brains. | Apple released Siri AI on Monday, in beta, English only, with French, Japanese, Korean, Portuguese and Spanish due next month. It does personal context across your messages, mail and photos, on-screen awareness, and in-app actions, plus Visual Intelligence, Writing Tools, Image Playground, Siri Recap and Live Rewind. Server-side operations carry daily usage limits while it is in beta. | The features run on "the next generation of Apple Foundation Models, custom-built in collaboration with Google and its Gemini models." Apple spent two years insisting the personal-context Siri was coming; it arrived leaning on a competitor's model. Supported hardware is iPhone 16 or later plus iPhone 15 Pro and Pro Max, M1-or-later iPads and Macs, MacBook Neo, Vision Pro, and Apple Watch Series 9 or later. It will not ship initially in the EU on iOS, iPadOS and watchOS, and will not ship in China while Apple works through regulatory requirements. | The same week Siri landed, X Corp and xAI dropped Apple from the antitrust suit they filed in August 2025 over ChatGPT's exclusive iPhone integration. They are still suing OpenAI. No reason given, no settlement disclosed. Make of that what you will now that Siri's default brain is not exclusively OpenAI's. | Read more. | | Jensen Huang breaks the “slowdown” consensus | At Salesforce's Dreamforce on Tuesday, Jensen Huang said the quiet part into a microphone: "Safety is an engineering problem, not a legal one. We don't need any new laws. We don't need new regulations." His proposed mechanism is product liability plus self-restraint: "if you build a product or a service, and you're not confident in its safety, then don't release it." He called the tradeoff between speed and safety "a false choice." | That is three days after Amodei's essay and two days after Altman, Musk and Nadella all signed on. It is also the opposite direction from where OpenAI spent Tuesday: policy chief Chris Lehane told reporters in Washington that OpenAI, Anthropic and Google DeepMind have been coordinating on safety for several weeks, that they do not need a government antitrust waiver to do it, and that OpenAI backs the FRONTIER Act provision requiring frontier labs to admit independent verification organizations. Amodei asked Washington for that waiver by name. OpenAI says it is unnecessary. | Huang was on that stage to sell something, for the record. Salesforce used the slot to launch Koa, its first CRM reasoning model, post-trained from NVIDIA's open Nemotron 3 Super on synthetic scenarios across 14-plus industries with no customer data. Salesforce claims Koa "matches or exceeds leading model performance on CRM actions with three times fewer errors" on its own internal benchmark, keeps the weights, and runs inference inside its own trust boundary. Pilots are running now with Formula 1, UChicago Medicine, Xero and 1-800Accountant, and general availability is slated for winter 2026 in US regions. | Read more. | | China's answer to pacing week: $5B and a 753B-parameter MIT license | Z.ai, the Hong Kong-listed lab behind the GLM family, closed roughly $5 billion on Monday: about $2B in a share placement and about $3B in convertible bonds. The stated use of proceeds is next-generation GLM foundation models, a fully self-training system, training and inference compute, and domestic-chip adaptation. SCMP reports the placement priced at HK$714 a share, a 9.96% discount, against HK$1,588 in July and a June peak of HK$2,980. The raise is enormous and the stock is down by three quarters. Both things are true. | A day before that, Shanghai AI Laboratory put Atria Dawn Preview on Hugging Face: 753B parameters, MIT license, 256K context, built on Z.ai's 744B GLM-5.2 MoE foundation. Posted scores include 96.0 on DeepSearchQA, 86.5 on CyberGym, 86.2 on MLE-bench Lite, 77.0 on BFCL v4 and 65.0 on Workspace-Bench. The accompanying paper claims top scores on five of sixteen benchmarks and reports that human participants rated about a third of completed AI-assisted tasks as infeasible without AI. Treat self-reported numbers as self-reported, but note the license: an agentic model at this scale, free to take. | The rest of the week's response came in four registers. China Daily called the slowdown push "a self-serving bid to preserve American dominance." The Foreign Ministry called it fearmongering. A DeepSeek engineer who worked on V4.1 posted that he especially does not want Anthropic holding the most advanced AI. And at the Xiangshan Forum, defense minister Dong Jun urged nations to "consult with one another on governance." Nobody in Beijing is pacing anything. | Read more. | | Sanders and Bannon shared a stage at a Future of Life Institute assembly of nearly 300, with Sanders pushing a Casar bill to permanently ban superintelligence and urging Trump to seek a pause treaty with Xi (NPR) OpenAI's Chris Lehane said the three labs have coordinated on safety for weeks and do not need an antitrust waiver, contradicting Amodei's ask (TechCrunch) China Daily called the US slowdown push "a self-serving bid to preserve American dominance" (Bloomberg) China's defense minister used the Xiangshan Forum to call for international consultation on AI governance (Bloomberg) A DeepSeek engineer who worked on V4.1 said he does not trust Anthropic or OpenAI to keep advanced AI open and affordable (South China Morning Post) US markets caught up Monday: Nvidia fell more than 3% and the Philadelphia Semiconductor Index almost 6%, while Alphabet, Microsoft and Meta all rose (Fortune) X Corp and xAI dropped Apple from their August 2025 antitrust suit over ChatGPT's iPhone exclusivity, keeping their claims against OpenAI alive (Al Jazeera) Meta is testing its third-generation MTIA chip, codenamed Arke, with a fourth generation called Astrid behind it, built with Broadcom and TSMC (Bloomberg) Salesforce launched Koa, a CRM reasoning model post-trained from NVIDIA's Nemotron 3 Super, claiming three times fewer errors on its internal benchmark (Salesforce) NVIDIA says Vera Rubin NVL72 delivers up to 30x higher throughput per megawatt than GB300 NVL72 on DeepSeek V4 Pro, per SemiAnalysis benchmarking (NVIDIA) Lambda ran 19 nodes inside a 16-node power budget using NVIDIA's DSX MaxLPS, lifting cluster throughput about 24% to roughly 5M tokens per second (NVIDIA) NVIDIA shipped CUDA-Q Logical for fault-tolerant quantum work, with Fermilab saying algorithm development went from five months to three weeks (NVIDIA) SK Hynix confirmed early talks with Intel on US memory production, potentially leasing part of Intel's Ohio fab (Korea Herald) Intel spinoff Cornelis Networks raised $205M led by IAG Capital for GPU-agnostic AI networking aimed at cutting GPU idle time (TechCrunch) TSMC reportedly plans to lift 2nm output from 90,000 to 110,000 wafers a month by mid-2027 and double CoWoS packaging to about 260,000 by end-2028 (TrendForce, citing Economic Daily News) Australian data center firm Firmus is reportedly lining up a roughly $5B ASX listing as early as late October, after an August round valued it near $10.5B (Reuters) 404 Media reported OpenAI contractors read and rate real ChatGPT conversations under "Project Lily," with OpenAI acknowledging sensitive details can still get through redaction (404 Media) Jensen Huang is reportedly attending Trump's state dinner for Xi Jinping this week (CNBC) xAI is running a three-day Grok Bot Galaxy build-along in San Francisco through Thursday (xAI) Huawei opened beta for XiaoYi Work, an agent that decomposes goals and hands tasks across HarmonyOS phones, tablets and PCs (TechNode) ByteDance shipped the consumer Doubao Phone Assistant with a beta screen-automation mode, debuting on the nubia NaviX Ultra (TechNode) Spurious Tool Use finds RL-trained agents fire tools on surface cues, with spurious invocation rates rising up to 39% under counterfactual tests (arXiv) Intel's BITCOS packs ternary weights at 1.485 bits each, under the 1.625 bits of five-trit packing, for up to 1.18x CPU and 1.27x GPU inference speedups (arXiv) Tsinghua's Gavel pulls a skill router out of a frozen LLM's own mid-layer states with two linear maps, beating retrieve-and-rerank by up to 21.9 points mid-rollout (arXiv) The Atria Dawn paper reports top scores on five of sixteen agentic benchmarks and humans rating about a third of AI-assisted tasks infeasible without AI (arXiv)
| | | Open Code Review is Alibaba's code review CLI: it reads git diffs, runs a configurable LLM agent with tool use, and emits line-level comments. It ships AACR-Bench alongside it, 200 real PRs across 50 repos and 10 languages with 1,505 ground-truth issues validated by 80-plus senior engineers, where Alibaba claims higher precision and F1 than general-purpose coding agents at about one ninth the tokens. Apache 2.0. | Colibri is a pure-C, zero-dependency inference engine that treats disk, RAM and VRAM as one memory hierarchy and streams MoE experts off disk, so a 744B model runs on hardware you already own. Author-reported: 5.8 to 6.8 tok/s decode on 6x RTX 5090, about 1.8 tok/s on a 128GB CPU-only desktop, 1.07 tok/s on a single RTX 5070 Ti laptop. Apache 2.0. | Perplexity Portable Computer came to Windows, running the agent loop entirely on-device on GeForce RTX and RTX PRO cards with 24GB or more of VRAM, default model Qwen 3.8 27B, with connectors for Outlook, OneDrive, Word, Google Drive, Gmail, Slack and GitHub. Work done locally does not burn Perplexity Computer credits. | Weave Router 2.0 is a drop-in endpoint for Claude Code, Codex and Cursor that classifies task difficulty and routes cheap work to DeepSeek or GLM while sending hard work to the frontier models, with cache-aware switching. Self-reported pass@2: Terminal-Bench 4.0 parity with GPT-6 Astra at 52% of the cost and 2.2x faster. Elastic License 2.0. | Jev is TypeSafe AI's attempt to kill LLM function calling: a purpose-built model that turns unstructured input straight into schema-guaranteed typed output, generated in parallel rather than token by token, in 70 to 500ms. The company claims 40x to 200x faster than comparable LLMs and flags its own workflow benchmark as possibly optimistic. $0.042 per million input tokens, output free, early access. | Capsule bundles an HTML and CSS app plus its data into one portable file backed by embedded SQLite, playable offline on desktop with no account. Rust and Tauri 2.0, free, version 0.4.0, and closer to a curiosity than a product right now. | | Thank you for reading today's edition. | | Your feedback is valuable. Respond to this email and tell us how you think we could add more value to this newsletter. | Interested in reaching smart readers like you? To become an AI Breakfast sponsor, reply to this email or DM us on X! | Thinking of starting your own newsletter? AI Breakfast readers who sign up with Beehiiv receive a 14-day free trial and 20% off for 3 months. |
|