Sign up | Follow us on X | Sponsor | | Together with | | | howdy, it’s Barsee again. | happy wednesday, AI family, and welcome back to AI Valley. | here are the biggest things worth knowing today: | Anthropic released Claude Fable 5.1 and Mythos 5.1 World Labs has a new world model called Atlas Meta has a new real-time speech model Plus trending AI tools, posts, and resources
| Let’s dive into the Valley of AI… |
|
| | | | | | WISPR FLOW | |  | Courtesy: Wispr Flow |
| Some problems don't solve themselves at a keyboard. Flow is perfect for walks: talk through the idea, the structure, the tradeoffs—then get back a clean draft you can paste into your doc or AI tool. It's like having a writing assistant that keeps up with your brain in motion. | Try Flow Free | *This is sponsored |
| |
| |
| | |
| | | | | | THROUGH THE VALLEY |  | Introducing Claude Fable 5.1 |
|
| 1/ Anthropic released Claude Fable 5.1 and Mythos 5.1 - they’re actually the same underlying model, but Fable is available to everyone while Mythos has fewer safeguards for cybersecurity and biology and is only available to vetted organizations. (read the announcement) | Fable got a pretty big science/coding bump. It scored 52.6% on Terminal-Bench-Science vs 24.7% for Fable 5, and Anthropic says Mythos is even better at some of the harder research stuff. It designed protein binders with a ~50% hit rate across 12 targets and sped up GPU kernels for biology models by up to 2.5x. | Pricing is still $10/$50 per million input/output tokens, but cache reads are 75% cheaper. Anthropic says that makes typical workloads ~25% cheaper and highly agentic ones up to ~45% cheaper. | Some fun Fable stuff already: | COLD WATCH is Ethan Mollick’s retro space-survival game built with Fable 5.1. A pretty fun way to see what a long-running coding agent can actually build. Cat Doom is exactly what it sounds like... a Doom-style FPS with cats. Fable 5.1 built the thing live on stream.
| 2/ World Labs has a new world model called Atlas - it can take images, video and camera movements and turn them into a 3D world you can move around in. You can basically give it a few photos, tell the camera where to go and generate new image/video frames from that exact viewpoint. (see the example here) (another one here) | They showed it generating a 1-minute 1440p video from seven reference images, and it can reconstruct explicit 3D from just one or a few images. More images = less guessing about what the world actually looks like. | The robotics use case is pretty cool too. Atlas can take a few photos of a real space and simulate what a robot’s cameras/depth sensors would see while moving through it. So instead of scanning every training environment with expensive equipment, you could potentially just take some photos and generate the rest. | It’s all one model, trained from scratch, that mixes ideas from LLMs and video models. World Labs says Atlas can perceive, generate and reason about both virtual + physical worlds. | Early access is coming in the next few weeks. (read the blog here) | 3/ Meta has a new real-time speech model - Muse Voice Transcribe is Meta’s first real-time audio perception model. It does speech-to-text, figures out who’s speaking and knows when someone has stopped talking, all in one model. | The clever bit is it decides how long to listen before writing each word. Meta says this gets it pretty close to the sweet spot between speed and accuracy. | It’s trained across 70+ languages, can handle people switching languages mid-sentence and apparently hour-long conversations with 20+ speakers. And it costs something like $3 per 1,000 voice minutes. | It’s already being used for dictation in Meta’s desktop app and voice input in Muse Code. Live on the API now too. (read the announcement by Zuck) |
| |
| |
| | |
| | | | | | Orbis 1.0 - create living worlds and stream them in real time, with persistent memory, interactivity, and physics-grounded generation of unbounded length Atlas - the world's first multimodal world model that generates image and video frames with pixel-perfect camera control and reconstructs them in 3D Claude Fable 5.1 - Anthropic’s new top-ranked model upgrade Dyson - launches a $499 AI toothbrush with a built-in camera that identifies gaps between teeth & automatically squirts them with mouthrinse Perplexity hybrid - this will allow Computer to orchestrate local models that can run locally on Mac, particularly for agent steps involving sensitive and private files Caddi - builds AI agents by recording narrated screenshares to automate repetitive back-office work across your existing software Google Pics - lets you edit individual objects, refine or translate text, and collaborate with your team
|
| |
| |
| | |
| | | | | | | OpenAI prepares to release Astra with "critical" cyber capabilities to answer Anthropic’s Mythos The world's first multimodal world model that generates image and video frames with pixel-perfect camera control and reconstructs them in 3D Anthropic made an evil model. Of course I have to do a video Nobody is talking seriously about AI demand Things are now happening much faster than AI 2027
|
| |
| |
| | |
| | | | | | | What’s trending on social today: | | | |
| |
| |
| | |
| | | | | | | | | | | REACH 100K+ READERS | Acquire new customers and drive revenue by partnering with us | Sponsor AI Valley and reach over 100,000+ entrepreneurs, founders, software engineers, investors, etc. | If you’re interested in sponsoring us, email barsee@aivalley.ai with the subject “AI Valley Ads”. |
| |
| |
| | |
|
|