- AI Valley
- Posts
- New Atlas world model is literally black magic
New Atlas world model is literally black magic
PLUS: Anthropic released Mythos 5.1 a
Together with
howdy, it’s Barsee again.
happy wednesday, AI family, and welcome back to AI Valley.
here are the biggest things worth knowing today:
Anthropic released Claude Fable 5.1 and Mythos 5.1
World Labs has a new world model called Atlas
Meta has a new real-time speech model
Plus trending AI tools, posts, and resources
Let’s dive into the Valley of AI…
WISPR FLOW
Some problems don't solve themselves at a keyboard. Flow is perfect for walks: talk through the idea, the structure, the tradeoffs—then get back a clean draft you can paste into your doc or AI tool. It's like having a writing assistant that keeps up with your brain in motion.
*This is sponsored
THROUGH THE VALLEY
1/ Anthropic released Claude Fable 5.1 and Mythos 5.1 - they’re actually the same underlying model, but Fable is available to everyone while Mythos has fewer safeguards for cybersecurity and biology and is only available to vetted organizations. (read the announcement)
Fable got a pretty big science/coding bump. It scored 52.6% on Terminal-Bench-Science vs 24.7% for Fable 5, and Anthropic says Mythos is even better at some of the harder research stuff. It designed protein binders with a ~50% hit rate across 12 targets and sped up GPU kernels for biology models by up to 2.5x.
Pricing is still $10/$50 per million input/output tokens, but cache reads are 75% cheaper. Anthropic says that makes typical workloads ~25% cheaper and highly agentic ones up to ~45% cheaper.
Some fun Fable stuff already:
COLD WATCH is Ethan Mollick’s retro space-survival game built with Fable 5.1. A pretty fun way to see what a long-running coding agent can actually build.
Cat Doom is exactly what it sounds like... a Doom-style FPS with cats. Fable 5.1 built the thing live on stream.
2/ World Labs has a new world model called Atlas - it can take images, video and camera movements and turn them into a 3D world you can move around in. You can basically give it a few photos, tell the camera where to go and generate new image/video frames from that exact viewpoint. (see the example here) (another one here)
They showed it generating a 1-minute 1440p video from seven reference images, and it can reconstruct explicit 3D from just one or a few images. More images = less guessing about what the world actually looks like.
The robotics use case is pretty cool too. Atlas can take a few photos of a real space and simulate what a robot’s cameras/depth sensors would see while moving through it. So instead of scanning every training environment with expensive equipment, you could potentially just take some photos and generate the rest.
It’s all one model, trained from scratch, that mixes ideas from LLMs and video models. World Labs says Atlas can perceive, generate and reason about both virtual + physical worlds.
Early access is coming in the next few weeks. (read the blog here)
3/ Meta has a new real-time speech model - Muse Voice Transcribe is Meta’s first real-time audio perception model. It does speech-to-text, figures out who’s speaking and knows when someone has stopped talking, all in one model.
The clever bit is it decides how long to listen before writing each word. Meta says this gets it pretty close to the sweet spot between speed and accuracy.
It’s trained across 70+ languages, can handle people switching languages mid-sentence and apparently hour-long conversations with 20+ speakers. And it costs something like $3 per 1,000 voice minutes.
It’s already being used for dictation in Meta’s desktop app and voice input in Muse Code. Live on the API now too. (read the announcement by Zuck)
TRENDING TOOLS
Orbis 1.0 - create living worlds and stream them in real time, with persistent memory, interactivity, and physics-grounded generation of unbounded length
Atlas - the world's first multimodal world model that generates image and video frames with pixel-perfect camera control and reconstructs them in 3D
Claude Fable 5.1 - Anthropic’s new top-ranked model upgrade
Dyson - launches a $499 AI toothbrush with a built-in camera that identifies gaps between teeth & automatically squirts them with mouthrinse
Perplexity hybrid - this will allow Computer to orchestrate local models that can run locally on Mac, particularly for agent steps involving sensitive and private files
Caddi - builds AI agents by recording narrated screenshares to automate repetitive back-office work across your existing software
Google Pics - lets you edit individual objects, refine or translate text, and collaborate with your team
WHAT I'M CONSUMING
OpenAI prepares to release Astra with "critical" cyber capabilities to answer Anthropic’s Mythos
The world's first multimodal world model that generates image and video frames with pixel-perfect camera control and reconstructs them in 3D
Anthropic made an evil model. Of course I have to do a video
Nobody is talking seriously about AI demand
Things are now happening much faster than AI 2027
THE VALLEY GEMS
THAT’S ALL FOR TODAY
Thank you for reading today’s edition. That’s all for today’s issue.

💡 Help me get better and suggest new ideas at [email protected] or @heyBarsee
👍️ New reader? Subscribe here
Thanks for being here.
REACH 100K+ READERS
Acquire new customers and drive revenue by partnering with us
Sponsor AI Valley and reach over 100,000+ entrepreneurs, founders, software engineers, investors, etc.
If you’re interested in sponsoring us, email [email protected] with the subject “AI Valley Ads”.


