A product discussed on Latent Space.

Moonlake: Interactive, Multimodal World Models — with Chris Manning and Fan-yun Sun
Apr 2, 2026 · 1:06:48
Moonlake AI founders Chris Manning and Fan-yun Sun argue that interactive, multimodal world models require structured symbolic reasoning over pure scale, enabling indefinite multiplayer gameplay and causal consistency that video generation models like Genie and Sora cannot achieve. Their approach uses code engines and physics simulators as cognitive tools, producing reasoning traces that handle geometry, physics, and logic, while a separate diffusion model (Reverie) handles pixel fidelity. They aim to replace traditional rendering and empower creators by allowing human intent to be injected at a symbolic layer. Manning contrasts this with Yann LeCun's JEPA, emphasizing language and abstraction over pixel-level prediction. Moonlake is hiring engineers at the intersection of code generation, computer vision, and graphics.

A Technical History of Generative Media
Sep 8, 2025 · 1:04:44
Fal.ai founders Gorkem and Batuhan detail their pivot from dbt pipelines to generative media inference, which now serves 2M developers and 350 models, crossed $100M ARR, and raised a $125M Series C. Key inflection points included Stable Diffusion 1.5 (company pivot), SDXL (first $1M revenue), Flux (jump from $2M to $10M monthly revenue), and Veo 3 (text-to-video with perfect lip-sync). They built a proprietary inference engine with 100+ custom kernels and a serverless GPU stack managing 10,000+ H100 equivalents across 6 cloud providers, typically delivering 1.5-10x speedups over stock PyTorch. Video models now drive 50% of revenue, up from 18% in February, fueled by open-source models like Hunyuan and partnerships with closed labs such as Play.ht and Google DeepMind. They argue advertising is the killer application for generative media and that image/video RL, specialized data pipelines, and cheaper conversational video models are underexplored startup opportunities.

AI Video Is Eating The World — Olivia and Justine Moore, a16z
Jul 9, 2025 · 49:28
Alessio and Shawn (hosts of Latent Space) talk with a16z partners Justine and Olivia Moore about the explosion of AI-generated video on TikTok, Instagram, and YouTube. The Moores trace how trends moved from Reddit to consumer platforms, with examples like Italian Brainrot, Kim the Gorilla, and fruit-slicing ASMR. They detail their own hands-on creation: Olivia used Veo 3 and MiniMax to make viral clips (gold bar squishing, Tide Pod consumption), spending up to 8 generations per usable video and hitting Veo 3’s $125/month plan limits. They explain monetization paths—creator fund payouts (~$20 per million views), merch (Breadclump sweatshirts), consulting, and licensing to Netflix—and recommend tools like Overlap (clips 30–140 seconds) and ComfyUI for control. The episode closes with prompt theory: AI characters questioning their existence, and a bet that Gen Alpha’s brainrot consumption will reshape long-form content.
Powered by PodHood