Published event
ArtificialIntelligence ProductUpdate 1 source(s)

Introducing Waypoint-1: Real-time interactive video diffusion from Overworld

Updated September 26, 2026 · 2:45 PM · source date January 20, 2026

Summary

Introducing Waypoint-1: Real-time interactive video diffusion from Overworld Waypoint-1: Real-time Interactive Video Diffusion from Overworld Published January 20, 2026 Update on GitHub Upvote 44 Andrew Lapp lapp0 Overworld Louis Castricato LouisCastricato Overworld Scott Fox ScottieFox Overworld Shahbuland Matiana shahbuland Overworld David Rossi xAesthetics Overworld Waypoint-1 Weights on the Hub Waypoint-1-Small Waypoint-1-Medium (Coming Soon!) Try Out The Model Overworld Stream: https://overworld.stream What is Waypoint-1? Waypoint-1 is Overworld’s real-time-interactive video diffusion model, controllable and prompted via text, mouse, and keyboard.

Why it matters

This ProductUpdate is relevant to the technology intelligence record because it involves GitHub. The source article should remain the factual reference for follow-up coverage.

Key facts
  • Waypoint-1: Real-time Interactive Video Diffusion from Overworld Published January 20, 2026 Update on GitHub Upvote 44 Andrew Lapp lapp0 Overworld Louis Castricato LouisCastricato Overworld Scott Fox ScottieFox Overworld Shahbuland Matiana shahbuland Overworld David Rossi xAesthetics Overworld Waypoint-1 Weights on the Hub Waypoint-1-Small Waypoint-1-Medium (Coming Soon!) Try Out The Model Overworld Stream: https://overworld.stream What is Waypoint-1?
  • Waypoint-1 is Overworld’s real-time-interactive video diffusion model, controllable and prompted via text, mouse, and keyboard.
  • You can give the model some frames, run the model, and have it create a world you can step into and interact with.
  • The backbone of the model is a frame-causal rectified flow transformer trained on 10,000 hours of diverse video game footage paired with control inputs and text captions.
  • Waypoint-1 is a latent model, meaning that it is trained on compressed frames.
  • The standard among existing world models has become taking pre-trained video models and fine-tuning them with brief and simplified control inputs.
Entities in this story
Related events