<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>Next-State-Prediction on MLLog.dev</title><link>https://mllog.dev/en/tags/next-state-prediction/</link><description>Recent content in Next-State-Prediction on MLLog.dev</description><image><title>MLLog.dev</title><url>https://mllog.dev/images/default_mllog.png</url><link>https://mllog.dev/images/default_mllog.png</link></image><generator>Hugo -- 0.147.9</generator><language>en</language><lastBuildDate>Sun, 05 Jul 2026 08:00:00 +0100</lastBuildDate><atom:link href="https://mllog.dev/en/tags/next-state-prediction/index.xml" rel="self" type="application/rss+xml"/><item><title>Orca: What If Next-Token, Next-Frame, and Next-Action Are the Same Task?</title><link>https://mllog.dev/en/posts/orca-world-foundation-model-next-state-prediction/</link><pubDate>Sun, 05 Jul 2026 08:00:00 +0100</pubDate><guid>https://mllog.dev/en/posts/orca-world-foundation-model-next-state-prediction/</guid><description>Orca (BAAI) replaces next-token, next-frame, and next-action prediction with a single Next-State-Prediction objective. A frozen 4B backbone pre-trained on 12.5K hours of video — with zero action labels — feeds three lightweight readouts, and the action readout, trained on just 200 trajectories per task, beats π0.5 on OOD robot manipulation (32.4 vs 29.4).</description></item></channel></rss>