Modern multimodal models aren't a single decode loop anymore; they're composite.
Modern multimodal models aren't a single decode loop anymore; they're composite. M* is one runtime that serves them all, and it matches or beats every specialized system: up to 2.7× on omni TTS, 12.5× on world-model rollouts. Learn more here: https:// ai.stanf