SemanTok: Predictable Semantic Tokens for Efficient Autoregressive Video Generation
Paper • 2610.00686 • Published • 10
Our vibrant communities consist of experts, leaders and partners across the globe. They are developing cutting-edge open AI models for Image, Language, Audio, Video, 3D and Biology.
4Director: Controlling Video World Models with Rigid 3D Geometry
SemanTok: Predictable Semantic Tokens for Efficient Autoregressive Video Generation