Yann LeCunZhang YimingFei-Fei LiAlphabetByteDancePicoMeta PlatformsDeepSeekApple Inc.Moonshot AIAlibaba Group Holding

ByteDance founder Zhang Yiming personally oversees new spatial-video AI model slated for October launch; takes on Meta and Alphabet in world models race

ByteDance founder Zhang Yiming is personally overseeing the development of a new spatial-video AI model, set for an October launch, aimed at competing with Meta and Alphabet. The model will enable real-time generation of interactive virtual worlds, enhancing ByteDance's capabilities in AI and cloud computing.

Bloomberg.com Bloomberg.com+1 source8 September 2026 · 05:10 UTC
CuriousCats Full Story

ByteDance is preparing to launch a new AI model for real-time spatial video generation, with founder Zhang Yiming personally overseeing its development. Slated for an October release, this model aims to compete with tech giants Meta and Alphabet in the burgeoning field of virtual and mixed reality.1456

The model, built on ByteDance's existing Seedance technology, will allow users to create interactive virtual worlds for applications such as live streams, short-form dramas, and games. Zhang is coordinating efforts across various business units, leveraging AI resources and computing capacity to enhance the model's capabilities.2

Zhang hopes this new AI model will position ByteDance as a key player in the world models arena, joining experts like Fei-Fei Li and Yann LeCun in exploring visual AI approaches crucial for robotics and gaming. The spatial-video model aims to create immersive experiences, placing users in three-dimensional environments similar to Google's Genie.

If successful, this initiative could open a new front in ByteDance's rivalry with Meta and Apple, both of which have heavily invested in virtual and mixed-reality technologies. The model is designed to respond to user interactions, offering on-demand videos with minimal latency, thus enhancing user engagement in spatial-computing environments.

Additionally, the model seeks to reduce the cost of VR adoption by shifting the computational load to the cloud, making it more accessible for users.

Key Insight
“The model, built on ByteDance's Seedance, would generate interactive virtual worlds for live streams and games, responding to Pico headset users' voices or movements. It offers on-demand videos with about 0.05 seconds latency at 20 frames per second, aiming to lower VR adoption costs by moving spatial content generation to the cloud.”
CuriousCats studied:
1
Bloomberg.comBloomberg.com
“is readying an AI model geared for real-time spatial video generation, taking on and in an arena with applications in robotics and autonomous systems.”
Bloomberg.com →
2
The Straits TimesThe Straits Times
“ByteDance is readying an AI model geared for real-time spatial video generation, taking on Meta Platforms and Alphabet in an arena with applications in robotics and autonomous systems.”
The Straits Times →
Ask CuriousCats
Who is Zhang Yiming?
What is ByteDance's new AI model?
Why is spatial-video important for VR?
Are Meta and Alphabet launching similar models?
How does this AI model compare to competitors'?
Get your CIA-level briefing,
in real time.
CuriousCats monitors the internet every minute for you and brings you the most personalized brief of videos, social media posts, news and more.
Download the App
Liked the depth here?
Get the full internet briefed for you any time of the day.
Get CuriousCats