@MaxForAI: Former NVIDIA and Meta researcher, xAI head of world model Ethan He just announced his departure. Many may not know exactly what he was responsible for. He was a core member of Grok Imagine (including video generation) from the 0-to-1 stage. He said when he joined xAI, Grok Imagine had nothing...
Summary
Ethan He, former head of world model at xAI and key contributor to Grok Imagine video generation, announced his departure. He built the multimodal video model from scratch in three months after joining xAI in July 2025.
View Cached Full Text
Cached at: 05/21/26, 06:39 AM
Former NVIDIA and Meta researcher, xAI world model lead Ethan He has just announced his departure.
Many may not know exactly what he was responsible for.
He was a core member of Grok Imagine (including video generation) from 0 to 1.
He himself said that when he joined xAI, Grok Imagine had nothing: no data, no infrastructure, no model.
Three months later, they shipped a multimodal video model, followed by reference-to-video and video extension.
On a broader scale, his work has always been related to media generation, VLM, and world models — not the typical product manager role for video generation.
He completed his undergraduate studies at Xi’an Jiaotong University, earning a Bachelor’s degree in Computer Science and Technology in 2018. He then pursued graduate studies at Carnegie Mellon University (CMU), entering the Master of Science in Computer Vision (MSCV) program at the Robotics Institute in 2018, and completed his degree in December 2019.
After that, Ethan He joined Meta AI (formerly Facebook AI Research, FAIR). His tenure at Meta AI primarily focused on multimodal learning systems and model optimization techniques for practical applications.
Ethan He joined NVIDIA in 2023 as a Staff Engineer, later serving as a Senior Deep Learning Algorithm Engineer, focusing on large-scale deep learning training frameworks, multimodal models, and mixture-of-experts architectures.
Part of He’s work at NVIDIA involved contributing to the development of the Cosmos world foundation model platform, aimed at accelerating the creation of customized world models for physical AI applications such as robotics and autonomous driving.
Ethan He joined xAI in July 2025, bringing his expertise from NVIDIA to contribute to the development of advanced AI models, particularly in video synthesis through the Grok Imagine project.
Grok Imagine v0.9 was released in early October 2025, improving visual quality, motion dynamics, local audio generation, and generation speed (less than 15 seconds per short video).
This multiplied Grok’s user count several times (though it also brought some controversies).
xAI has indeed seen a number of departures recently, and Ethan’s is one of the more notable among them.
Wishing him all the best in his next chapter.
Ethan He (@EthanHe_42): I’ve left xAI. It’s been quite a journey. I joined when xAI was about to build Grok Imagine from 0 to 1 - no data, no infra, no model. Three months later, we shipped our multimodal video model, followed by reference-to-video and video extension. I’m grateful for the opportunity
Similar Articles
@shen_zheng25741: Today is my last day at xAI. My time at xAI has been intense and incredibly fruitful. I had the chance to work on many …
Shen Zheng announced his departure from xAI, where he worked on improving Grok's search and DeepResearch capabilities.
@0xLogicrw: Noam Shazeer, Google AI key figure and Gemini model technical lead, leaves Google again and officially joins rival OpenAI. OpenAI announced to employees that Shazeer will focus on finding entirely new underlying architectures for large models and advancing the Transformer...
Noam Shazeer, co-author of the Transformer architecture and technical lead of Google's Gemini model, has left Google again and officially joined OpenAI. He will focus on discovering new underlying architectures for large models and driving research into the evolution of Transformers.
Why Video Agent models are next — Ethan He, xAI Grok Imagine (98 minute read)
Ethan He from xAI discusses why video agent models are the next frontier, arguing that video models derive intelligence from LLMs and that the evolution of video generation will mirror AI coding, shifting from one-shot output to multi-turn planning and execution.
@MaxForAI: According to SemiAnalysis @SemiAnalysis_, Jia Yangqing, former Facebook (now Meta) AI Architecture Director, Alibaba Vice President of Technology, President of Alibaba Cloud Intelligent Computing Platform, now founder and CEO of LeptonAI, and NVIDIA Vice President of System Software, @jiayq, has left NVIDIA.
Jia Yangqing, former Facebook AI Architecture Director and LeptonAI founder, left NVIDIA one year after NVIDIA acquired his startup team. According to SemiAnalysis, reasons include DGX Lepton's performance not meeting expectations and disagreements with Jensen Huang over open-source commitments.
@VraserX: https://x.com/VraserX/status/2093563301330346314
Elon Musk testified that xAI used OpenAI's models to train Grok through model distillation, highlighting the ethical and legal controversies surrounding such practices in the AI industry.