🪄 Bring your 3D characters to life with words! Introducing UniMate: one unified model for text-driven animation across diverse skeletons—from humans and animals to articulated objects. 🎬 A rigged asset + a text prompt → motion. No per-skeleton retraining needed! 🌐 linzhanmou.com/unimate/
25
83
920
111,939
🧱 Why just look at a LEGO set when you can take it apart and watch it build itself? I built an interactive 3D playground to explore sets, inspect every brick, and replay their assembly, right in your browser. Take the Helicarrier for a spin! 🚁 linzhanmou.com/resources/leg…
50
Encoder-free, pixel-space, unified understanding + generation for images and video. Very cool work from @CongWei1230
Let's remove VAEs and ViTs from video models! 🚀 Introducing 𝗣𝗶𝘅𝗲𝗹𝗨𝗠𝗠: an encoder-free unified multimodal model for image and video understanding and generation, directly in pixel space. Code and model available today! 🌐 nv-tlabs.github.io/PixelUMM 📄 arxiv.org/abs/2609.38597
1
26
4,173
Linzhan Mou retweeted
⚡ One policy, millions of embodiments, over 200 robot models. Can we add yours? We're building γ₀, a generalist RL policy for motion control trained across millions of randomized embodiments derived from a growing collection of more than 200 robot models.
14
41
357
39,641
Linzhan Mou retweeted
New Open-Source AI Animator: Text-to-Animation for Humans, Animals, Creatures, Robots & More UniMate generates motion from text prompts for rigged 3D characters. Describe an action and turn it into animation. Highlights: • One model, different skeletons—no separate retraining for each rig. • Generate transitions between existing keyframes. • Edit motion with text while keeping selected joints unchanged. • Extend animations with a sequence of prompts. • Export animated meshes as FBX and GLB. • MIT-licensed code + downloadable preview weights. github.com/Friedrich-M/UniMa…
41
260
2,773
238,197
Linzhan Mou retweeted
This was a really challenging task when we first started thinking about a unified generative motion prior for large-scale, diverse animation skeletons! It took Linzhan months to prepare the first version of the dataset, and a lot of effort to figure out how to model these irregular structures. Now the data and models are all available!
🪄 Bring your 3D characters to life with words! Introducing UniMate: one unified model for text-driven animation across diverse skeletons—from humans and animals to articulated objects. 🎬 A rigged asset + a text prompt → motion. No per-skeleton retraining needed! 🌐 linzhanmou.com/unimate/
1
19
2,237
Linzhan Mou retweeted
One model for text-driven motion, without per-rig retraining.
Linzhan Mou
10
104
5,991
Linzhan Mou retweeted
🔊 My first tweet, for the shared work! Excited to share our work, UniMate, with labmate @LinzhanMou and the team! 🧑‍🎨 One unified model, diverse rigs, prompt in, motion out. A step toward building efficient, high-fidelity tools that free artists from tedious work and give them more time to create. Checkout our work at SIGGRAPH Asia, which I will be attending this Dec. Feel free to reach out! 🤖 Looking ahead, a shared representation across diverse embodiments could extend beyond animation to support robot control and action prediction. #SIGGRAPHAsia2026 #Animation #GenerativeAI #3D #MotionGeneration #AI #AIGC #Robotics #RobotLearning #EmbodiedAI
🪄 Bring your 3D characters to life with words! Introducing UniMate: one unified model for text-driven animation across diverse skeletons—from humans and animals to articulated objects. 🎬 A rigged asset + a text prompt → motion. No per-skeleton retraining needed! 🌐 linzhanmou.com/unimate/
1
11
871
Linzhan Mou retweeted
One model to control/animate them all: humans, animals, and articulated objects... Check out @LinzhanMou’s UniMate: linzhanmou.com/unimate/ One unified model can animate diverse rigged assets from text prompts, with no per-skeleton retraining. This idea could go beyond kinematic motion generation— a unified representation across different embodiments could also be very useful for robot control and action prediction. Very nice work from @LinzhanMou! Congrats! #SIGGRAPHAsia2026 #Animation #Robotics #RobotLearning #EmbodiedAI #GenerativeAI #3D #MotionGeneration #AI #AIGC
🪄 Bring your 3D characters to life with words! Introducing UniMate: one unified model for text-driven animation across diverse skeletons—from humans and animals to articulated objects. 🎬 A rigged asset + a text prompt → motion. No per-skeleton retraining needed! 🌐 linzhanmou.com/unimate/
2
5
41
11,349
🪄 Bring your 3D characters to life with words! Introducing UniMate: one unified model for text-driven animation across diverse skeletons—from humans and animals to articulated objects. 🎬 A rigged asset + a text prompt → motion. No per-skeleton retraining needed! 🌐 linzhanmou.com/unimate/
25
83
920
111,939
UniMate will be presented at @SIGGRAPHAsia 2026! Code, dataset, and model checkpoints are now released! 🎮 Demo: linzhanmou.com/unimate/inter… 📄 Paper: arxiv.org/abs/2609.05415 💻 Code: github.com/Friedrich-M/UniMa… 🤗 Dataset & checkpoints: huggingface.co/collections/L… Huge thanks to our amazing team: @JiahuiLei1998, @frankzydou, @chenyueccai, @chaoyue_song, Adam Finkelstein, and Szymon Rusinkiewicz! Special thanks to @frankzydou for his valuable insights and suggestions on the demo design!
3
4
68
8,452
Linzhan Mou retweeted
GPT-6 Astra scored 95% on a robot control task, up from Fable 5.1's 40%, with 6.2x fewer output tokens at 2.3x lower cost. 🧵
167
636
5,767
1,972,818
Linzhan Mou retweeted
Introducing GEN-1.5, a one-shot learner. It can learn new tasks in a few seconds. Show it what to do, and it generalizes. This capability emerged from pretraining on physical data at scale, as a step towards our mission of building general intelligence for the physical world.
322
1,687
12,163
3,422,684
Linzhan Mou retweeted
Robot policies can move but can't think. LLMs can think but can't move. So we connected them. Real robot: 16.7% → 97.3% Sim (LIBERO-PRO): 12.8% → 53.3%
24
54
305
95,804
Linzhan Mou retweeted
In my recent blog post, I argue that "vision" is only well-defined as part of perception-action loops, and that the conventional view of computer vision - mapping imagery to intermediate representations (3D, flow, segmentation...) is about to go away. vincentsitzmann.com/blog/bit…
43
162
1,054
397,610
Linzhan Mou retweeted
Step inside Project Genie: our experimental research prototype that lets you create, edit, and explore virtual worlds. 🌎
960
4,189
33,923
13,487,736
Linzhan Mou retweeted
A SINGLE encoder + decoder for all the 4D tasks! We release 🎯 D4RT (Dynamic 4D Reconstruction and Tracking). 📍 A simple, unified interface for 3D tracking, depth, and pose 🌟 SOTA results on 4D reconstruction & tracking 🚀 Up to 100x faster pose estimation than prior works
18
70
444
119,736
Linzhan Mou retweeted
Check out DIMO (linzhanm.github.io/dimo/, Highlight) 🌀at Booth #410 at #ICCV2025 (Wed, morning session). From a single image, we distill video-model priors into a motion latent space to sample diverse 3D motions (neural keypoint trajectories + 3DGS).
8
54
4,265
Linzhan Mou retweeted
Introducing DINOv3: a state-of-the-art computer vision model trained with self-supervised learning (SSL) that produces powerful, high-resolution image features. For the first time, a single frozen vision backbone outperforms specialized solutions on multiple long-standing dense prediction tasks. Learn more about DINOv3 here: ai.meta.com/blog/dinov3-self…
343
734
4,381
901,530
Linzhan Mou retweeted
Another one. Already a powerful painting, but moving around it yourself gives a totally different feeling. Jacques Louis David's "The Death of Socrates" => #Genie3
133
296
2,664
322,013