MMaDA-VLA
🦾
1
Diffusion VLA predicting robot actions and future frames
Generate AI-powered YouTube Shorts videos
Race a 3D car through a sunset track
Multilayer-geometry 3D from a single image or 16-frame clip
MolmoPoint - Image & Video Pointing & Tracking