I am currently with the Mamoda team at ByteDance, where I work on unified omni-modal models for understanding and generation. We are building a real-time long-video model that unifies video understanding and generation, driven by new model architectures and scaling-law-guided training. Our goal is to push the frontier of both generation quality and multimodal understanding.
Last updated: Aug. 12, 2026 · Feel free to reach out via email
ICML
arXiv
ICLR
arXiv
arXiv
CVPR
AAAI
CVPR
CVPR
ICLR
arXiv
CVPR
IROS
ROBIO
ICIA
Powered by Jekyll and Minimal Light theme.