Yichi Zhang
Hello! My name is Yichi Zhang, and I am doing multimodal research at ByteDance Seed. My research focuses on real-time multimodal agentic systems that integrate perception, reasoning, and interaction across streaming vision and audio. I aim to build proactive, full-duplex agents that can continuously understand and act in both digital and physical worlds, collaborating with people naturally and at the right moment.
During my Ph.D., I also worked on embodied AI agents that follow natural-language instructions and collaborate with people in situated, interactive environments. I received my Ph.D. in Computer Science and Engineering from the University of Michigan, where I was advised by Professor Joyce Chai as a member of the SLED lab. In 2023, I led Team SEAGULL to win first place in the inaugural Amazon Alexa Prize SimBot Challenge.
Before joining UMich, I obtained my Master’s in Information and Communication Engineering at Tsinghua University in 2020, advised by Professor Zhijian Ou. In 2019, I worked with Professor Zhou Yu as a visiting scholar on task-oriented dialog systems. I received my Bachelor’s in Electronic Information Science and Technology from Tsinghua University in 2017.
News
| Aug 05, 2026 | Excited to release SeedRealtime, a native audio-visual full-duplex LLM. As a core contributor, I worked on visual-based proactive interaction and duplex RL training, enabling the model to continuously watch, listen, and respond at the right moment. |
|---|---|
| Jul 01, 2026 | We released the Seed2.0 Model Card, presenting a frontier model for real-world complexity. I contributed to its streaming-video and omni-modal understanding capabilities. |
| Apr 15, 2026 | We released the Seedance 2.0 paper, presenting a unified multimodal audio-video generation model. I contributed to its omni-modal data pipeline. |
| Mar 21, 2026 | We released the Seed1.8 Model Card, introducing a generalized agentic model for real-world scenarios. I contributed to its streaming-video understanding and proactive visual interaction capabilities. |
| May 23, 2025 | Excited to share that I’ve joined the Multimodal Interaction & World Model team at ByteDance! Looking forward to tackling new challenges ahead. I’m open to collaborations and mentoring interns — feel free to reach out! |