search
home
research
people
contact
donate
system
All topics
MM
Multimodal Learning
Joint learning across vision, language, and other modalities.
Papers
Continual Visual and Verbal Learning Through a Child's Egocentric Input
2026-06-03
Xavier (Xiaoyang) Jiang, Yanlai Yang, Kenneth A. Norman, Brenden Lake, and Mengye Ren
MA-EgoQA: Question Answering over Egocentric Videos from Multiple Embodied Agents
2026-03-10
Kangsan Kim, Yanlai Yang, Suji Kim, Woongyeong Yeo, Youngwan Lee, Mengye Ren, and Sung Ju Hwang
StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding
2025-08-21
Yanlai Yang, Zhuokai Zhao, Satya Narayan Shukla, Aashu Singh, Shlok Kumar Mishra, Lizhu Zhang, and Mengye Ren
LifelongMemory: Leveraging LLMs for Answering Queries in Long-form Egocentric Videos
2023-12-07
Ying Wang, Yanlai Yang, and Mengye Ren
Prev
Next
Learning Paradigms
Continual Learning
Self-Supervised Learning
Test-time Learning
In-Context Learning
Creative Exploration
Meta-Learning
Multi-Agent
Local Learning
Reinforcement Learning
Models & Representations
Concept Learning
Hierarchical Abstraction
World Models
Data & Applications
LLM Reasoning
Egocentric Video
Embodied AI
Multimodal Learning
Forecasting
Perspectives
Human-like Learning
Philosophy of AI
AI Safety