Skip to content

NewsMore

Research

Multimodal Large Language Model

Multimodal Large Language Model

Multimodal large language models (MLLMs) are driving AI toward cross-modal understanding, generation, and interaction. Our research focuses on key challenges in modality fusion, cross-domain generalization, and interpretable reasoning. We develop integrated approaches to representation, architecture, and data for more capable and human-centered AI systems.