Multimedia Semantic Analytics Lab
The Multimedia Semantic Analytics Lab (MSALab), led by Prof. Yunhai Tong, is a research group at Peking University dedicated to advancing the frontier of Artificial Intelligence through high-quality research, education, and real-world impact.
MSALab is a center of excellence focusing on semantic understanding across multiple modalities.
Our mission is to explore state-of-the-art AI methods that enable machines to better understand and generate meaning from vision, language, and video.
We conduct fundamental and applied research to push forward the capabilities of modern AI systems in perceiving and reasoning about the world.
Our research interests include, but are not limited to:
- Multi-modal Learning
- Visual Perception
- Video Content Analysis
- Language and Semantic Understanding
We aim to bridge different modalities and uncover the semantic structures behind diverse data sources.
We are currently focusing on the following key directions:
- Multi-modal Large Language Models (MLLMs)
- Text-to-Image and Text-to-Video Generation & Editing
- Large Language Models (LLMs)
MSALab is actively recruiting self-motivated research interns and Ph.D. students.
- Have strong coding skills
- Possess a solid mathematical background
- Show strong interest in:
- Multi-modal learning
- Large Language Models
- Text-to-Image / Text-to-Video generation and editing
📍 Location: Peking University
If you are passionate about cutting-edge AI research, we warmly welcome you to join our team.
Recent research works can be found here
(Link to be updated)
For collaboration or recruitment inquiries, please contact us through official channels.
Multimedia Semantic Analytics Lab · Peking University