Core Research
Multimodal Emotion Understanding
Modeling emotion from speech, vision, and text with depth-aware representations.
Hi, I'm
Ph.D. candidate in Computer Science, focusing on multimodal learning, affective computing, and robust speech understanding.
NowCurrently at PolyU · Hong Kong from Sep 2026
I am a Ph.D. candidate in Computer Science at Sichuan University. My research focuses on multimodal learning, affective computing, and robust speech understanding, with particular interests in multimodal sentiment analysis, cross-modal contrastive optimization, noise-resilient speech recognition, and emotion-aware agent workflows. I care about research that remains interpretable and generalizes under real-world conditions.
Core Research
Modeling emotion from speech, vision, and text with depth-aware representations.
Systems
Task-adaptive routing and efficient expert collaboration for better generalization.
Impact
Bridging research and deployment through reliable workflows and automation.
Hierarchical emotion modeling with adaptive multi-level mixture-of-experts.
End-to-end pipeline for multimodal emotion analysis and conversational AI.
Hierarchical cross-modal denoising for robust audio-visual speech representation under noisy real-world conditions.
Toolchain for transforming research prototypes into reproducible demos.