Haolin HE
Harland
AI & ML interests
Large Audio Language Models
Recent Activity
upvoted a paper 2 days ago
VisionWeave: Weaving Elastic Visual Representations as a Native Capability of MLLMs upvoted a paper 13 days ago
OmniVChat: Synthesizing, Benchmarking, and Training for Native Audio-Visual Dialogue upvoted a paper 17 days ago
HappyWorld-Bench