arxiv:2509.21060
Haolin HE
Harland
AI & ML interests
Large Audio Language Models
Recent Activity
upvoted a paper about 4 hours ago
VisionWeave: Weaving Elastic Visual Representations as a Native Capability of MLLMs upvoted a paper 11 days ago
OmniVChat: Synthesizing, Benchmarking, and Training for Native Audio-Visual Dialogue upvoted a paper 15 days ago
HappyWorld-Bench