Lasha Koroshinadze
lashahub
AI & ML interests
Large Audio-Language Models
Recent Activity
authored a paper about 10 hours ago
Audio-Visual Flamingo: Open Audio-Visual Intelligence for Long and Complex Videos authored a paper 3 months ago
Audio Flamingo Next: Next-Generation Open Audio-Language Models for Speech, Sound, and Music authored a paper 4 months ago
MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos