Submitted by
Hanoona Rasheed
AI & ML interests
Natural Language Processing, Machine Learning, and Computer Vision
Recent Activity
View all activity
Papers
Locate Anything in Videos: Rethinking Efficient Generative Spatio-Temporal Video Grounding
Training-Free Speech-Centric Omni Understanding with Frozen VLMs