Umberto Cappellazzo
hisoka94
AI & ML interests
Multimodal Large Language Models and audio-visual speech processing at @ Imperial College London.
Recent Activity
authored a paper about 11 hours ago
Listening Forward: Next Patch Embedding Prediction Enables Scalable Audio Learners upvoted a paper about 12 hours ago
Listening Forward: Next Patch Embedding Prediction Enables Scalable Audio Learners submitted a paper about 12 hours ago
Listening Forward: Next Patch Embedding Prediction Enables Scalable Audio LearnersOrganizations
None yet