Guangyu Sun
imguangyu
AI & ML interests
Multi-modal learning, foundation model, federated learning
Recent Activity
authored a paper about 17 hours ago
FLAT: Resampling Image and Text into 1D Flexible-Length Aligned Transmodal Tokens for Retrieval and Generation authored a paper about 17 hours ago
Diagnosing Visual Reasoning: Challenges, Insights, and a Path Forward authored a paper about 17 hours ago
From Frames to Clips: Efficient Key Clip Selection for Long-Form Video
UnderstandingOrganizations
None yet