Model and data for ReflectiVA: Augmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering [CVPR 2025]
Federico Cocchi
fede97
AI & ML interests
Multimodal LLM - Computer Vision
Recent Activity
updated
a model
1 day ago
aimagelab/LLaVA_MORE-gemma_2_2b-finetuning
published
a model
1 day ago
aimagelab/LLaVA_MORE-gemma_2_2b-finetuning
updated
a model
6 months ago
aimagelab/LLaVA_MORE-gemma_2_9b-dinov2-finetuning