Model and data for ReflectiVA: Augmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering [CVPR 2025]
Federico Cocchi
fede97
AI & ML interests
Multimodal LLM - Computer Vision
Recent Activity
updated
a model
about 1 month ago
aimagelab/LLaVA_MORE-gemma_2_9b-dinov2-finetuning
published
a model
about 1 month ago
aimagelab/LLaVA_MORE-gemma_2_9b-dinov2-finetuning
updated
a model
about 1 month ago
aimagelab/LLaVA_MORE-gemma_2_9b-dinov2-pretraining