Model and data for ReflectiVA: Augmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering [CVPR 2025]
Federico Cocchi
fede97
AI & ML interests
Multimodal LLM - Computer Vision
Recent Activity
updated
a dataset
about 2 months ago
aimagelab/RAID
updated
a collection
about 2 months ago
RAID
updated
a collection
about 2 months ago
RAID