FastVLM Collection Efficient Vision Encoding for Vision Language Models • 9 items • Updated 6 days ago • 90
MobileCLIP2 Collection MobileCLIP2: Mobile-friendly image-text models with SOTA zero-shot capabilities trained on DFNDR-2B • 31 items • Updated 6 days ago • 47