Benchmarking Zero-Shot Recognition with Vision-Language Models: Challenges on Granularity and SpecificityPublished in CVPR Workshop on Multimodal Foundation Models, 2024Share on Twitter Facebook LinkedIn Previous Next