atlas

Research map

Vision-language and multimodal: research map

Multimodal LLMs, CLIP-style models, VQA and embodied tasks.

7,482 accepted papers at ICML, NeurIPS, ICLR, CVPR and AAAI (2016–2026), in 6 topics. Within all five venues, its share grew from 6.0% in 2023–24 to 10.5% in 2025–26 (1,587 → 4,783 papers at ICML, NeurIPS, CVPR and AAAI, the venues with data for all four years).

2016: 62162017: 88172018: 131182019: 149192020: 157202021: 230212022: 295222023: 520232024: 1067242025: 1731252026: 305226

Explore Vision-language and multimodal in the interactive map

Topics

Most cited papers in Vision-language and multimodal