atlas

Research map / Architectures and efficiency

Attention and transformers: research map

2,296 accepted papers on Attention and transformers in Architectures and efficiency, from ICML, NeurIPS, ICLR, CVPR and AAAI (2016–2026), grouped into 6 clusters and 25 approaches. The busiest year so far is 2026.

Within Architectures and efficiency, its share grew from 29.8% in 2023–24 to 32.9% in 2025–26 (583 → 1,166 papers at ICML, NeurIPS, CVPR and AAAI, the venues with data for all four years).

2016: 8162017: 24172018: 33182019: 46192020: 100202021: 137212022: 199222023: 221232024: 362242025: 504252026: 66226

Explore Attention and transformers in the interactive map

Working on something in this topic? Describe your idea in scime atlas to see which approach it falls under, the closest papers by meaning and how crowded the spot has become.

Approaches and key papers

kv · cache · long context · 531 papers

Approaches in this cluster:

Most cited and most cited since 2024:

vision · vit · transformers · 384 papers

Approaches in this cluster:

Most cited and most cited since 2024:

self attention · attention mechanism · softmax · 369 papers

Approaches in this cluster:

Most cited and most cited since 2024:

sequence · long · modeling · 364 papers

Approaches in this cluster:

Most cited and most cited since 2024:

transformers · language · icl · 349 papers

Approaches in this cluster:

Most cited and most cited since 2024:

global · local · spatial · 299 papers

Approaches in this cluster:

Most cited and most cited since 2024:

Related topics in Architectures and efficiency