paper
referenced-only
2024
paper:doi-10-1016-j-patter-2025-101176

Attention heads of large language models

ByZifan Zheng·Yezhaohui Wang·Yuxin Huang·Shichao Song·Bo Tang·Feiyu Xiong+1 more
Original abstract (expand)

Summary Large language models (LLMs) have demonstrated performance approaching human levels in tasks such as long-text comprehension and mathematical reasoning, but they remain black-box systems. Understanding the reasoning bottlenecks of LLMs remains a critical challenge, as these limitations are deeply tied to their internal architecture. Attention heads play a pivotal role in reasoning and are thought to share similarities with human brain functions. In this review, we explore the roles and mechanisms of attention heads to help demystify the internal reasoning processes of LLMs. We first introduce a four-stage framework inspired by the human thought process. Using this framework, we review existing research to identify and categorize the functions of specific attention heads. Additionally, we analyze the experimental methodologies used to discover these special heads and further summarize relevant evaluation methods and benchmarks. Finally, we discuss the limitations of current research and propose several potential future directions.

Similar preprints — Semantic Scholar

Cited by (1)