Word lengths are optimized for efficient communication
Massachusetts Institute of Technology
Indexed incrossrefpubmed
Abstract
We demonstrate a substantial improvement on one of the most celebrated empirical laws in the study of language, Zipf's 75-y-old theory that word length is primarily determined by frequency of use. In accord with rational theories of communication, we show across 10 languages that average information content is a much better predictor of word length than frequency. This indicates that human lexicons are efficiently structured for communication by taking into account interword statistical dependencies. Lexical systems result from an optimization of communicative pressures, coding meanings efficiently given the complex statistics of natural language use.
Citation impact
706
total citations
- FWCI
- 32.53
- Percentile
- 100%
- References
- 46
Citations per year
Authors
3Topics & keywords
Topics
Keywords
- Zipf's law
- Word (group theory)
- Word length
- Computer science
- Coding (social sciences)
- Natural language processing
- Word lists by frequency
- Human language
UN Sustainable Development Goals
- Quality Education
No related works found for this paper.