Understanding Transformers via N-gram Statistics (2407.12034v2)

Published 30 Jun 2024 in cs.CL, cs.AI, and cs.LG

Abstract: Transformer based large-LLMs display extreme proficiency with language yet a precise understanding of how they work remains elusive. One way of demystifying transformer predictions would be to describe how they depend on their context in terms of simple template functions. This paper takes a first step in this direction by considering families of functions (i.e. rules) formed out of simple N-gram based statistics of the training data. By studying how well these rulesets approximate transformer predictions, we obtain a variety of novel discoveries: a simple method to detect overfitting during training without using a holdout set, a quantitative measure of how transformers progress from learning simple to more complex statistical rules over the course of training, a model-variance criterion governing when transformer predictions tend to be described by N-gram rules, and insights into how well transformers can be approximated by N-gram rulesets in the limit where these rulesets become increasingly complex. In this latter direction, we find that for 79% and 68% of LLM next-token distributions on TinyStories and Wikipedia, respectively, their top-1 predictions agree with those provided by our N-gram rulesets.

Citations (2)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Tweets

https://twitter.com/fly51fly/status/1813840395555708969

https://twitter.com/realmofresearch/status/1813996996577083863

https://twitter.com/thodupunoori/status/1825711691591762420

https://twitter.com/betterhn20/status/1923929030241190031

https://twitter.com/betterhn50/status/1923945136331030768

HackerNews

Understanding Transformers via N-gram Statistics (139 points, 16 comments)

Understanding Transformers via N-gram Statistics (1 point, 0 comments)
Understanding Transformers via N-gram Statistics (1 point, 1 comment)

Understanding Transformers via N-gram Statistics (2407.12034v2)

Summary

Related Papers

Tweets

HackerNews

Reddit