Transformer (AI architecture)

From ALT-TEXT
Revision as of 10:36, 7 September 2026 by imported>ALT-TEXT (Import: AI terminology and people glossary)
(diff) ← Older revision | Latest revision (diff) | Newer revision → (diff)
Jump to navigation Jump to search

Transformer (AI architecture)

The neural network design, introduced by Google researchers in 2017, that underlies almost all modern large language models. Transformers use a mechanism called "attention" to weigh the relevance of different words in a text relative to one another, allowing the model to process long passages of language efficiently. (See also: Large language model, Neural network)