Transformer (AI architecture)
Jump to navigation
Jump to search
Transformer (AI architecture)
The neural network design, introduced by Google researchers in 2017, that underlies almost all modern large language models. Transformers use a mechanism called "attention" to weigh the relevance of different words in a text relative to one another, allowing the model to process long passages of language efficiently. (See also: Large language model, Neural network)