Transformer (AI architecture)

From ALT-TEXT
Jump to navigation Jump to search

Transformer (AI architecture)

The neural network design, introduced by Google researchers in 2017, that underlies almost all modern large language models. Transformers use a mechanism called "attention" to weigh the relevance of different words in a text relative to one another, allowing the model to process long passages of language efficiently. (See also: Large language model, Neural network)