Skip to content

Large Language Model

A neural network trained on massive text corpora to predict the next token, enabling it to generate and understand language.

Last updated

A large language model (LLM) is a neural network, usually based on the transformer architecture, trained to predict the next token in a sequence. Because that simple objective is practised on billions of words, the model picks up grammar, facts, reasoning patterns and even coding ability. LLMs power chat assistants, summarisation, translation and code generation. Their quality depends heavily on model size, training data and the fine-tuning that follows.

Related terms

Read more about Large Language Model

Keep exploring

Browse the full AI glossary or compare AI tools that use this technology.

Report an issue with this page