Skip to content
large language model

File:The_number_of_publications_about_Large_Language_Models_by_year.png · Wikimedia Commons · See Wikimedia Commons

EntityQ115305900· pop 66· linked from 1,009 articles

large language model

Sign in to save

Also known as LLM, LLMs, large language models, word soup machine, word soup model, word salad machine, word salad model

language model built with very large amounts of texts

AI overview

A large language model is a computer system trained on vast amounts of text data to understand and generate human language. It matters because it can perform a wide range of language tasks—like answering questions, writing, and translation—which makes it useful for many practical applications.

AI-generated from the Wikipedia summary — may contain errors.

Described at

Link to a page describing this subject · not written by Vinony

Wikidata facts

Subclass of
language model
Show 10 more facts
Commons category
Large language models
short name
LLM
on focus list of Wikimedia project
WikiProject Artificial Intelligence
topic's main category
Category:Large language models
Sources (3)

via Wikidata · CC0

~40 min read

Encyclopedic overview

A large language model (LLM) is a neural network trained on a vast amount of text for natural language processing tasks, especially language generation. LLMs can typically generate, summarize, translate and analyze text in many contexts, and are a foundational technology behind modern chatbots. Biased or inaccurate training data can make an LLM's output less reliable.

As of 2026, the most capable LLMs are based on transformer architectures, which, according to the 2017 paper "Attention Is All You Need", can be more efficient and parallelizable than earlier statistical and recurrent neural network models.

Excerpted from Wikipedia’s “large language model” article, available under the CC BY-SA 4.0 licence.

Gallery (7)