Search references for LANGUAGE MODEL. Phrases containing LANGUAGE MODEL
See searches and references containing LANGUAGE MODEL!LANGUAGE MODEL
Statistical model of language
A language model is a computational model that predicts sequences in natural language. Language models are useful for a variety of tasks, including speech
Language_model
Type of machine learning model
A large language model (LLM) is an AI model (typically a neural network) trained on a vast amount of text for natural language processing tasks, especially
Large_language_model
Series of language models developed by Google AI
Bidirectional encoder representations from transformers (BERT) is a language model introduced in October 2018 by researchers at Google. It learns to represent
BERT_(language_model)
Large language model developed by Google
Gemini is a family of multimodal large language models (LLMs) developed by Google DeepMind, and the successor to LaMDA and PaLM 2. Comprising Gemini Pro
Gemini_(language_model)
Large language model by Meta AI (2023–2026)
Llama ("Large Language Model Meta AI" serving as a backronym) is a family of large language models (LLMs) released by Meta AI starting in February 2023
Llama_(language_model)
Family of large language models by Google
Gemma is a series of source-available large language models developed by Google DeepMind. It is based on similar technologies as Gemini. The first version
Gemma_(language_model)
Type of artificial intelligence model
Small language models (SLM) or compact language models are artificial intelligence language models designed for human natural language processing including
Small_language_model
Large language model designed for reasoning tasks
Reasoning language models (RLMs) or large reasoning models (LRMs) are large language models that are trained to solve tasks that require several steps
Reasoning_model
A large language model (LLM) is a type of machine learning model designed for natural language processing tasks such as language generation. LLMs are language
List_of_large_language_models
Foundation model allowing control of robot actions
robot learning, a vision–language–action model (VLA) is a class of multimodal foundation models that integrates vision, language and actions. Given an input
Vision–language–action_model
Type of large language model
A generative pre-trained transformer (GPT) is a type of large language model (LLM) that is widely used in generative artificial intelligence chatbots.
Generative pre-trained transformer
Generative_pre-trained_transformer
Language model by DeepMind
Chinchilla is a family of large language models (LLMs) developed by the research team at Google DeepMind, presented in March 2022. It is named "chinchilla"
Chinchilla_(language_model)
Microsoft free and open-source large language model
Phi is a series of large language models developed by Microsoft that are open-weights and can run locally on one's device. The initial version, Phi-1
Phi_(language_model)
Standardized AI performance test
A language model benchmark is a standardized test designed to evaluate the performance of language models on various natural language processing tasks
Language_model_benchmark
Type of artificial intelligence system
A vision–language model (VLM) is a type of artificial intelligence system that can jointly interpret and generate information from both images and text
Vision-language_model
Multilingual open-access large language model
Large Open-science Open-access Multilingual Language Model (BLOOM) is an open-access large language model (LLM) released in 2022. It was created by a
BLOOM_(language_model)
Series of large language models developed by Google AI
is a series of large language models developed by Google AI introduced in 2019. Like the original Transformer model, T5 models are encoder-decoder Transformers
T5_(language_model)
Notation expressing information under a rule set
A modeling language is a notation for expressing data, information or knowledge or systems in a structure that is defined by a consistent set of rules
Modeling_language
Code-generating large language model by OpenAI
OpenAI Codex is a large language model developed by OpenAI for translating natural-language prompts into source code. Announced in 2021, it was a modified
OpenAI_Codex_(language_model)
Software design modeling notation
The Unified Modeling Language (UML) is a general-purpose, object-oriented, visual modeling language that provides a way to visualize the architecture
Unified_Modeling_Language
Artificial intelligence model paradigm
learning model trained on vast datasets so that it can be applied across a wide range of use cases. Generative AI applications like large language models (LLMs)
Foundation_model
Artificial intelligence LLM (released 2026)
large language model created by Thinking Machines company, first released on July 15, 2026. It is published under an Apache 2.0 license. The model allows
Inkling (large language model)
Inkling_(large_language_model)
Large language model developed by Google
PaLM (Pathways Language Model) is a 540 billion-parameter dense decoder-only transformer-based large language model (LLM) developed by Google AI. Researchers
PaLM
Protocol for communicating between LLMs and applications
standardize the way artificial intelligence (AI) systems like large language models (LLMs) integrate and share data with external tools, systems, and data
Model_Context_Protocol
General-purpose modeling language
The systems modeling language (SysML) is a general-purpose modeling language for systems engineering applications. It supports the specification, analysis
Systems_modeling_language
Purely statistical model of language
A word n-gram language model is a statistical model of language which calculates the probability of the next word in a sequence from a fixed size window
Word_n-gram_language_model
Open-source large language model
Jais is an open-source large language model launched in August 2023. Developed as a collaboration between Emirati AI company G42, the Mohamed bin Zayed
Jais_(language_model)
A cache language model is a type of statistical language model. These occur in the natural language processing subfield of computer science and assign
Cache_language_model
Degradation of AI models trained on synthetic data
In artificial intelligence, model collapse, also known as "AI inbreeding", "AI cannibalism", "Habsburg AI", and "model autophagy disorder" or "MAD", is
Model_collapse
Language model of pre-1931 English
Talkie (stylized as talkie) is a small language model developed by Nick Levine, David Duvenaud, and Alec Radford. It was announced in April 2026 and described
Talkie_(language_model)
2020 text-generating language model
Transformer 3 (GPT-3) was a large language model released by OpenAI in 2020 as part of the company's GPT series of models. Like its predecessor, GPT-2, it
GPT-3
Term used in machine learning
learning, the term stochastic parrot is a metaphor that frames large language models as systems that statistically mimic text without real understanding
Stochastic_parrot
Informative representation of an entity
software Economic model, a theoretical construct representing economic processes Language model, a probabilistic model of a natural language, used for speech
Model
Large language model and AI chatbot by Z.ai
General Language Model, is a series of open weight large language models developed by Chinese software company Z.ai. Though the first GLM model was published
GLM_(AI)
Large language model and AI chatbot by Anthropic
Claude is a series of large language models (LLMs) developed by American software company Anthropic. Claude was released as an AI-based chatbot in March
Claude_(AI)
Shading language
to augment the shader assembly language, and went on to become the required shading language for the unified shader model of Direct3D 10 and higher. It
High-Level_Shader_Language
American machine learning company
release an open large language model. In 2022, the workshop concluded with the announcement of BLOOM, a multilingual large language model with 176 billion
Hugging_Face
Algorithm for modelling sequential data
variations have been widely adopted for training large language models (LLMs) on large (language) datasets. Modern transformer designs are commonly grouped
Transformer_(deep_learning)
Type of information retrieval using LLMs
Retrieval-augmented generation (RAG) is a technique that enables large language models (LLMs) to retrieve and incorporate new information from external data
Retrieval-augmented generation
Retrieval-augmented_generation
short for Open Language Model, is a family of large language models developed by the Allen Institute for AI (Ai2). The first models, OLMo 7B, were released
OLMo_(language_model)
Structuring text as input to generative artificial intelligence
process of structuring natural language inputs (known as prompts) to produce specified outputs from a generative AI model. Context engineering is the related
Prompt_engineering
Conformance of AI to intended objectives
distributions. Empirical research in 2024 found that advanced large language models (LLMs) such as OpenAI o1 or Claude 3 sometimes engage in strategic
AI_alignment
Indian artificial intelligence company
the company develops large language models (LLMs) such as Indus and multimodal AI systems with a focus on Indian languages and region-specific use cases
Sarvam_AI
Machine learning model
model is a machine learning model which takes an input natural language prompt and produces an image matching that description. Text-to-image models gradually
Text-to-image_model
Large language model with ternary weights
A 1.58-bit large language model (also known as a ternary LLM) is a type of large language model (LLM) designed to be computationally efficient. It achieves
1.58-bit_large_language_model
Abstract model
programming languages. Data models are often complemented by function models, especially in the context of enterprise models. A data model explicitly determines
Data_model
Artificial intelligence chatbot by Moonshot AI
series of large language models developed by Chinese company Moonshot AI. Its Kimi K3, released July 2026, is the largest open weights model ever, at 2.8
Kimi_(AI)
Concept in artificial intelligence
code-base developed by human engineers that equips an advanced future large language model (LLM) built with strong or expert-level capabilities to program software
Recursive_self-improvement
Software platform for running large language models
2023 for running and managing large language models on local GPU infrastructure and through hosted cloud models. It provides a command-line interface
Ollama
2017 research paper by Google
These convinced the team that the Transformer is a general-purpose language model, and not just good for translation. As of 2026,[update] the paper has
Attention_Is_All_You_Need
Concept in information theory
{\displaystyle q={\tilde {p}}} . In natural language processing, a corpus is a set of sentences or texts, and a language model is a probability distribution over
Perplexity
Family of large language models by Alibaba
large and small language models (LLM and SLM) developed by Alibaba Cloud. Its latest version, Qwen3.8; the 2.4 trillion parameter model was the second
Qwen
Machine learning technique
model, produces a fine-tuned model. It allows for performance that approaches full-model fine-tuning with lower space requirements. A language model with
Fine-tuning_(deep_learning)
Chinese artificial intelligence company
company; their flagship product is the GLM (General Language Model) family of open weights large language models (LLMs). Formerly known as Zhipu AI outside China
Z.ai
Generative AI chatbot by OpenAI
Originally released on November 30, 2022, the product uses large language models—specifically generative pre-trained transformers (GPTs)—to generate
ChatGPT
Large language model family by SpaceXAI
large language models developed by SpaceXAI. It was launched in November 2023 by Elon Musk as an initiative based on the large language model (LLM) of
Grok_(chatbot)
Processing of natural language by a computer
can achieve state-of-the-art results in many natural language tasks, e.g., in language modeling and parsing. This is increasingly important in medicine
Natural_language_processing
Training methods used after LLM pretraining
Post-training of large language models is a term used in recent technical literature for training applied to a large language model (LLM) after its initial
Post-training of large language models
Post-training_of_large_language_models
2026 large language model by OpenAI
GPT-6 is a family of large language models (LLM) developed by American artificial intelligence firm OpenAI. GPT-6 Astra was released to the general public
GPT-6
line. Command Code — CLI and desktop AI agent DeepSeek Coder – large language model designed for code generation and AI-assisted software development. goose
Lists of open-source artificial intelligence software
Lists_of_open-source_artificial_intelligence_software
Artificial intelligence division of Meta Platforms
produces the Muse family of generative AI models, including the Muse AI agent, the Muse Spark large language model (LLM), Muse Image and Muse Video. MSL was
Meta_Superintelligence_Labs
Technology made by American organization
debuted its GPT series of large language models (LLMs) in 2018, and has since released multiple versions of its models. GPT is an acronym for generative
Products and applications of OpenAI
Products_and_applications_of_OpenAI
Google large language models family
LaMDA (Language Model for Dialogue Applications) is a family of conversational large language models developed by Google. Originally developed and introduced
LaMDA
Machine learning technique
proposed mixture of softmaxes for autoregressive language modelling. Specifically, consider a language model that given a previous text c {\displaystyle c}
Mixture_of_experts
Machine learning technique
including natural language processing tasks such as text summarization and conversational agents, computer vision tasks like text-to-image models, and the development
Reinforcement learning from human feedback
Reinforcement_learning_from_human_feedback
Artificial intelligence model developed by Meta
Muse Spark is a large language model (LLM) developed by Meta through its Meta Superintelligence Labs (MSL). It was introduced in April 2026 and launched
Muse_Spark
Standard data modeling language for product data
standard for generic data modeling language for product data. EXPRESS is formalized in the ISO Standard for the Exchange of Product model STEP (ISO 10303), and
EXPRESS (data modeling language)
EXPRESS_(data_modeling_language)
Language model benchmark
Measuring Massive Multitask Language Understanding (MMLU) is a popular benchmark for evaluating the capabilities of large language models. It inspired several
MMLU
Technique using a large language model as an evaluator
LLM-based evaluation or language model-based evaluation) is a technique in natural language processing in which a large language model (LLM) is used to assess
LLM-as-a-Judge
Tel Aviv-based company
launches NLP model to rival GPT-3". Globes. 2021-11-08. Retrieved 2023-02-14. "Watch out, GPT-3, here comes AI21's 'Jurassic' language model". ZDNET. Retrieved
AI21_Labs
Large language model
series of large language models developed by Anthropic. It is the most complicated set of models in the Claude product line. The first model in the series
Claude_Mythos
Gestural communication system
certain "resilient" properties of language whose development can proceed without guidance of a conventional language model. More recent studies of deaf children's
Home_sign
Specification language
model transformation language in systems and software engineering is a language intended specifically for model transformation. The notion of model transformation
Model_transformation_language
Chinese technology company
benchmarked against OpenAI's GPT-4 Turbo. The Xinghuo large language model is a large language model developed by iFlytek in 2024 that solely relies on Huawei's
IFlytek
2023 text-generating language model
Transformer 4 (GPT-4) is a large language model developed by OpenAI and the fourth in its series of GPT foundation models. GPT-4 is preceded by GPT-3.5 and
GPT-4
Erroneous AI-generated content presented as true
computers. Symbolic artificial intelligence models generally do not produce hallucinations, unlike large language models. In computer vision, since the 2000s
Hallucination (artificial intelligence)
Hallucination_(artificial_intelligence)
Theoretical framework
Conceptual model is any model that is the direct output of a conceptualization or generalization process. Conceptual models are often abstractions of things
Conceptual_model
Topics referred to by the same term
large language model that is widely used in generative artificial intelligence chatbots GPT (OpenAI model), a series of large language models developed
GPT
2019 text-generating language model
Transformer 2 (GPT-2) is a large language model (LLM) by OpenAI and the second in their foundational series of GPT models. GPT-2 was pre-trained on a dataset
GPT-2
ASCII-based file format for describing graphs
Graph Modeling Language (GML) is a hierarchical ASCII-based file format for describing graphs. It has been also named Graph Meta Language. A simple graph
Graph_Modelling_Language
Third astrological sign of the zodiac
zodiac Circle of stars Cusp (astrology) Elements of the zodiac Gemini (language model) Digital twin Air (classical element) Astronomical Applications Department
Gemini_(astrology)
2026 large language model by OpenAI
Transformer 5.6) is a set of large language models (LLMs) developed by OpenAI and released on July 9, 2026. It is a family of models that comes in three distinct
GPT-5.6
AI chatbot developed by Yandex
assistant Alice to interact with the new language model. On June 15, 2023, Yandex added the YandexGPT language model to the image generation application Shedevrum
Alice_AI_(AI_model_family)
Temporal limit of a model's knowledge
time beyond which a large language model has not been trained on new data. Since large language models are pretrained, any model's knowledge is fixed at what
Knowledge_cutoff
Chinese Internet technology company
proprietary large language model while seed-oss is an open-source large language model. Both are available in the US. Seedream is a text-to-image model. It has
ByteDance
Software platform for running artificial intelligence models locally
LM Studio is large language model inference software for running and using large language models locally on personal computers. It provides tools for
LM_Studio
Language model benchmark
Humanity's Last Exam (HLE) is a language model benchmark consisting of 2,500 questions across a broad range of subjects. It was created jointly by the
Humanity's_Last_Exam
Intelligence in machines
research could model. Humans solve most of their problems using fast, intuitive judgments. Reasoning models, a type of large language model (LLM) trained
Artificial_intelligence
Artificial intelligence tool
first announced by GitHub on 29 June 2021. Users can choose the large language model used for generation. On June 29, 2021, GitHub announced GitHub Copilot
GitHub_Copilot
Statistical law in machine learning
resources and time required for model training. With the "pretrain, then finetune" method used for most large language models, there are two kinds of training
Neural_scaling_law
Graphical notation for user interactions with a software system
The Interaction Flow Modeling Language (IFML) is a standardized modeling language in the field of software engineering. IFML includes a set of graphic
Interaction Flow Modeling Language
Interaction_Flow_Modeling_Language
Topics referred to by the same term
zodiac Gemini (astrology), an astrological sign Gemini (language model), series of large language models developed by Google Google Gemini, chatbot Gemini twins
Gemini
Open-source software for large language model inference
software framework for inference and serving of large language models and related multimodal models. Originally developed at the University of California
VLLM
AI research laboratory
development of two large language model families: the proprietary Gemini and open-weight Gemma, as well as other generative AI models, such as Imagen (text-to-image)
Google_DeepMind
Text-to-video model by Google
VideoPoet is a large language model developed by Google Research in 2023 for video making. It can be asked to animate still images. The model accepts text, images
VideoPoet
Open-source framework for large language model inference
Structured Generation Language) is an open-source framework for programming and serving large language models and multimodal models. It was introduced by
SGLang
Adjacent characters (tokens) merge-based compression algorithm
table. A slightly modified version of the algorithm is used in large language model tokenizers. The original version of the algorithm focused on compression
Byte-pair_encoding
Artificial intelligence division of Meta Platforms
under Meta Superintelligence Labs, which brought together FAIR, large language model development teams, and other AI research and product groups. Meta AI
Meta_AI
Machine learning methods using multiple input modalities
transfer learning. The LLaVA was a vision-language model composed of a language model (Vicuna-13B) and a vision model (ViT-L/14), connected by a linear layer
Multimodal_learning
Swiss free and open-source large language model
Apertus is a public large language model, developed by the Swiss AI Initiative (a collaboration between EPFL, ETH Zurich, and the Swiss National Supercomputing
Apertus_(LLM)
travel, tourism, insurance
LANGUAGE MODEL
LANGUAGE MODEL
LANGUAGE MODEL
LANGUAGE MODEL
LANGUAGE MODEL
LANGUAGE MODEL
LANGUAGE MODEL
LANGUAGE MODEL
LANGUAGE MODEL
travel, tourism, insurance