Large Language Model (LLM)
A deep learning model trained on massive text datasets to understand, generate, and manipulate human language. LLMs power modern chatbots, code generators, and text-based AI tools.
A Large Language Model (LLM) is a neural network model trained on vast amounts of text data, capable of understanding and generating human-like text. These models use the transformer architecture and are trained with billions of parameters.
Notable LLMs include OpenAI’s GPT-4 and GPT-4o, Anthropic’s Claude, Google’s Gemini, and Meta’s LLaMA. LLMs can perform a wide range of language tasks including translation, summarization, question answering, code generation, and creative writing.
Key concepts in modern LLMs include context windows (the amount of text the model can consider at once), tokens (the basic unit of text processing), and temperature (controlling randomness in outputs).
Related Terms
Get smarter about AI
The sharpest AI news, curated daily. Delivered free to your inbox.