Berk Bayri

LLM

Large language model: an AI model trained on very large amounts of text to predict and generate language, which can be used to answer, summarise, write, reason about and act on text.

A large language model is a neural network trained on very large collections of text to predict what comes next. From that simple objective, scaled up, it learns to answer questions, summarise documents, translate, write code and follow instructions. Models such as the GPT family, Claude and Gemini are LLMs; the broader term foundation model also covers systems trained on images, audio and other data.

What an LLM is, and is not

An LLM on its own produces text from its training and its prompt. Most of what makes an AI product useful and reliable comes from the system built around it: retrieval (RAG), tools, memory, checks and the harness that coordinates them. This is also where the main risks sit, such as hallucination.

Evaluating one

A model name is not a complete description of what you will run. The same label can be served through different routes and wrapped in different harnesses, so evaluate the configured system, not the label: its runtime address. Systems that use LLMs to plan and act over multiple steps are discussed under agentic AI.