Glossary
LLM Inference
Inference is the process of running a trained language model to produce an output from a prompt and its available context. During inference, model parameters are generally not retrained; the response is generated from existing parameters plus any instructions, conversation history and retrieved evidence.
In plain terms
Inference is the moment a trained model creates an answer.
Why it matters
It separates what happens during a live query from model training, preventing false claims that a page edit immediately teaches the underlying model.
How to apply it
- Describe visibility changes as retrieval or answer changes unless training is documented.
- Record the model and settings used during tests.
- Use citations to identify live evidence.
Example
A newly published page influences a cited answer through retrieval during inference without changing the model’s trained parameters.
Sources
Related reading
Related terms
Back to the full glossary (75 terms).
Ranking is no longer enough
You need to be cited, mentioned, and recommended.
Being cited, mentioned, and recommended are three different outcomes, and most brands only ever achieve the first one. Ranking is no longer enough because AI engines answer buyers directly and name only a short list of vendors as the recommendation — everyone else is cited in passing, if at all. Get a free AI Visibility Report to see exactly where your brand appears today across ChatGPT, Google AI Overviews, Gemini, Perplexity, and Copilot, where competitors are winning the recommendation instead, and what's keeping you from moving up the shortlist.