Concept
LLM
Large language models predict next tokens from prompts, tool outputs, and retrieved context windows. Fine-tuning steers tone while guardrails reduce unsafe completions. Cost, latency, and memory dominate hosting decisions at scale.
5 documentation pages cover this concept. Editorial glossary