Skip to content

Knowledge catalog

Generated reference

Read-only pages rendered from the Quake AI knowledge graph.

Concept

LLM

Large language models predict next tokens from prompts, tool outputs, and retrieved context windows. Fine-tuning steers tone while guardrails reduce unsafe completions. Cost, latency, and memory dominate hosting decisions at scale.