Glossary · Models and generation
Developer message
Application-level instructions placed below system policy and above ordinary user content in some model APIs. Its exact priority and availability depend on the provider.
Terms this definition uses
Models and generation
How a model produces text, what its training and decoding controls change, and what none of them can prove about a source.
Autoregressive model · Transformer · Attention · Self-attention · Parameter · Inference · Pretraining · Fine-tuning · Supervised fine-tuning · Instruction tuning · RLHF · Preference optimization · Knowledge distillation · Quantization · Mixture of experts · Multimodal model · Vision-language model · Small language model · Tokenizer · Temperature · Top-p sampling · Deterministic decoding · System prompt · Tool call · Function calling · Structured output · Constrained decoding · Chain of thought · Reasoning trace · Context caching · Context compaction · Model routing · Model drift · Model version
← System prompt · Tool call →
See it in the full glossary · 579 terms across 19 areas. Scan a site to see which of these apply to it.