Pidoku
Glossary

Glossary

Short definitions of the terms this course uses. The lesson in brackets is where each is explained.

TermMeaning
A2AAgent2Agent protocol: how one agent delegates a task to another across a boundary. (Tools and Protocols)
ActivityIn durable execution, a recorded step with side effects, such as a model or tool call. (Orchestration and Durable Execution)
AgentA model choosing its own next action in a loop, run by a harness. (Anatomy of an Agent)
Agent CardAn A2A document describing an agent’s skills and endpoint; signed in v1.0. (Tools and Protocols)
Agent SandboxThe Kubernetes API for stateful, isolated, idle-heavy agent workspaces. (Sandboxes and Tool Execution)
AI gatewayThe single entry point for model traffic: identity, token limits, routing, fallback. (Model Access and Gateways)
Blast radiusThe worst outcome if a component is fully compromised. (Trust Boundaries)
CheckpointSaved agent state from which a task can resume. (Orchestration and Durable Execution)
ChunkA retrievable piece of a document, a few hundred tokens long. (Context and Retrieval)
Circuit breakerSkips a failing route until probes succeed again. (Reliability Patterns)
CompactionSummarising older context so a long task fits the window. (Context Engineering and Memory)
Context engineeringDeciding what enters the context window on each call. (Context and Retrieval)
Context windowEverything the model sees on one call; a fixed budget. (What a Model Changes)
Control planeComponents that decide what should run; not on the request path. (The Layers)
Data planeComponents that carry requests. (The Layers)
Deterministic controlA control enforced by code or configuration that the model cannot talk its way past. (Trust Boundaries)
DRADynamic Resource Allocation: Kubernetes API for requesting devices by attribute. (The Cluster)
Durable executionRecording each step so a task resumes exactly where it stopped. (Orchestration and Durable Execution)
Egress controlRestricting where a component may send network traffic. (Sandboxes and Tool Execution)
EmbeddingA vector representing the meaning of a piece of text. (Context and Retrieval)
Endpoint pickerThe component that chooses a model-server replica per request. (The Serving Layer)
EvaluationA dataset plus graders that measure quality as a rate. (Evaluation and Quality)
GoodputThroughput that still meets the latency targets. (Compute and Capacity)
HarnessThe code around the model in an agent: context, tools, policy, budgets, state. (Anatomy of an Agent)
Vector and keyword search combined. (Context and Retrieval)
Idempotency keyAn identifier that makes a repeated write take effect once. (Orchestration and Durable Execution)
InferencePoolThe Gateway API resource grouping pods that serve one model. (The Serving Layer)
Judge modelA model used to grade another model’s output. (Evaluation and Quality)
Lethal trifectaPrivate data, untrusted content and an outbound channel in one agent context. (Trust Boundaries)
MCPModel Context Protocol: how an agent reaches tools and data. (Tools and Protocols)
OccupancyThe share of capacity in use; the main driver of self-hosted unit cost. (The Layers)
Policy layerCode that approves or denies each proposed tool call. (Anatomy of an Agent)
Prefill / decodeReading the prompt, then generating tokens one at a time. (The Serving Layer)
Prefix cacheReuse of the processed form of a repeated prompt prefix. (State, Memory and Caching)
Prompt injectionText in the context that redirects the model. (What a Model Changes)
RAGRetrieval-augmented generation: answering from fetched context. (Context and Retrieval)
RerankerA model that scores candidate chunks against the question. (Context and Retrieval)
Reciprocal rank fusionA way to merge rankings without comparable scores. (Context and Retrieval)
Rule of twoAn agent holds at most two legs of the trifecta without human approval. (Trust Boundaries)
SandboxAn isolated environment for running untrusted code. (Sandboxes and Tool Execution)
Semantic cacheReturns a stored answer for a similar question; risky across users. (State, Memory and Caching)
SkillA folder of instructions and scripts an agent loads when relevant. (Anatomy of an Agent)
Sub-agentAn agent given a narrow task and a fresh context by another agent. (Context Engineering and Memory)
TokenThe unit a model reads and writes; the unit of cost, latency and limits. (What a Model Changes)
Tool callingThe model requests a function call; the harness executes it. (Tools and Protocols)
TPOTTime per output token. (What a Model Changes)
TTFTTime to first token. (What a Model Changes)
WorkflowModel calls arranged in a control flow written in advance. (Anatomy of an Agent)

↑↓ navigate↵ openesc close

drag to pan · scroll to zoom