Contents
Map

20 · Capstones

Appendix - Summary & Key Terms

View as:

Appendix - Capstones

What We Learned

  • Integrate pretraining, post-training and evaluation into a reproducible model project.
  • Serve under realistic load and defend optimizations using quality, latency and goodput measurements.
  • Ship a tool-using agent with permission boundaries, approval, persistence and telemetry.
  • Report baselines, uncertainty, costs, limitations and failure cases alongside the code.
  • Explain architecture decisions and the business outcome with reviewable evidence.

Key Acronyms, Concepts & Jargon

TermShort meaning
SFT / DPO / GRPOSupervised Fine-Tuning / Direct Preference Optimization / Group Relative Policy Optimization.
MFUModel FLOPs Utilization: useful model compute as a fraction of accelerator peak compute.
Model cardDocuments intended use, training, evaluation and limitations of a model.
TTFT / TPOTTime To First Token / Time Per Output Token: serving responsiveness metrics.
SLO / goodputService Level Objective / work completed within defined service targets.
Open-loop load testSchedules arrivals independently of response completion to expose queueing.
MCP / HITLModel Context Protocol / Human In The Loop: tool integration / human review boundary.
Durability / idempotencyResume after interruption / retries preserve one intended effect.
Threat modelMaps attackers, assets, entry points and controls.
CI / pass^kConfidence Interval / success across all k repeated trials.
ADR / ROIArchitecture Decision Record / Return On Investment: design rationale / business value measure.
ReproducibilityEnough versioned code, data, configuration and environment detail to repeat the work.

Back to section overview

⚡AI-assisted content - always verify, always explore multiple perspectives·