Skip to main content
Living editorial map

What we have explored—and what still needs a story

This page is generated from the project’s canonical curriculum state. Research depth and story coverage are tracked separately: a subject can be well mapped before a story is commissioned.

Evidence reviewed through State version 1.11.0

4schools
14definitive subjects
3subject deep dive complete
49deep-dived content areas
14subjects touched by a story
17deep-dive areas with a story
13reader-visible stories
11current workflow completed

How to read the state​

A research badge reports how far the territory has been explored. Inside each subject, stories are grouped by Beginner, Advanced, and Professional reasoning level; only populated levels appear.

Landscape mappedDeep dive completeNo story yetNext in pipelineStory stagedStory available
School 1

How software executes and changes​

Where does work run, what do boundaries promise, and how can behavior evolve safely?

School 2

Where truth lives and moves​

Which state is authoritative, how is it found, and what survives replay or partial failure?

School 3

How production systems are operated and trusted​

How do platforms remain observable, economical, recoverable, and authorized?

07
Landscape mappedStory available

Architecture, platforms, and cloud control planes

Which boundaries let teams change, operate, and recover independently?

Beginner1 story
09
Landscape mappedStory available

Performance, capacity, and technology economics

Which resource limits useful work, and what does a successful outcome really cost?

Advanced1 story
School 4

How intelligent systems are built and governed​

How do probabilistic components become dependable, bounded products?

12
Deep dive completeStory available

Generative model behavior, adaptation, and inference

Which model behaviors can the application constrain, measure, or afford?

16 defined content areas

Beginner2 stories
Advanced1 story
Professional1 story
13
Deep dive completeStory available

Agentic and multimodal product systems

How do models, tools, state, protocols, media, and people complete bounded work?

20 defined content areas

Advanced1 story
Professional3 stories
Deep-dive inventory

Subject 11: ML systems and decision quality​

All content areas below are research-ready. Story status shows which parts of the subject have been turned into public teaching material.

Evidence and model foundations​

11.1No story yet

Decision problem framing, baselines, and loss

Define the population, unit of decision, desired outcome, observable proxy, action set, intervention owner, time horizon, baseline, and cost of each error before selecting a model.

No story has been commissioned for this area.

11.2Story available

Label, outcome, and observation contracts

Specify how an outcome becomes an authoritative, versioned label, including observation delay, adjudication, correction, missingness, censoring, and the effect of intervention.

Beginner1 story
11.3No story yet

Feature lineage, point-in-time correctness, and leakage

Build reproducible features and semantic representations from only the evidence available at decision time, with the same meaning across training and serving.

No story has been commissioned for this area.

11.4No story yet

Model evaluation, calibration, uncertainty, and abstention

Evaluate whether scores, rankings, classes, or generated signals support the intended decisions across relevant slices and uncertainty regions.

No story has been commissioned for this area.

Operational decision systems​

11.5No story yet

Decision policies, thresholds, and intervention boundaries

Convert model evidence into a deterministic, versioned policy for action, non-action, evidence gathering, abstention, review, escalation, and appeal.

No story has been commissioned for this area.

11.6Story available

Operational thresholds, reviewer queues, and feedback bias

Co-design threshold, prioritization, deferral, and reviewer capacity so the deployed system remains timely, learnable, and safe under the actual score and uncertainty distribution.

Advanced1 story
11.7No story yet

Delayed and selective ground truth, intervention feedback, and counterfactual evidence

Evaluate and learn when actions determine which outcomes occur, which outcomes become visible, and which counterfactuals remain unknowable.

No story has been commissioned for this area.

11.8No story yet

Segment behavior, fairness, contestability, and appeal

Define and operate decision quality for materially different populations, contexts, and harms, with a usable path to challenge or correct a decision.

No story has been commissioned for this area.

Production learning and change​

11.9No story yet

Production monitoring, drift, quality objectives, and diagnosis

Detect loss of decision validity and identify whether the cause is population, data, feature, label, model, calibration, policy, threshold, reviewer, intervention, or outcome change.

No story has been commissioned for this area.

11.10No story yet

Experiments, shadowing, canaries, and safe intervention rollout

Compare decision systems without exposing users or operations to uncontrolled interventions or evaluating a shadow system on outcomes it did not cause.

No story has been commissioned for this area.

11.11No story yet

Retraining, reproducibility, rollback, and model-system inventory

Reproduce and change the complete decision system across data, features, labels, preprocessing, model, calibration, policy, threshold, reviewer workflow, and environment versions.

No story has been commissioned for this area.

11.12No story yet

Generative-model signals in decision systems

Use LLM or multimodal outputs as versioned, evaluated decision evidence without treating structured generation, explanation, self-confidence, or a model judge as authoritative truth.

No story has been commissioned for this area.

11.13No story yet

Professional capstone — mixed-version decision loops

Defend a production decision system whose population, labels, features, models, policies, reviewers, interventions, and governance controls change on different clocks.

No story has been commissioned for this area.

Deep-dive inventory

Subject 12: Generative model behavior, adaptation, and inference​

All content areas below are research-ready. Story status shows which parts of the subject have been turned into public teaching material.

Behavior and controlled generation​

12.1No story yet

Capability profiles, request contracts, and authority boundaries

Define supported tasks, inputs, segments, outputs, failure states, and the deterministic authority that may accept a generation.

No story has been commissioned for this area.

12.2No story yet

Tokenization, chat templates, context allocation, and truncation

Reproduce what the model received and test representation, position, language, stop, cost, and evidence-loss effects.

No story has been commissioned for this area.

12.3No story yet

Decoding, stopping, stochasticity, and repeatability

Choose generation controls while keeping business invariants outside sampling assumptions and seeds.

No story has been commissioned for this area.

12.4Story available

Structured generation, semantic validation, bounded repair, and abstention

Use constrained decoding for structure while enforcing semantic and domain rules before authority; bound repair and expose unsupported output.

Beginner1 story
12.5Story available

Effective context, reasoning effort, and test-time compute policy

Allocate context and reasoning by task value, quality evidence, deadline, cache, and shared-capacity impact.

Advanced1 story

Adaptation, data, and feedback​

12.6No story yet

Adaptation choice and behavior-update lifecycle

Choose prompting, examples, SFT, PEFT, preferences, reward adaptation, or no training from the behavior, evidence, update rate, and rollback contract.

No story has been commissioned for this area.

12.7No story yet

Training and synthetic-data provenance, diversity, contamination, and evaluation separation

Preserve source rights, generation lineage, coverage, filtering, deduplication, diversity, and frozen evaluation boundaries.

No story has been commissioned for this area.

12.8No story yet

Feedback, preference, reward, and grader specification

Translate intended behavior into bounded feedback without mistaking a proxy or verifier for the task objective.

No story has been commissioned for this area.

12.9No story yet

Continual/domain adaptation, forgetting, adapter composition, and rollback

Change one capability without silently erasing another; operate compatibility, coexistence, and rollback.

No story has been commissioned for this area.

Inference policy, serving, and change​

12.10Story available

Model portfolios, routing, cascades, risk vetoes, and fallback semantics

Route only within an eligible set; distinguish route, cascade, fallback, retry, and abstention; evaluate rare slices and full-path cost.

Advanced1 story
12.11Story available

Serving phases, workload shape, latency objectives, and admission

Decompose prefill and decode, latency, sequence lengths, concurrency, warm-up, and admission against useful-output service objectives.

Advanced1 story
12.12No story yet

KV and prefix state, continuous batching, locality, and goodput

Balance reuse against queue fairness, load, eviction, trust boundaries, and in-service-level accepted outcomes.

No story has been commissioned for this area.

12.13No story yet

Quantization, speculative decoding, parallelism, and quality validation

Treat serving optimizations as behavioral releases validated on the real runtime and workload.

No story has been commissioned for this area.

12.14No story yet

Disaggregated inference, model/adapter placement, and topology-aware transfer

Compare colocated, partial, and separated execution using interference, KV transfer, locality, compatibility, and workload evidence.

No story has been commissioned for this area.

12.15Story available

Behavioral versioning, provider change, coexistence, migration, and rollback

Operate the complete behavior tuple across canaries, aliases, mixed versions, provider change, and rollback.

Professional1 story
12.16Story available

Professional capstone — mixed-version adaptive inference fleets

Defend a fleet whose behavior, routing, serving, adaptation, evidence, and tenant commitments change independently.

Professional1 story
Deep-dive inventory

Subject 13: Agentic and multimodal product systems​

All content areas below are research-ready. Story status shows which parts of the subject have been turned into public teaching material.

Shared control foundations​

13.1No story yet

Autonomy envelope and deterministic control

Place each decision in deterministic code, an explicit workflow, a constrained model decision, or a bounded agent loop.

No story has been commissioned for this area.

13.2Story available

Durable execution and authoritative workflow state

Survive loss, deployment, waiting, and failover while separating history, checkpoints, workflow state, and business truth.

Professional1 story
13.3Story available

Effect contracts, uncertain commits, and reconciliation

Convert probabilistic intent into an attributable, retry-safe or reconcilable operation.

Professional1 story
13.4Story available

Hierarchical budgets, deadlines, cancellation, and termination

Bound time, cost, steps, fan-out, authority, and blast radius across a task tree—and verify what actually stopped.

Advanced1 story
13.5Story available

Delegated identity, approval binding, and containment

Prove who acts, on whose behalf, for which target, and under which current effect-specific approval.

Professional1 story
13.6Story available

Semantic observability and trajectory evaluation

Reconstruct decisions, checks, effects, and progress; evaluate both outcome and path across repeated trials.

Professional1 story

Agentic systems specialization​

13.7No story yet

Multi-agent coordination and task ownership

Choose a coordination topology only when its benefit exceeds added cost, authority, and failure modes.

No story has been commissioned for this area.

13.8No story yet

MCP and A2A operational semantics

Operate tool and remote-agent protocols as lifecycle, trust, authorization, cancellation, and reconciliation contracts.

No story has been commissioned for this area.

13.9No story yet

Memory, artifacts, and delayed influence

Define what may be remembered, for how long, with what provenance, and how retraction propagates.

No story has been commissioned for this area.

13.10No story yet

Capstone — mixed-version production autonomy

Roll out, roll back, govern, and incident-manage an agent whose independently versioned parts outlive releases.

No story has been commissioned for this area.

Realtime multimodal systems specialization​

13.11No story yet

Media transport and regional session ownership

Place media edges, session owners, reconnect state, and regional boundaries deliberately.

No story has been commissioned for this area.

13.12No story yet

Cascaded, native, and hybrid speech architectures

Choose speech architecture from latency, evidence, control, accessibility, and operating constraints.

No story has been commissioned for this area.

13.16No story yet

Temporal grounding across audio, image, screen, and video

Bind interpretations and actions to evidence from the correct moment and modality.

No story has been commissioned for this area.

13.17No story yet

Latency, backpressure, admission, and graceful degradation

Allocate latency and capacity across media, perception, generation, tools, and response.

No story has been commissioned for this area.

13.18No story yet

Ephemeral session credentials and continuous action authorization

Separate short-lived media access from current authorization to cause business effects.

No story has been commissioned for this area.

13.19No story yet

Consent, synthetic media, and accessible completion

Preserve consent, identity, disclosure, and an equivalent completion path across modalities.

No story has been commissioned for this area.

State, not a hand-maintained report

Every count, subject, content area, badge, and story link on this page comes from one structured curriculum record. Updating that record updates this dashboard on the next site build.