* feat(site): interactive training-foundations figures in 5 lessons
Add five theme-aware interactive widgets to lesson-figures.js, embedded
via the existing ```figure fence:
- gradient-descent (P1.08 optimization): drag learning rate, watch the
descent path converge or diverge past lr > 1
- softmax-temperature (P3.04 activations): divide logits by T, reshape
the distribution from argmax to uniform
- bias-variance (P2.10): slide model complexity across the U-shaped
test-error curve, see the sweet spot move
- l2-regularization (P3.07): raise lambda, watch every weight shrink
- lr-schedule (P3.09): compare warmup, cosine, step, exponential decay
Validated headless: all five mount with no console errors, sliders and
selects drive re-render, both light and dark themes render correctly.
* feat(site): interactive LLM-internals figures in 5 lessons
Batch 2, building on the same widget system:
- sampling-decoder (P10.04 mini-gpt): temperature then top-k then top-p
filtering over the logits, survivors renormalized
- scaling-laws (P7.13): Chinchilla loss from params and tokens, with the
20-tokens-per-parameter compute-optimal rule
- quantization (P10.11): bits per weight against model size and the
precision lost at fp16/int8/int4/int2
- rope-explorer (P7.04): rotary frequencies across position and dimension,
base controls wavelength and usable context
- lora-params (P11.08): rank against the 2r/d trainable fraction
Validated headless: all five mount with no console errors, sliders and
selects drive re-render, both light and dark render correctly.
* feat(site): interactive evaluation and representation figures in 5 lessons
Batch 3, same widget system:
- precision-recall-threshold (P2.09 model-evaluation): slide the cutoff
across two class distributions, watch precision/recall/F1 trade
- cross-entropy-loss (P3.05 loss-functions): -log(p_true), the price of
being confident and wrong
- cosine-similarity (P11.04 embeddings): the angle between two vectors is
the similarity, magnitude drops out
- tokenizer-tradeoff (P10.01 tokenizers): vocab size against tokens-per-word
and the embedding table cost
- rag-chunking (P11.06 rag): chunk size, overlap, and top-k against chunk
count and context tokens per query
Validated headless: all five mount with no console errors, math checks out
(thr 0.8 -> P 1.00/R 0.11, -ln(0.05)=2.996, cos 90 deg = 0, 224 chunks),
sliders drive re-render, both light and dark render correctly.
* feat(site): interactive figure system — 74 new widgets across 11 phases
Expand the lesson-figure system from a handful of widgets into a curriculum-wide
library. Refactor lesson-figures.js to expose a shared LF toolkit (el, svgEl,
slider, select, fmtInt, clamp, lerp, raf, register) and split widgets into eight
per-phase module files that plug in via LF.register.
New module files (3,682 LOC) and the concepts they make draggable:
- figures-math.js (P1, 11): vector projection, matrix transform + determinant,
eigenvectors, derivative tangent, chain rule, gaussian, bayes update,
entropy/KL, PCA axes, fourier synthesis, convex vs nonconvex
- figures-ml.js (P2, 10): regression fit/MSE, logistic boundary, SVM margin,
kNN smoothness, k-means steps, tree depth, feature scaling, naive bayes,
class imbalance, k-fold CV
- figures-dl.js (P3, 9): perceptron boundary, MLP forward pass, vanishing
gradients, optimizer trajectories, weight-init variance, dropout, batchnorm,
learning curves, gradient clipping
- figures-vision-speech.js (P4/P6, 8): convolution kernel, pooling, receptive
field, conv output size, CNN params, spectrogram window, mel scale, aliasing
- figures-transformers.js (P5/P7, 9): attention heatmap, multihead split, causal
mask, sqrt(d_k) scaling, word2vec arithmetic, BPE merges, GQA sharing,
residual stream, flash-attention memory
- figures-genai-rl.js (P8/P9, 9): diffusion denoise, noise schedule, VAE latent,
GAN minimax, Q-learning gridworld, value iteration, epsilon-greedy, discount
horizon, policy-gradient ascent
- figures-llms-systems.js (P10/P12/P13, 9): beam search, speculative decoding,
MoE routing, context window, perplexity, continuous batching, ViT patches,
multimodal fusion, MCP round trip
- figures-agents-alignment.js (P11/P14/P16/P18, 9): agent loop, ReAct trace,
tool routing, swarm message scaling, supervisor tree, RLHF reward-KL,
DPO margin, context budget, guardrail gates
Each widget embedded in its lesson via the figure fence (74 lessons). All
theme-aware through CSS vars, vanilla ES5, no dependencies.
Validated headless: all 90 registered figures (16 prior + 74) mount with zero
console errors in a master harness; rich SVG visualizations (attention heatmap,
gridworld policy, convolution feature map, swarm graphs) render correctly in
both light and dark.
* feat(site): 44 more interactive figures — NLP, LLM internals, infra, autonomy
Wave 2 extends the figure system into the phases that were still bare,
plus deeper coverage of the large NLP and LLM phases. Five new module
files (2,219 LOC), each plugging into the shared LF toolkit:
- figures-math2.js (P1, 9): SVD low-rank reconstruction, tensor broadcasting,
log-sum-exp stability, Lp unit balls, monte-carlo pi, system conditioning,
random-walk diffusion, roots of unity, graph degree
- figures-nlp2.js (P5, 8): BoW/TF-IDF, RNN unroll, LSTM gates, seq2seq
alignment, edit distance, n-gram backoff, BIO tagging, sentiment logits
- figures-llms2.js (P10, 9): RMSNorm vs LayerNorm, SwiGLU, RLHF pipeline,
DPO loss, paged KV cache, expert capacity, sliding-window attention,
differential attention, weight tying
- figures-infra.js (P17, 9): data/tensor/pipeline parallelism, ZeRO sharding,
GPU memory breakdown, throughput-latency, autoscaling, cost-per-token,
roofline
- figures-frontier.js (P15/P19, 9): task decomposition, reflection loop,
memory consolidation, world-model rollout, autonomy oversight, pass@k,
eval-harness matrix, canary rollout, trace spans
Embedded in 44 lessons via the figure fence. Validated headless: all 134
registered figures (16 core + 118 module) mount with zero console errors in
a full harness; pipeline-bubble, SVD energy, and trace-span visualizations
render correctly in light and dark.
* fix(site): address review findings on figure widgets
- sampling-decoder: formula now reads 'cumulative >= p' (nucleus keeps the
smallest set covering p, matching the implementation)
- supervisor-hierarchy: drop the dead capped-total accumulator; show the exact
geometric total and note when the diagram caps a level at 64 so the number
and the drawn nodes stay consistent; handle b=1 (total = depth + 1) instead
of the closed form that is undefined at b=1
- image-patch-tokens: use ceil(size/patch) so non-divisible sizes count the
partial patch row; formula shows the ceil and meta notes the padded size
- debugging-neural-networks: normalize the one-off Type 'Practice' to 'Build'
Verified in browser: all three widgets render with the corrected text/math,
no console errors.
Skipped: the 'figure fence is not an approved language tag' findings. lesson.html
keys on codeLang === 'figure' to emit the widget mount point; the fence body is
the figure id. Renaming the fence to the figure id would stop it rendering.
There is no fence-language allowlist for these lesson docs.
About page shipped without the inline theme bootstrap every other page
has, so the theme toggle was dead and the page was stuck on light. Add
the same localStorage/matchMedia bootstrap + toggle wiring.
It also cleared the 64px fixed header with only 64px top padding, so the
eyebrow tucked under the header. Bump .about top padding to 100px (80px
mobile) to match the glossary page.
Close three harness-engineering gaps in the Agent Workbench mini-track:
- 33 (Instructions): progressive disclosure — thin AGENTS.md router + tiered docs
- 36 (Scope Contracts): feature_list.json as the project-level scope primitive
- 40 (Handoff): leave a clean state — cleanup phase before the handoff packet
* feat(site): add About page + nav/footer links + /about rewrite
* fix(site): add command palette trigger to About page header
About page loaded cmdpalette.js but had no [data-cmd-palette] trigger,
unlike the other five pages. Insert the same search-toggle button
between </nav> and the theme toggle so Cmd-K and click both work.
* feat(site): interactive lesson figures + KV-cache sizer
Adds an in-lesson interactive figure layer. Authors drop a fenced block in
docs/en.md:
```figure
kv-cache
```
which the lesson renderer hydrates into a real widget (sliders, live output),
theme-aware via the site's CSS vars. First widget: a KV-cache sizer — drag
sequence length, batch, layers, kv-heads, head-dim, dtype and watch the cache
size cross a single GPU's memory. Wired into 07/12 (KV cache & FlashAttention).
Mechanism: `figure` fenced block -> <div class="lesson-figure" data-figure>,
mounted by lesson-figures.js after render. No deps; figures live in lessons,
not on the homepage. Validated interactivity + light/dark parity.
* feat(site): animated figures in lesson content + delegate from fenced block
The fenced ```figure``` block now mounts both interactive widgets (defined in
lesson-figures.js) and the animated SVG explainers (figures.js), via one
syntax. Embeds animated figures directly in lesson bodies:
- attention-matrix -> 07/02 self-attention
- transformer-block -> 07/05 full transformer
- tokenizer-bpe -> 10/01 tokenizers
- kv-cache-sizer -> 07/12 (interactive sliders)
Animated figures render live in normal browsers and fall back to a clean
static frame under prefers-reduced-motion. Validated all four mount via the
lesson path; light/dark parity.
* docs(07/02): replace ASCII pipeline with mermaid flowchart
* fix: update dataset path for Rotten Tomatoes in load_and_inspect and stream_dataset functions
* fix: improve formatting of dataset split print statements
* fix: update Hugging Face IDs for dataset recommendations in prompt-data-helper
* fix: update Hugging Face IDs and configurations in dataset recommendations
CodeRabbit on #256:
- Insert-or-replace the README STATS block: if the markers are missing or
mangled, re-insert before "## How this works" instead of silently doing
nothing, so the README can't drift from site/stats.json.
- Replace the empty catch with a console.warn so a malformed stats.json is
visible. Kept it a warning, not a CI hard-fail: bad analytics JSON should
not break the whole site build.
145,598 readers and 234,496 page views (last 30 days) now show under the
hero. The numbers live in a single source (site/stats.json) and build.js
regenerates the README block on each build; it also keeps the lessons
badge in sync with the live count. Vercel has no analytics API, so refresh
stats.json from the dashboard and re-run build to propagate.
PR #246 indexed 35 lessons (468 -> 503) but the marketing counts stayed
hardcoded: index/catalog/lesson/prereqs HTML and the command palette still
showed 473/435 lessons and 489 outputs, so the live site read stale.
build.js now rewrites every "<n> lessons / phases / outputs" string to the
count it computes, so these never drift again. Synced all pages to
503 lessons, 20 phases, 499 outputs.
35 completed lessons were merged to main but never wired into the site
source (README), so the catalog and lesson pages could not reach them.
- Phase 19 capstones 58-87 added to README (tracks E. Multimodal VLM,
F. Advanced RAG, G. Eval framework, H. Distributed train, I. Safety
harness); the phase-19 summary now reads 85 lessons / 9 deep-build tracks
- Phase 19 capstones 20-87 added to ROADMAP (it only listed 01-17)
- 5 cross-phase orphans wired into README + ROADMAP:
07/15 attention-variants, 07/16 speculative-decoding,
08/19 visual-autoregressive-var, 10/25 speculative-decoding,
10/34 gradient-checkpointing
- Corrected stale counts; rebuilt site/data.js (468 -> 503 lessons)
Tables (closes#193):
- Parse table rows with an escape/code-aware splitter so escaped pipes
(P(X\|Y)) and pipes inside inline-code spans no longer shatter cells
- Normalize each row to the header column count
- Split markdown on /\r?\n/ so CRLF docs do not leak a trailing \r
Diagrams (closes#233):
- Replace the non-rendering block-beta protocol landscape in
phase 16/03 with a standard flowchart, and convert the stateDiagram
note from literal \n to block-note syntax
- Add a hue-preserving HSL fill transform to the mermaid renderer:
in dark mode light fills are darkened, in light mode dark fills are
lightened, so hard-coded node colors stay legible in both themes
- Check the issuer allow-list before any JWKS lookup/refresh so an
untrusted iss returns "iss not allowed" and never triggers a refresh
- Map each tool to its required scope; destructive notes.delete now needs
mcp:tools.delete (outside the minimal scopes_supported), so the step-up
bonus actually demonstrates an incremental scope gate
- Replace asserts in Client.discover/register with explicit ValueErrors so
the contract checks survive python -O
- Add Client ID Metadata Documents (CIMD) as the recommended default
enrollment path; reframe Dynamic Client Registration as a backwards-
compatible fallback (spec demoted DCR from SHOULD to MAY)
- Fix JWKS handling: separate authorization-server key rotation from
resource-server cache refresh. The cache-miss fallback now re-fetches
(idempotent) instead of minting a key, closing an attacker-controlled
key-creation DoS and a logic bug where the minted kid never matched
- Emit resource_metadata in WWW-Authenticate per RFC 9728 5.1
- Correct canonical resource URI guidance (path may be retained when it
identifies an individual server)
- Add RFC 9207 iss / mix-up attack coverage; distinguish audience replay
(access-token privilege restriction) from the confused-deputy problem
- Refresh IdP capability matrix, key terms, exercises, further reading
- Rewrite code/main.py and the skill output to match
Regenerate site/data.js.