Gyroscope report

Influence functions: what you did, and where to restart

Covers Wed 9 Sep – Wed 16 Sep 2026 (weeks 37–38). Generated Thu 17 Sep 2026 (week 38), 09:45 PDT on the Intel MacBook Pro. Sources: Codex and Claude Code session logs, the live Google Doc, GitHub, check-ins, Tickertape, Gmail and earlier reports. The M5 was asleep, so its local session logs were not read.

Bottom line

Where everything lives

What Where State
Your project doc project: influence functions Created Sat 12 Sep 15:38, last edited 16:46. Nothing since
Prep reports (Wed 9 Sep) Grosse + Wick–KFAC plan · paper author histories · prior contact + ASTRA authors ⚠️ The Wick pitch is superseded
Background notes (Wed 9 Sep) influence_functions_notes.md (Grosse 2023, why the Hessian matters) · astra_notes.md (the ASTRA paper) In Wednesday's day dir, synced
MNIST ASTRA lab (Codex) ~/git/influence-functions: tutorial · results · W&B project 🛑 No git, Intel only
Kaczmarz tutorial (Codex) linear-influence On GitHub, live
3 tutorial pages (Fable, second Claude account, M5) The Kaczmarz Way · The Price of Influence · The Response Curve On GitHub, live
MNIST influence atlas (Codex) mnist-influence-atlas (needs sign-in); source in the lab's mnist_experiment/ Never checked against retraining
Mini-lab (scheduled pass) mini-lab report Ran in a scratchpad; only the report was saved
Two toy scripts wick_vs_kfac_check.py, astra_preconditioner_toy.py in ~/.codex/attachments/b4345e40-7c7b-48e5-8e07-9ad065ef981f/ The only copies I found

The sessions

When (PDT) Tool What you asked
Wed 9 Sep 13:03–16:31 Claude Code 84b11447, Intel Find the 11 Jun Grosse call and how to prepare; what influence functions are; paper author histories; past contact with Roger
Wed 9 Sep 16:29 → Thu Claude Code 1d58dad2, Intel Who you know at Anthropic → people map
Wed 9 Sep ~17:15 claude.ai chat, M5 Understanding influence functions and Hessian computation
Wed 9 Sep 17:18–17:47 Codex, 5 sessions in the lab Reuse the Ciresan code from gradient-dissent, implement ASTRA on MNIST, keep Modal under $10, log to W&B, write a tutorial
Wed 9 Sep 17:43–17:57, 20:09 Codex Costs, uses, and the name for how predictions move after one step. Answer: the NTK, or TracIn for loss changes
Wed 9 Sep 18:20–18:29 Codex Interactive 2-D Kaczmarz tutorial: add a line, scale its weight, watch the other residuals
Sat 12 Sep 15:39 Claude Code eb908a8e, Intel "Remind me what I did on Wednesday". It found only the Claude half
Sat 12 Sep 16:13–16:39 Fable, session The three tutorial pages
Sat 12 Sep 16:55–17:07 Codex, 4 sessions The atlas
Mon 14 → Tue 15 Sep 04:33 Scheduled agent-requests pass The mini-lab, from your Monday shower thought

Day by day

What the work found

  1. ASTRA solves more accurately, but no method predicted retraining. The lab ran the 11.97M-parameter Ciresan MLP through EK-FAC → ASTRA → PCG and 24 half-data retrains. Mean LDS: EK-FAC 0.014, ASTRA −0.033, PCG 0.030, and every per-query 95% interval includes zero. The Wednesday Modal bill was $0.34. SOURCE was not implemented.
  2. At this scale the ground truth is noise. In the mini-lab, two seeds of the same leave-one-out run agree at Spearman −0.007. Score against PBRF or dattri's pre-retrained MNIST models, not cold-start retrains.
  3. Wick/Isserlis is neither new nor large. The cross-covariance is the layer's mean gradient, so it vanishes at convergence (0.08%).
  4. What survives: K-FAC's error at convergence is the basis, not the eigenvalues. On a small MLP's last layer, relative error is EK-FAC 0.671, best rank-1 Kronecker 0.557, rank-2 0.434, and 45% of the mass lies where EK-FAC can't reach. Hong, Eschenhagen, Mlodozeniec & Turner (2509.23437) say the opposite, but their own share falls 60 → 58 → 41% with training. Runa Eschenhagen is on Roger's team. A per-layer, per-epoch sweep settles it in about an hour on the laptop. Kronecker sums have prior art (Koroko et al., 2201.10285).
  5. ⚠️ The 9 Sep "half-hour falsification test" is a trap as written. MNIST's constant border pixels give a fake ρ₁ ≈ 0.55.
  6. The atlas uses a CNN at 0.39% test error. It shows the top-5 training digits for the 100 most ambiguous test digits. It's a good picture for the post, but nothing validates it.

What you planned, and where it stands

Plan Source Status
Research → blog post → apply in week 38 Your doc, Sat 12 Sep Research done; no post; not applied
"understand influence functions more, do influence on MNIST" Your doc's last line Done since, by the atlas and the mini-lab
Post sent to Roger raw, before Fri 18 Sep Kateryna Peters, Fri 11 Sep Not done Sat or Wed; planned for today
Apply, "maybe… pushed to Friday" Check-in, Wed 16 Sep 10:53 Posting live today
08:40–11:45 post; 13:20–16:30 application Today's directions Today

Restart: suggested order

  1. Commit the lab (10 min). Run git init, add the two toy scripts from ~/.codex/attachments/…, and push. The existing .gitignore should keep .venv, node_modules and checkpoints out.
  2. Apply (20 min): Research Engineer / Scientist, Alignment, San Francisco. It doesn't need the post, and the post has already slipped twice.
  3. Write the post from what exists, and send it raw in Roger's "catch up?" thread. - Intuition from the Kaczmarz pages, cost from The Price of Influence, one atlas example. - Then the honest MNIST result: better solves, no LDS gain, noisy ground truth. - End on the basis-vs-eigenvalue question as your next experiment. Leave Wick out.

Corrections made this pass: the lab tutorial is 2,760 words, not the "20,000" the mini-lab report gave (that was its size in bytes). The 9 Sep report now marks its Wick pitch and draft message as superseded.

Not checked: the M5's local logs (ssh timed out), the transcripts of the claude.ai chat and the Fable session (web only), and the atlas page itself (sign-in).

Total life satisfaction

This is one of the two things this week is for. The research part is finished. What's left is a 20-minute application and one raw email, and on 11 Sep you said yourself that the blocker is showing raw work, not the work. Sending both before Kateryna Peters tomorrow turns eight days of preparation into the result you set out to get.