W4M · Research Notes

Research Notes · The series so far

Three weeks, measured

Every result we have published, on one page — a dated timeline of the cadence, and a map of the whole GoM stack. Nothing here is new; every number links back to the note that reports it.

W4M Research · July 2026 · the series so far

14 notes
published research notes between Jun 30 and Jul 18, 2026 — every figure below links to the note that reports it
6 in a day
the founding drop: Notes 1–6 all went out on Jun 30, 2026 — then the cadence kept going
+57.5 points
just published in Note 14 — the frozen Nemotron Nano's GSM8K number climbed again under a leaner brain: 25.0 → 82.5% (n=200). The pace is accelerating.
How to read this page Two pictures. First, a timeline: what landed, when, in the order it published — the point is the density of the cadence, not any single line. Second, a map of the stack: the small models we train ourselves, the brains that bolt onto someone else's frozen model, and the innovations around both. Bolded numbers are all previously published; each card links to its source note.

Visual 1 — Three weeks, measured

A vertical timeline of the series. Same-day clusters are drawn together on purpose: on Jun 30 six notes went out at once, and the drumbeat has not stopped since.

Honest note This page adds no new claims. Every bolded figure is lifted from a note already published above, and each card carries the link so you can check it in context — we do not put figures on this site before the note that reports them.

Visual 2 — The GoM stack

The same body of work, arranged by what it is rather than when it shipped: the models we train ourselves (small to large), the brains that attach to a frozen third-party model, and the innovations that surround both.

For the technical reader Sources, left to right in the stack. Column 1: 98K/400 KB Sudoku solver at 4,865/4,865 (Note 11); ~37M Hanoi core, 3–8-disk training generalizing to the 30-disk optimum (Notes 3 & 9); ARC program-writer with a 96.7% fresh-transfer verified-program rate (Note 13); 30M→962M ladder at β = 0.245, R² = 0.938 (Note 1). Column 2: 376M brain on the frozen 3B-active Nemotron Nano, GSM8K 25.0→82.5 (a one-week climb of +12.5 → +44.5 → +57.5, completed in Note 14); the same brain line on the 120B Super, 53.5→76.5 (+23, n=200, interim); a 27B host at +26.5 on-device at 4-bit; harm-free across the ARC-E/ARC-C/PIQA/HellaSwag row (Note 13). Column 3: on-device assistant at 42→68% with zero confidently-wrong (Note 6); spaced memory 81% vs 32% at day ten and self-teaching 87→100% (Note 7); portable memory transplant 0→100% and swarm sharing 45.6→100% ×3 (Note 8). Full evidence, configs and logs are available to qualified partners under NDA.

Three weeks. Every number logged, reproducible, and linked.

This page is the map; the notes are the evidence. Qualified partners can request the full results behind each figure.

w4m.ai