<feed xmlns='http://www.w3.org/2005/Atom'>
<title>prime-grokking.git/NOTES-E6.md, branch main</title>
<subtitle>Prime Grokking — can minimal architectures grok next-prime? Pre-registered negative-class experiment (weight-tied RNN vs transformer).</subtitle>
<id>http://git.jayrup.me/c/prime-grokking.git/atom?h=main</id>
<link rel='self' href='http://git.jayrup.me/c/prime-grokking.git/atom?h=main'/>
<link rel='alternate' type='text/html' href='http://git.jayrup.me/c/prime-grokking.git/'/>
<updated>2026-08-18T07:44:34Z</updated>
<entry>
<title>E6 Phase 5: complete 16/16 cells (no O1, all P4, no sieve rank)</title>
<updated>2026-08-18T07:44:34Z</updated>
<author>
<name>Void Agent</name>
<email>void@jayrup.hermes</email>
</author>
<published>2026-08-18T07:44:34Z</published>
<link rel='alternate' type='text/html' href='http://git.jayrup.me/c/prime-grokking.git/commit/?id=e0fec42563680396f51a07805187bd0a8539ea15'/>
<id>urn:sha1:e0fec42563680396f51a07805187bd0a8539ea15</id>
<content type='text'>
Safety-net run: rsync'd the remaining 8 cells from ichi (wd-1.0 seeds {1,2},
wd-3.0, batch-128), classified the full grid against Addendum 6. Verdict
unchanged and now complete: no O1 anywhere, P4 out-of-range in all 16 cells,
sieve-rank ladder resolves to 'no sieve' (smallest in-range composite
predictions 1001/1003/1010/1079/1099, all small-factor). In-range only: low-wd
RNN O-PARTIAL (seed-0), wd-1.0 replication seed-stable O4, batch-128 transformer
destabilizes late (train 1.0-&gt;0.55, val 0.83-&gt;0.45). Tracks e6 summary.csv.
</content>
</entry>
<entry>
<title>E6 Phase 5: classify 8/16 finished cells (no O1, all P4, no sieve rank)</title>
<updated>2026-08-17T22:41:02Z</updated>
<author>
<name>Void Agent</name>
<email>void@jayrup.hermes</email>
</author>
<published>2026-08-17T22:41:02Z</published>
<link rel='alternate' type='text/html' href='http://git.jayrup.me/c/prime-grokking.git/commit/?id=6ed4c6ee43907bd78c90b53a57ca2db411467dbe'/>
<id>urn:sha1:6ed4c6ee43907bd78c90b53a57ca2db411467dbe</id>
<content type='text'>
- Apply Addendum 6 P-ladder on [1001,2000] probe (computed on CPU; eval.py still hardcodes [101,200])
- wd 0.01/0.1/0.3: O-PARTIAL + P4 (val EM 0.77-0.84 in-range); wd 1.0: O4 + P4
- Sieve rank: undefined in all cells (smallest composite pred has small factors; no P5(k)/P6)
- 960-&gt;967 in all cells is training-split memorization, not sieve evidence
- Verdict: data pressure did not change outcome class; walls algorithmic, not data-bound
- 4 cells running (wd1.0 seeds 1,2), 4 queued (wd3.0, batch-128) -&gt; 08:30 safety-net cron
</content>
</entry>
</feed>
