<feed xmlns='http://www.w3.org/2005/Atom'>
<title>prime-grokking.git/src, branch main</title>
<subtitle>Prime Grokking — can minimal architectures grok next-prime? Pre-registered negative-class experiment (weight-tied RNN vs transformer).</subtitle>
<id>http://git.jayrup.me/c/prime-grokking.git/atom?h=main</id>
<link rel='self' href='http://git.jayrup.me/c/prime-grokking.git/atom?h=main'/>
<link rel='alternate' type='text/html' href='http://git.jayrup.me/c/prime-grokking.git/'/>
<updated>2026-08-30T21:04:02Z</updated>
<entry>
<title>fix(eval): structured-mode greedy decode capped at in-range layout bound (369) — probe-window traces reach 633 tokens and were silently truncated; cap now keys to positional table (1024) so OOD traces decode fully. Regression-verified vs stub model.</title>
<updated>2026-08-30T21:04:02Z</updated>
<author>
<name>Void Agent</name>
<email>void@jayrup.hermes</email>
</author>
<published>2026-08-30T21:04:02Z</published>
<link rel='alternate' type='text/html' href='http://git.jayrup.me/c/prime-grokking.git/commit/?id=3135fea726157f8ea596758ce246dfb74779c865'/>
<id>urn:sha1:3135fea726157f8ea596758ce246dfb74779c865</id>
<content type='text'>
</content>
</entry>
<entry>
<title>fix(transformer): increase positional embedding and causal mask buffer to 1024 to support worst-case probe traces</title>
<updated>2026-08-29T22:35:10Z</updated>
<author>
<name>CaptainJack2491</name>
<email>jayrupnakawala@gmail.com</email>
</author>
<published>2026-08-29T22:35:10Z</published>
<link rel='alternate' type='text/html' href='http://git.jayrup.me/c/prime-grokking.git/commit/?id=fc56dd007142264ef029eb8c0f9db11a6cf3d298'/>
<id>urn:sha1:fc56dd007142264ef029eb8c0f9db11a6cf3d298</id>
<content type='text'>
</content>
</entry>
<entry>
<title>optimize E8: precompute causal mask buffer, add eval early exit, and enable compile_model in e8.csv</title>
<updated>2026-08-29T22:32:30Z</updated>
<author>
<name>CaptainJack2491</name>
<email>jayrupnakawala@gmail.com</email>
</author>
<published>2026-08-29T22:32:30Z</published>
<link rel='alternate' type='text/html' href='http://git.jayrup.me/c/prime-grokking.git/commit/?id=9a5912f7764dbc942b56f8114f2c878c1a6d44ff'/>
<id>urn:sha1:9a5912f7764dbc942b56f8114f2c878c1a6d44ff</id>
<content type='text'>
</content>
</entry>
<entry>
<title>implement Addendum 8: 6-arm token-space recurrence suite (arms A, B, C, D1, D2, D3), tests, and jobs/e8.csv</title>
<updated>2026-08-29T22:19:29Z</updated>
<author>
<name>CaptainJack2491</name>
<email>jayrupnakawala@gmail.com</email>
</author>
<published>2026-08-29T22:19:29Z</published>
<link rel='alternate' type='text/html' href='http://git.jayrup.me/c/prime-grokking.git/commit/?id=7b0fddb02b089b82b8d12b2dcac17bf9527817da'/>
<id>urn:sha1:7b0fddb02b089b82b8d12b2dcac17bf9527817da</id>
<content type='text'>
</content>
</entry>
<entry>
<title>eval: strip torch.compile _orig_mod./module. prefixes when loading checkpoints (salvages compiled-GPU runs)</title>
<updated>2026-08-24T23:33:03Z</updated>
<author>
<name>Void Agent</name>
<email>void@jayrup.hermes</email>
</author>
<published>2026-08-24T23:33:03Z</published>
<link rel='alternate' type='text/html' href='http://git.jayrup.me/c/prime-grokking.git/commit/?id=4b0c3cdc3db5b29c9d770082bf6caeb677221bfa'/>
<id>urn:sha1:4b0c3cdc3db5b29c9d770082bf6caeb677221bfa</id>
<content type='text'>
</content>
</entry>
<entry>
<title>fix(train): enforce epoch permutation without replacement, add max cache OOM guard, and handle empty split tensors</title>
<updated>2026-08-20T18:25:57Z</updated>
<author>
<name>CaptainJack2491</name>
<email>jayrupnakawala@gmail.com</email>
</author>
<published>2026-08-20T18:25:57Z</published>
<link rel='alternate' type='text/html' href='http://git.jayrup.me/c/prime-grokking.git/commit/?id=615c87402138da053abe2cd0ab897eb6a5af9d67'/>
<id>urn:sha1:615c87402138da053abe2cd0ab897eb6a5af9d67</id>
<content type='text'>
</content>
</entry>
<entry>
<title>perf(train): VRAM-resident tensor caching, zero-copy GPU batch sampling, and adaptive eval schedule</title>
<updated>2026-08-20T16:37:46Z</updated>
<author>
<name>CaptainJack2491</name>
<email>jayrupnakawala@gmail.com</email>
</author>
<published>2026-08-20T16:37:46Z</published>
<link rel='alternate' type='text/html' href='http://git.jayrup.me/c/prime-grokking.git/commit/?id=be8d62351497ed2e5ee136ae533ee9c5ec4491f2'/>
<id>urn:sha1:be8d62351497ed2e5ee136ae533ee9c5ec4491f2</id>
<content type='text'>
</content>
</entry>
<entry>
<title>fix(eval): correct sieve-rank indexing + test expectations</title>
<updated>2026-08-20T16:02:15Z</updated>
<author>
<name>Void Agent</name>
<email>void@jayrup.hermes</email>
</author>
<published>2026-08-20T16:02:15Z</published>
<link rel='alternate' type='text/html' href='http://git.jayrup.me/c/prime-grokking.git/commit/?id=cf83937689bb30e2b5fa6e3efaa2f115016a030b'/>
<id>urn:sha1:cf83937689bb30e2b5fa6e3efaa2f115016a030b</id>
<content type='text'>
- _compute_sieve_rank: find LARGEST k (not smallest) whose signature
  contains all error preds — fixes P5(4) vs P5(1) bug
- _sieve_rank_signature: iterate all small_primes (not just p²≤n) to
  catch composites like 209=11*19 where 11²&gt;209
- Test expectations updated for [101,200] range (209 outside it)
- is_prime task: P1→P3 (no sieve-rank for classification tasks)
48/48 green.
</content>
</entry>
<entry>
<title>fix(eval): P5/P6 sieve-rank ladder replaces hardcoded FLAGGED_SIEVE_PREDS</title>
<updated>2026-08-20T15:55:49Z</updated>
<author>
<name>Void Agent</name>
<email>void@jayrup.hermes</email>
</author>
<published>2026-08-20T15:55:49Z</published>
<link rel='alternate' type='text/html' href='http://git.jayrup.me/c/prime-grokking.git/commit/?id=f8995751de925571d4eec669ab7ab8b1c3d76cb9'/>
<id>urn:sha1:f8995751de925571d4eec669ab7ab8b1c3d76cb9</id>
<content type='text'>
- probe_report: dynamic lo/hi (no more hardcoded [101,200])
- main(): probe range derived from range_end ([range_end+1, range_end+1000])
- P6 = exact (no probe misses), P5(k) = errors match rank-k sieve signature
- _sieve_rank_signature: computes composites with all factors &gt; p_k
- _compute_sieve_rank: identifies rank from model's error predictions
- build_report.py: probe_fig uses sieve rank for coloring
- Tests updated for P5/P6 codes (48/48 green)
</content>
</entry>
<entry>
<title>perf(cuda): add AMP fp16 autocast with GradScaler and torch.compile reduce-overhead mode</title>
<updated>2026-08-17T15:35:28Z</updated>
<author>
<name>CaptainJack2491</name>
<email>jayrupnakawala@gmail.com</email>
</author>
<published>2026-08-17T15:35:28Z</published>
<link rel='alternate' type='text/html' href='http://git.jayrup.me/c/prime-grokking.git/commit/?id=7166dc47c9bd764e2cd1d7c3d9e046cdc439b58c'/>
<id>urn:sha1:7166dc47c9bd764e2cd1d7c3d9e046cdc439b58c</id>
<content type='text'>
</content>
</entry>
</feed>
