D1Kaggle as the platform: 30 GPU-hours per week, 2x T4held
D2From-scratch only; no fine-tuning, no distillationheld
D358M params: 10 layers x 512 dim, context 1024held
D4TinyStories + FineWeb-Edu corpusheld
D5One change per version, incremental ladderheld
D6Micro-ablation before every full runheld
D7Pre-register experiments with hashed predictionsheld
D8Journal every session; three sources per research claimheld
D9Micro-batch 8 x accum 32: OOM fix at equal effective batchheld
D10Browser-started runs only: API runs cannot read secretsheld
D11Freeze v0.1.0 as the equal-token baselineheld
D12Screen changes at 30M params, 50M tokens, ~30 minheld
D13Journal everything; verify every claimheld
D14Implement Muon+ as experiment E1validated, below threshold
D15E1 at -0.015 misses the 0.02 threshold; do not promoteenforced
D16Fix resume race: single downloader, atomic swap, barrierfixed
D17Expert reprioritization: sweeps first, recipe lock, data before architectureheld
D18Finish v0.1.2 instead of abandoning at 71%held
D19v2 mixture: 60/20/15/5 FineWeb-Edu / TinyStories / Cosmopedia / codebuilt
D20Punch above weight: token efficiency and data quality before scaleheld
D21Distillation rejected: KALIA stays a pure from-scratch lineageclosed
D22Adopt test-time training and self-teaching research trackqueued
D23Invention target: entity memory + in-place test-time weightsqueued
D24Looped depth promoted to front of queuetested, rejected
D25RL self-play and MoE/DSA rejected at this scale, on evidenceclosed
D26Continual-learning recipe planned for post-trainingplanned
D27Publish minimum footprint, honest only, never redistribute dataenforced
D28Rebuild v2 code shard with a per-file license filterdone
D29Build an evaluation harness: probes, bpB, comparison tooldone
D30Reasoning policy: public methods, our own weights, no RLVR yetheld
D31Freeze Muon LR at 0.02; sweep completeclosed
D32Three ancient-text-inspired designs queued (Kautilya, Utsarga, Apoha)queued
D33Four more queued; Jata-patha becomes the reversal experiment X16queued
D34Three Veda-derived designs recordedqueued
D35Gita compilation adapted; prior art cited, not claimed as novelqueued
D36Entity-consistency harness built; novelty claim narrowed after prior-art checkactive
D37Chakravyuha dismissal corrected; Abhimanyu-gap metric builtactive
D38Hash-anchored pre-registration and pre-registered stopping adoptedactive
D39Strategic audit: finish, test, publish, then v0.2.0; freeze the queueheld
D40Quota-exhaustion response: defer the experiment, publish assets nowsuperseded
D41Stop v0.1.2 at step 3,478: converged within noise, quota to v0.2.0done
D42X16 rejected: reversal made the gap worse on both seeds; frozen recipe to v0.2.0enforced
D43The rebuilt corpus redefines the yardstick: v2b val is canonical, v0.1.2 re-baselined on itactive
D44Licence finding disclosed; publish the filtered corpus plus the filter, never the raw corpusactive
D45Accuracy thresholds must sit above the benchmark noise floor; X18 and S-A were both under-poweredactive
D48No architecture arm from a single seed; the gated-residual line is closed as a null resultactive