LMSteinshark
From corpus collection and a billion-parameter first model to a smaller, better-instrumented training system.
Each series is arranged chronologically. Earlier pages describe the original implementation; the latest page records the current architecture and status.
From corpus collection and a billion-parameter first model to a smaller, better-instrumented training system.
From the research idea to supervised bootstrapping and a complete self-play pipeline.