A small, fully local project that measures — not just claims — the effect of common cost/memory optimizations used in large-scale LLM training, and empirically reproduces the loss-vs-compute scaling ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results