Just want to add that @tydsh himself is a great example of down-to-earth do-er! Last week he asked me for some model checkpoints, and immediately ran a bunch of experiments to discover some important findings. You can’t be a great LLM researcher without being down-to-earth! 🫡
Excited to share these preliminary results on our internal autoresearch system @Recursive_SI, where we achieve SOTA on nanochat / nanogpt speedrun / kernel benchmarks using the same underlying system without task-specific adaptations.
blog: