gitmyhub

llm-speedrunner

★ 0 updated 10mo ago ⑂ fork

The Automated LLM Speedrunning Benchmark measures how well LLM agents can reproduce previous innovations and discover new ones in language modeling.

No plain-English explanation yet — one is being written right now. Check back in a minute.