llm-speedrunner
★ 0
updated 10mo ago
⑂ fork
The Automated LLM Speedrunning Benchmark measures how well LLM agents can reproduce previous innovations and discover new ones in language modeling.
No plain-English explanation yet — one is being written right now. Check back in a minute.