Members
-
horizon
Capable, auditable coding that runs fully offline on a 16 GB machine. A verification-first layer (hard test execution, symbolic checking, agentic repair) that takes a local 7B to parity with its 671B teacher on verifiable tasks. MIT, pre-registered, reproducible.
Python ★ 22 16d agoExplain → -
vexp-swe-bench
Open benchmark for AI coding agents on SWE-bench Verified. Compare resolution rates, cost, and unique wins.
Shell ★ 13 2mo agoExplain →
No repos match these filters.