ASSERT
★ 0
updated 12d ago
⑂ fork
Requirement-driven evaluation harness for AI agents and LLM applications. Generate behavior-specific test cases, run them against any target (hosted models, callable wrappers, OTel-traced agents), and inspect local-first artifacts.
No plain-English explanation yet — one is being written right now. Check back in a minute.