19-day current streak·31-day longest streak
-
beam
No description.
Python ★ 16 5mo agoExplain → -
bibtex-mcp
No description.
Python ★ 9 1y agoExplain → -
Athena
No description.
Python ★ 2 1y agoExplain → -
yet-another-applied-llm-benchmark ⑂
A benchmark to evaluate language models on questions I've previously asked them to solve.
★ 1 1y agoExplain → -
teleport-contest ⑂
No description.
JavaScript ★ 0 7h agoExplain → -
discord-py-self-mcp
Read-only Discord MCP server
Python ★ 0 10d agoExplain → -
verifiers ⑂
Our library for RL environments + evals
★ 0 9h agoExplain → -
harbor ⑂
Harbor is a framework for running agent evaluations and creating and using RL environments.
★ 0 1mo agoExplain → -
rlhf-book ⑂
Textbook on reinforcement learning from human feedback
★ 0 1mo agoExplain → -
dehub ⑂
A TUI to de-GitHub yourself. Control PRs, Actions, Issues, Notifications.
Go ★ 0 1mo agoExplain → -
PostTrainBench ⑂
Measuring how well CLI agents like Claude Code or Codex CLI can post-train base LLMs on a single H100 GPU in 10 hours
★ 0 1mo agoExplain → -
tracked-models ⑂
List of HuggingFace models tracked by Interconnects AI for open model analytics
★ 0 6mo agoExplain → -
lighteval ⑂
Lighteval is your all-in-one toolkit for evaluating LLMs across multiple backends
★ 0 8mo agoExplain → -
openbench ⑂
Provider-agnostic, open-source evaluation infrastructure for language models
★ 0 7mo agoExplain → -
openrouter-tool-check
No description.
Python ★ 0 10mo agoExplain → -
tools
No description.
Python ★ 0 1y agoExplain → -
blogcaster ⑂
Python tools for easily translating your blog content to podcasts & YouTube
Python ★ 0 1y agoExplain → -
fh-bootstrap ⑂
fasthtml wrapper for bootstrap
★ 0 1y agoExplain → -
ttok ⑂
Count and truncate text based on tokens
★ 0 1y agoExplain →
No repos match these filters.