1-day current streak·30-day longest streak
Active across 50+ public developer, AI, open source, and research surfaces Code, datasets, package registries, preprints, reproducible demos, agent workflows, and OSS contribution traces. Open Source · Publications · Recently…
Active across 50+ public developer, AI, open source, and research surfaces
Code, datasets, package registries, preprints, reproducible demos, agent workflows, and OSS contribution traces.

























[Open Source](#open-source-focus) · [Publications](#publications) · [Recently Shipped](#recently-shipped) · [Packages](#published-packages) · [Projects](#featured-projects) · [Impact](#impact-at-a-glance) · [Experience](#experience) · [Stats](#github-stats)
8+ years building production systems at Fortune 100 scale
Former SDE at Amazon Web Services • Currently at Southwest Airlines
Deep expertise in ML systems, distributed architectures, and full-stack engineering
<!-- now:start -->
Now: shipped ragdrift (five-dimensional RAG drift detection on crates.io + PyPI), the rust-llm-stack of <!-- rust-llm-stack-count -->5<!-- /rust-llm-stack-count --> small Rust crates, and the mcp-stack of <!-- mcp-stack-count -->14<!-- /mcp-stack-count --> MCP servers in the official MCP Registry — 4 RAG/agent helpers + 10 reliable transforms LLMs reach for tools instead of imagining (CSV, regex, JMESPath, diff, SQL formatting, shell escaping, JSON5, TOML/YAML/JSON, IANA timezones, HTML→Markdown). Plus the @mukundakatta/agent* reliability stack (fit → guard → snap → vet → cast → budget), 6 earlier MCP servers (also in the Registry), GitHub Actions on the Marketplace, <!-- pypi-count -->52<!-- /pypi-count --> PyPI packages, and <!-- ext-open -->288<!-- /ext-open -->+ open PRs across MCP SDKs, FastMCP, claude-code-action, and Anthropic's agent SDK (<!-- ext-merged -->185<!-- /ext-merged -->+ already merged upstream).
<!-- now:end -->
---
Portfolio at a Glance
PUBLIC REPOS
1396
ORIGINALS
871
ACTIVE PROJECTS
834
FORKS
525
ARCHIVED
304
Every repo is indexed in claude-workspace - wired for Multica, Claude Code, Codex, OpenClaw, and Cursor to reason across the portfolio.
---
Profile Maintenance
The profile README is partly managed by scheduled automation. Pull requests run
profile checks that compile the Python refresh scripts and verify the managed
README markers plus the cached stats files before changes merge.
Latest Drop · The Agent Reliability Stack
> 🌐 Live at mukundakatta.github.io/agent-stack - single landing page for the whole 196+ package ecosystem (npm + PyPI + MCP Registry + GitHub Marketplace).
>
> 🤗 Try it live on the HuggingFace Space · jailbreak fixtures on the HF Dataset.
> Six small, focused npm packages that fix the boring problems every long-running agent eventually hits.
> Pure ESM JavaScript, zero runtime deps, TypeScript types in the box. Designed to compose into a pipeline:
> fit → guard → snap → vet → cast → budget.
---
Hackathon Submissions
> <!-- hackathon-entries -->33<!-- /hackathon-entries -->+ entries across <!-- hackathon-events -->17<!-- /hackathon-events --> events shipped end-to-end during contest
> periods: original repos, live demos on Cloud Run / HF Spaces / dev.to,
> narrated demo videos, Apache 2.0. Each entry targets a different
> partner MCP, a different domain problem, or a different writing-contest
> angle. Click the dropdown for the full list.
📋 Full hackathon participation list — click to expand
Google Cloud Agent Builder (ADK) + Gemini 2.5 agents
| Project | Hackathon · Track | What it does | Live |
| --- | --- | --- | --- |
| gemini-splunk-devx-agent | Splunk Agentic Ops · Platform & Developer Experience + Splunk MCP Server bonus | Splunk admin copilot on the Splunk MCP — walks list_apps / list_savedsearches / get_savedsearch / list_kvstore_collections / audit_knowledge_objects and emits a ranked cleanup punch-list for an inherited Splunk Cloud tenant. | Cloud Run · Demo |
| gemini-splunk-security-agent | Splunk Agentic Ops · Security + Splunk MCP Server bonus | Splunk ES + SOAR triage agent on the Splunk MCP — visibly self-corrects from "TRUE POSITIVE" to "FALSE POSITIVE — sanctioned admin activity" when the threat-intel feed, asset owner, change window, and running SOAR playbook all overturn the surface SIEM signal. | Cloud Run · Demo |
| gemini-splunk-agent | Splunk Agentic Ops · Observability + Splunk MCP Server bonus | Production observability agent on the Splunk MCP — walks list_alerts / get_detector / run_search / run_observability_query to diagnose a firing alert end-to-end, quoting alert IDs, detector rules, and SPL output verbatim. | Cloud Run · Demo |
| protocol-sift-agent | FIND EVIL · SANS SIFT | Forensic incident-response agent that visibly self-corrects from "MALWARE CONFIRMED" to "FALSE POSITIVE — sanctioned admin activity" when deeper evidence overturns surface indicators. | Cloud Run · Devpost |
| gemini-connector-agent | Google Cloud Rapid Agent · Fivetran | Triages connector health on the Fivetran MCP — picks the broken connector by name + quotes the verbatim error. | Cloud Run |
| gemini-pipeline-agent | Google Cloud Rapid Agent · GitLab | Walks down from a failed GitLab pipeline to the failing job + stage + log excerpt with a one-shot remediation. | Cloud Run |
| gemini-search-agent | Google Cloud Rapid Agent · Elastic | Turns plain-English log/product questions into Elasticsearch queries via the Elastic MCP, answers with counts copied verbatim. | Cloud Run |
| gemini-data-agent | Google Cloud Rapid Agent · MongoDB | NL-to-MongoDB query agent — discovers collections, reads schema, runs aggregations, copies counts verbatim. | Cloud Run |
| gemini-eval-agent | Google Cloud Rapid Agent · Arize | LLM-evaluation auditor over the Arize Phoenix MCP — picks the worst-performing model + cites the metric and number. | |
| gemini-ops-agent | Google Cloud Rapid Agent · Dynatrace | Production incident investigator on the Dynatrace MCP — answers "what's broken right now" with a structured triage. | |
| gemini-bright-agent | Bright Data Web Data UNLOCKED (lablab.ai) | Research analyst that walks Bright Data MCP — SERP search → Web Unlocker scrape → structured-dataset lookup — with byte-for-byte verbatim citations from the unlocked pages. | Cloud Run · Demo |
| gemini-rpa-agent | UiPath AgentHack 2026 · workflow-automation theme | RPA workflow-diagnosis agent on an n8n-style MCP — pins the failing step in a workflow run, quotes the verbatim error payload, emits the canonical retry. | Cloud Run · Demo |
| gemini-mantle-agent | DoraHacks Mantle Turing Test 2026 ($100K) | On-chain query agent on Mantle MCP — block height, contract reads, TVL, tx receipts, all verbatim from chain state. | Cloud Run |
Web3 + on-chain agents
| Project | Hackathon · Track | What it does | Live |
| --- | --- | --- | --- |
| mantle-agent-attest | DoraHacks Mantle Turing Test 2026 ($100K) | Verifiable AI-agent-run attestations on Mantle EVM L2 — Merkle-root the JSONL audit log, sign with the agent's EVM key, post (runId, root, sig) on-chain. | DoraHacks BUIDL |
Open-weight + edge AI
| Project | Hackathon · Track | What it does | Live |
| --- | --- | --- | --- |
| briefing-32 | Build Small · Backyard AI | 32B-class AI-news briefing — Qwen3-32B + Gradio. Down-port of an every-2hr cron from Groq Llama-3.3-70B onto an open-weight model on a laptop. | HF Space · Demo |
| gemma4-safe-agent | Gemma 4 Challenge · Build track | Tool-using research agent on Gemma 4 via e2b sandbox — five-piece safety stack on top of the open-weight model. | GitHub |
MCP-agent reliability primitives
| Project | Hackathon · Track | What it does | Live |
| --- | --- | --- | --- |
| agent-stack | DevNetwork [AI+ML] 2026 · TrueFoundry: Resilient Agents | Six MCP-agent reliability primitives (fit · guard · snap · vet · cast · budget) for healthcare AI agents. | Devpost |
| ragvitals | DevNetwork [AI+ML] 2026 · TrueFoundry | RAG drift detector — composable detectors for embedding / retrieval / response / latency drift into one DriftReport. | |
| geminilens | DevNetwork [AI+ML] 2026 · TrueFoundry observability | Local-first observability for Vertex AI Gemini agents — per-call USD cost, drift detection, tool egress audit. | Devpost |
Ha
…
-
karna ★ PINNED
Karna — Your Loyal AI Agent Platform. Self-hosted personal AI assistant with multi-channel messaging, extensible skills, and semantic memory.
TypeScript ★ 3 1mo agoExplain → -
rnht ★ PINNED
Rudra Narayana Hindu Temple — community platform for events, donations, and priest scheduling
TypeScript ★ 1 5d agoExplain → -
astra-agent ★ PINNED
Standalone AI agent runtime — tool execution, context management, and multi-model routing for autonomous assistants
TypeScript ★ 2 8d agoExplain → -
chetana ★ PINNED
Chetana — AI Consciousness Research Platform. Test AI models against 14 consciousness indicators from 6 scientific theories.
TypeScript ★ 2 1mo agoExplain → -
hermes-agentmemory
Pull-model episodic memory plugin for Hermes Agent. Real deletes, audit trace, BYO Claude. MIT.
Python ★ 9 7d agoExplain → -
mcp-config-check
Linter for MCP (Model Context Protocol) config files. Python, CLI + library. Supports Claude Desktop, Cursor, Cline, Windsurf, Zed.
Python ★ 3 7d agoExplain → -
buildcost
BuildCost — Construction Cost Estimator. AI-powered construction cost estimation from blueprints.
Python ★ 3 1mo agoExplain → -
crammr
Crack the entrance exam. — AI-tutored prep for JEE, NEET, MCAT, CAT, and the bar. One subject at a time, one question at a time
TypeScript ★ 2 1mo agoExplain → -
aphrodite ▣
Aphrodite — AI UI Component Library. AI-generated UI component library
★ 2 3mo agoExplain → -
demeter ▣
Demeter — AI-Native CMS. AI-native content management system
★ 2 3mo agoExplain → -
inkwise
Design your tattoo. Before the needle. AI tattoo designer with flash from actual shops. See it on your skin before you commit.
TypeScript ★ 2 1mo agoExplain → -
SwarmKit
Swarm intelligence multi-agent framework — orchestrate collaborative AI agent swarms with voting and consensus
Python ★ 2 7d agoExplain → -
MukundaKatta
Profile README — AI/ML engineer portfolio, open-source contributions, and featured projects
Python ★ 1 18h agoExplain → -
fanout
Agentic content studio: one product → 5 platform-tailored drafts → posted from your own browser session. 15 channels, no third-party API keys.
Python ★ 1 1d agoExplain → -
gemma4-safe-agent
Gemma 4 (2B) on Ollama wired through five tiny reliability libs. Build-track entry for the Gemma 4 DEV Challenge.
JavaScript ★ 1 14d agoExplain → -
tool-arg-fuzzy-rs
Rust port of tool-arg-fuzzy. Fuzzy-match LLM args to JSON Schema enum values. No Levenshtein. dep: serde_json.
Rust ★ 1 1mo agoExplain → -
mukunda-ai
Personal portfolio — mukundakatta.dev | AI/ML Engineer
TypeScript ★ 1 7d agoExplain → -
oss-contributions
Public hub for my AI SDK, MCP, eval, and developer tooling contributions
Python ★ 1 8d agoExplain → -
Oradent
Oradent — dental practice management platform for appointments, patient records, and billing
TypeScript ★ 1 1mo agoExplain → -
artigen
AI art community platform built with React Native, Expo, and Supabase
TypeScript ★ 1 1mo agoExplain → -
sadhak
Sadhak (साधक) — AI-powered job search pipeline built on Claude Code. The Seeker.
JavaScript ★ 1 1mo agoExplain → -
bedrock-kit
Small, opinionated AWS Bedrock client wrapper: adaptive throttle, cache-aware cost tracking, structured-output parse-and-repair. Single-cloud, single-purpose.
Python ★ 1 7d agoExplain → -
awesome-prompt-injection-defense
A curated list of tools, papers, datasets, and guardrails for defending LLMs against prompt injection.
★ 1 1mo agoExplain → -
bustrack
Real-time school bus tracking with AI arrival prediction
TypeScript ★ 1 7d agoExplain → -
twincast
Clone yourself on video. — Record once. Speak any script, in any language, on every platform. Made for creators, priced for cre
TypeScript ★ 1 1mo agoExplain → -
triaged
Your inbox, on autopilot. — AI reads every email, sorts it, summarizes threads, and drafts replies in your voice. All for five d
TypeScript ★ 1 1mo agoExplain → -
tintalk
Therapy. Built for teenagers. A text-first mental health app that actually talks like a teen — and escalates to a real human when
TypeScript ★ 1 1mo agoExplain → -
thumblift
Thumbnails that actually get clicked. Describe your video. Get ten MrBeast-quality thumbnails in thirty seconds, every time.
TypeScript ★ 1 1mo agoExplain → -
signoff
Read contracts like a lawyer. Paste a contract. Get the red flags, the boilerplate, and the terms worth negotiating. In plain Engl
TypeScript ★ 1 1mo agoExplain → -
sendrail
Send money home. Instantly. Cheaply. — Stablecoin rails from the US and GCC to India and the Philippines. Settles in seconds, not days.
TypeScript ★ 1 1mo agoExplain → -
scholar
Your all-nighter research assistant. Pick a topic. Scholar reads fifty sources, synthesizes the arguments, and hands you an outline with
TypeScript ★ 1 1mo agoExplain → -
samecrew
Therapy that gets it. — AI-first emotional support for people who share your context. New dads. Immigrants. First-time found
TypeScript ★ 1 1mo agoExplain → -
rightside
Know your rights. In plain English. — AI legal copilot for landlords, tenants, freelancers, and small claims. Cheaper than a lawyer. Faste
TypeScript ★ 1 1mo agoExplain → -
riffly
Sing it in anyone's voice. Covers, duets, parodies. Generate studio-quality vocals from your favorite artists — ethically licen
TypeScript ★ 1 1mo agoExplain → -
restwise
The sleep coach your watch forgot. — Plug in your Oura, Whoop, or Apple Watch. Get an AI coach that tells you what to actually do differe
TypeScript ★ 1 1mo agoExplain → -
replynow
Every Google review, answered overnight. We draft a response to every review — 5 star and 1 star — in your voice. You approve in bulk each mo
TypeScript ★ 1 1mo agoExplain → -
relood
The resale app for kids stuff. — Buy and sell outgrown clothes, toys, and gear. Local pickup. Fixed prices. No haggling, no scams.
TypeScript ★ 1 1mo agoExplain → -
readmate
The reading coach for your first-grader. Kids read aloud. ReadMate listens, gently corrects, and celebrates. Parents get a progress report ev
TypeScript ★ 1 1mo agoExplain → -
quotebolt
Contractors — stop writing quotes by hand. Describe the job in plain words. Get an itemized quote with labor, materials, and markup, ready to s
TypeScript ★ 1 1mo agoExplain → -
quilled
Your newsletter. Ghostwritten. Tell us what happened this week. Wake up tomorrow to a full issue, in your voice, ready to send.
TypeScript ★ 1 1mo agoExplain → -
primeint
Nail your next interview. — AI-powered mock interviews that feel real. Role-specific questions. Feedback that actually helps you
TypeScript ★ 1 1mo agoExplain → -
prdforge
Cursor, but for product managers. — Synthesize customer calls. Draft PRDs. Ship roadmaps. The AI workbench PMs have been waiting for.
TypeScript ★ 1 1mo agoExplain → -
postops
Your social media team, in a box. — Photos in. Posts out. AI writes the captions, picks the hashtags, and schedules it all for you.
TypeScript ★ 1 1mo agoExplain → -
polydub
Dub your videos. In every language. — Upload once. Get perfectly lip-synced dubs in 40 plus languages, at creator-friendly prices.
TypeScript ★ 1 1mo agoExplain → -
platewise
Your dietician, in your camera roll. Snap your plate. See macros, calories, and nutrient gaps. Get a friendlier suggestion for tomorrow.
TypeScript ★ 1 1mo agoExplain → -
platemap
Your menu. Digitized. Nutrition-calculated. Photograph your handwritten menu. Get a clean digital version with calories, allergens, and QR-ready
TypeScript ★ 1 1mo agoExplain → -
pitchr
Agency proposals, written in ten minutes. Tell us about the client and the scope. Get a polished proposal with pricing and timeline, on your l
TypeScript ★ 1 1mo agoExplain → -
pindrop
The editor for indie podcasters. Drop your raw recording. Get a clean cut with chapters, show notes, and a transcript — ready to publ
TypeScript ★ 1 1mo agoExplain → -
patchly
The code review bot that actually helps. Reviews every PR in minutes. Flags the bugs, suggests the fix, and explains why — like a senior engi
TypeScript ★ 1 1mo agoExplain → -
moontrack
Your cycle. Understood. A cycle tracker that respects your privacy and actually gives useful coaching — for energy, sleep, a
TypeScript ★ 1 1mo agoExplain → -
memoryloop
Bring old photos back to life. — Restore grainy family photos in one tap. Turn them into short videos you actually want to share.
TypeScript ★ 1 1mo agoExplain → -
mathkin
Math that clicks. One AI tutor per grade, K through 10. Homework help. Test prep. Gentle drills. All in one app.
TypeScript ★ 1 1mo agoExplain → -
headshotly
Studio-grade headshots. From your selfies. Forty professional headshots in twenty minutes. No photographer, no studio, no suit required.
TypeScript ★ 1 1mo agoExplain → -
gridcraft
An Instagram grid that actually looks good. Plan six months of posts visually. See your grid before you post it. Never break the aesthetic again
TypeScript ★ 1 1mo agoExplain → -
gigledger
Taxes for the 1099 life. — Track income, expenses, and estimated taxes for rideshare, freelance, and every side hustle in betwe
TypeScript ★ 1 1mo agoExplain → -
fridgechef
What's for dinner? Ask your fridge. Open the door. Snap a photo. Get three dinner ideas using only what's inside. No grocery trip requir
TypeScript ★ 1 1mo agoExplain → -
followy
Meetings that actually ship. — An AI teammate that joins every call, files Jira tickets, Slacks the follow-ups, and books the next
TypeScript ★ 1 1mo agoExplain → -
flickup
One podcast. Sixty shorts. Drop your hour-long podcast in. Get sixty scroll-stopping vertical clips with captions, tuned per pl
TypeScript ★ 1 1mo agoExplain → -
feather
The CRM built for a team of one. Freelancers, consultants, solo founders. Track clients, deals, and follow-ups without Salesforce-lev
TypeScript ★ 1 1mo agoExplain → -
codegarden
Coding for teens. That doesn't feel like school. Your kid builds a game, a Discord bot, or a mod — guided by an AI tutor that speaks their language.
TypeScript ★ 1 1mo agoExplain → -
chartly
Your AI scribe for the exam room. Listens to the patient conversation. Writes the SOAP note. Cuts physicians' documentation time by 75
TypeScript ★ 1 1mo agoExplain → -
chaptr
Your paid community, done right. — A clean, fast home for subscribers. Messaging, events, gated content. No Discord chaos.
TypeScript ★ 1 1mo agoExplain → -
cancely
Cancel every subscription you forgot. Connect your card. See the 23 things you're paying for. Cancel any in one tap. Keep what you love.
TypeScript ★ 1 1mo agoExplain → -
breathly
Five minutes. Regulate your nervous system. Guided breathwork for anxiety, sleep, and focus. No subscription paywall on the first week.
TypeScript ★ 1 1mo agoExplain → -
boxit
The receipts. All of them. Sorted. Email them, photograph them, forward them. Boxit files each one under the right category — ready for
TypeScript ★ 1 1mo agoExplain → -
applybot
One form. Thousand applications. — Fill out your profile once. Our AI finds and applies to every job that fits, with tailored cover let
TypeScript ★ 1 1mo agoExplain → -
answeroo
The AI receptionist that never sleeps. — Dentists, salons, law firms, trades. An AI that answers every call, books appointments, and sends re
TypeScript ★ 1 1mo agoExplain → -
animeify
Turn any photo into anime. One tap. Six seconds. Ready to post. Studios Ghibli, Shinkai, shonen — pick your vibe.
TypeScript ★ 1 1mo agoExplain → -
petrait
Your dog. As Napoleon. AI portraits of your pet in any style — royal, Renaissance, astronaut, wizard. Made for your wall.
TypeScript ★ 1 1mo agoExplain → -
pricematch
Stop undercharging. Freelancers, we see you. Tell us the project, the city, the seniority. We tell you what to charge — backed by real industry r
TypeScript ★ 1 7d agoExplain → -
markdown-strip
Strip Markdown formatting (headers, bold, italic, links, code, blockquotes) to plain text. Conservative, fast, zero deps.
Rust ★ 1 1mo agoExplain → -
aws-cdk-guide ⑂
User guide for the AWS Cloud Development Kit (CDK).
★ 0 2mo agoExplain → -
agenttap-rs
Wire-level prompt introspection for LLM SDK calls. See exactly what was sent, with credentials redacted by default.
Rust ★ 0 1mo agoExplain → -
js-recon ⑂
Automate JS analysis on modern JS apps
★ 0 2mo agoExplain → -
agent-turn-limit
No description.
Python ★ 0 1mo agoExplain → -
trace-field-normalize
No description.
Python ★ 0 1mo agoExplain → -
tool-error-classify-rs
No description.
Rust ★ 0 7d agoExplain → -
agent-health-check-rs
Health status tracker for LLM agent components
Rust ★ 0 1mo agoExplain → -
llm-response-cache-rs
In-memory LRU cache for LLM responses keyed by request hash
Rust ★ 0 1mo agoExplain → -
llm-mock-rs
Deterministic mock LLM for testing agent pipelines
Rust ★ 0 1mo agoExplain → -
agent-event-log-rs
No description.
Rust ★ 0 1mo agoExplain → -
agent-epoch-counter-rs
No description.
Rust ★ 0 1mo agoExplain → -
secret-mask
Mask known secret patterns (API keys, JWTs, AWS access keys, GitHub tokens) in log lines before they reach stdout/files/sinks. Zero deps.
Rust ★ 0 1mo agoExplain → -
rate-limit-class
Parse 429 / rate-limit responses from Anthropic, OpenAI, Google Gemini, and AWS Bedrock into a unified shape: retry_after, kind (RPM/TPM/concurrent), provider.
Rust ★ 0 1mo agoExplain → -
bedrock-production-stack
Landing repo for the bedrockcache + bedrockstack + ragvitals Python libraries: production-grade AWS Bedrock + Anthropic Claude.
★ 0 7d agoExplain → -
hermes-safety-rig
Drop-in safety layer for Hermes Agent: arg validation + egress allowlist + budget cap + structured output. Built for the Hermes Agent Challenge on dev.to.
Python ★ 0 1mo agoExplain → -
agentidemp-py
Python port of agentidemp-rs: idempotency key helpers for LLM and agent calls. sha256-hex, UUIDv5, scoped helpers.
Python ★ 0 7d agoExplain → -
tool-output-truncate-py
Python port of tool-output-truncate: truncate LLM tool output with head/tail/middle/middle_lines strategies. UTF-8 safe, zero deps.
Python ★ 0 7d agoExplain → -
prompt-cache-warmer
Pre-warm Anthropic prompt cache before user traffic. Pairs with cachebench.
Python ★ 0 1mo agoExplain → -
llm-fallback-router
Multi-provider failover for LLM calls. Anthropic to OpenAI to Gemini on retryable errors.
Python ★ 0 1mo agoExplain → -
anthropic-batch-kit
Thin Anthropic Message Batches helper. Submit, poll, retrieve, with cost tally.
Python ★ 0 1mo agoExplain → -
tool-loop-guard
Detect when an LLM agent gets stuck calling the same tool with the same args. Sliding-window loop detector. Zero deps.
Python ★ 0 1mo agoExplain → -
prompt-template-version
Semver-pin LLM system prompts. Register name+version+content, retrieve by pin, persist to JSON. Zero deps.
Python ★ 0 1mo agoExplain → -
agent-deadline
Cooperative per-task deadline primitive for agent workflows. check_or_raise, remaining_seconds, intersect for nested ops. Zero deps.
Python ★ 0 8d agoExplain → -
prompt-token-counter
Approximate token counts for LLM messages, system prompts, and tools. chars/4 by default, BYO tokenizer callable. Zero deps.
Python ★ 0 1mo agoExplain → -
tool-result-cache
Content-addressable LRU cache for LLM agent tool calls. Hash(tool_name, args) → cached result. TTL + LRU eviction. Zero deps.
Python ★ 0 1mo agoExplain → -
tool-schema-from-fn
Generate Anthropic / OpenAI tool schemas from a Python function signature plus a Google-style docstring. Zero deps.
Python ★ 0 1mo agoExplain → -
agent-fn-registry
One registry per LLM agent tool: pair a callable with its schema, side-effect tags, and default args. Dispatch by name; build Anthropic/OpenAI tools list. Zero deps.
Python ★ 0 1mo agoExplain → -
llm-tool-timeout-py
Sync and async timeout enforcement for LLM agent tool functions — zero dependencies
Python ★ 0 7d agoExplain → -
tool-timeout-wrap
No description.
Python ★ 0 8d agoExplain → -
llm-message-validator-py
Validate LLM conversation messages before sending — Anthropic and OpenAI rules, zero dependencies
Python ★ 0 8d agoExplain → -
agent-config-loader
No description.
Python ★ 0 7d agoExplain → -
agent-scratchpad-py
No description.
Python ★ 0 7d agoExplain → -
agent-run-context
No description.
Python ★ 0 7d agoExplain → -
prompt-assembly-py
No description.
Python ★ 0 7d agoExplain → -
agent-memory-store-py
No description.
Python ★ 0 7d agoExplain → -
agent-step-logger
No description.
Python ★ 0 7d agoExplain → -
llm-response-parser
No description.
Python ★ 0 7d agoExplain → -
llm-conversation-history
No description.
Python ★ 0 8d agoExplain → -
tool-call-rate-limiter
No description.
Python ★ 0 7d agoExplain → -
agent-work-queue
No description.
Python ★ 0 8d agoExplain → -
agent-tool-mock
No description.
Python ★ 0 8d agoExplain → -
agent-thought-chain
No description.
Python ★ 0 7d agoExplain → -
llm-system-prompt-builder
No description.
Python ★ 0 8d agoExplain → -
llm-few-shot-store
No description.
Python ★ 0 7d agoExplain → -
agent-observation-buffer
No description.
Python ★ 0 7d agoExplain → -
agent-handoff
No description.
Python ★ 0 7d agoExplain → -
agent-goal-tracker
No description.
Python ★ 0 7d agoExplain → -
agent-persona
No description.
Python ★ 0 7d agoExplain → -
agent-run-tags
Tag and filter agent runs with arbitrary string labels
Python ★ 0 8d agoExplain → -
agent-state-machine
No description.
Python ★ 0 7d agoExplain → -
llm-response-metadata
Extract usage, model, and stop-reason metadata from LLM API responses
Python ★ 0 7d agoExplain → -
agent-input-sanitizer
Sanitize user input before sending to an LLM
Python ★ 0 7d agoExplain → -
llm-conv-stats
Per-turn conversation statistics for LLM applications
Python ★ 0 8d agoExplain → -
agent-run-stats
No description.
Python ★ 0 7d agoExplain → -
agent-output-schema
No description.
Python ★ 0 7d agoExplain → -
llm-provider-config
No description.
Python ★ 0 7d agoExplain → -
agent-conversation-state
No description.
Python ★ 0 8d agoExplain → -
agent-progress-tracker
No description.
Python ★ 0 7d agoExplain → -
agent-plan-executor
No description.
Python ★ 0 7d agoExplain → -
llm-prompt-scaffold
No description.
Python ★ 0 7d agoExplain → -
llm-tool-schema-validator
No description.
Python ★ 0 7d agoExplain → -
tool-call-middleware
No description.
Python ★ 0 7d agoExplain → -
agent-message-filter
No description.
Python ★ 0 7d agoExplain → -
agent-run-ledger
No description.
Python ★ 0 7d agoExplain → -
tool-call-audit
No description.
Python ★ 0 7d agoExplain → -
llm-tool-call-extractor
No description.
Python ★ 0 7d agoExplain → -
agent-tool-trace
No description.
Python ★ 0 8d agoExplain → -
llm-tool-manifest
No description.
Python ★ 0 7d agoExplain → -
agent-context-trim
No description.
Python ★ 0 7d agoExplain → -
llm-conversation-exporter
Export LLM conversations to markdown, HTML, plain text, and JSON
Python ★ 0 8d agoExplain → -
agent-capability-registry
Registry of named agent capabilities with enable/disable, tags, and prerequisites
Python ★ 0 7d agoExplain → -
agent-slot-filler
No description.
Python ★ 0 7d agoExplain → -
agent-prompt-builder
Composable system prompt builder for LLM agents
Python ★ 0 7d agoExplain → -
agent-run-plan
Lightweight structured plan with step status tracking for agent runs
Python ★ 0 8d agoExplain → -
agent-prompt-sections
No description.
Python ★ 0 7d agoExplain → -
agent-cost-tracker
No description.
Python ★ 0 8d agoExplain → -
agent-turn-builder
No description.
Python ★ 0 8d agoExplain → -
agent-tool-registry
No description.
Python ★ 0 7d agoExplain → -
pii-sentry-py
Python port of @mukundakatta/pii-sentry: detect and redact PII and secret-like values.
Python ★ 0 1mo agoExplain → -
consent-redaction-log-py
Record consent-aware redactions for privacy review trails. Python port of @mukundakatta/consent-redaction-log.
Python ★ 0 1mo agoExplain → -
eval-dataset-smith-py
Generate balanced eval cases from bugs, docs, examples, and policies. Python port of @mukundakatta/eval-dataset-smith.
Python ★ 0 1mo agoExplain → -
vector-poison-score-py
Python port of @mukundakatta/vector-poison-score: detects vector-text mismatch, zero/NaN vectors, instruction-like payloads, link farms.
Python ★ 0 1mo agoExplain → -
kavach-py
Threat-scoring library for AI-app security monitoring. Python port of @mukundakatta/kavach.
Python ★ 0 1mo agoExplain → -
agenttap
Wire-level prompt introspection for LLM SDK calls. See exactly what was sent, with credentials redacted by default. Anthropic, OpenAI, any httpx-based client.
Python ★ 0 7d agoExplain → -
cachebench-rs
Prompt-cache observability for LLM APIs. Per-call hit ratio, cost saved, regression alerts. Anthropic, OpenAI, Bedrock.
Rust ★ 0 1mo agoExplain → -
cost-meter
Aggregate LLM API cost across providers, models, and time windows. Provider-agnostic. Zero deps.
Rust ★ 0 1mo agoExplain → -
llmfleet-rs
Fleet-level batch dispatcher for LLM APIs. Pool requests across tasks, route to provider Batch APIs, save 50% on cost without rewriting your agent loops.
Rust ★ 0 1mo agoExplain → -
llm-error-class
Classify LLM provider error responses (rate-limit, auth, server, context-window, content-policy). Anthropic, OpenAI, Google, AWS Bedrock. Zero deps.
Rust ★ 0 1mo agoExplain → -
token-budget-pool
Shared token + dollar budget across concurrent LLM tasks. Thread-safe, zero deps.
Rust ★ 0 1mo agoExplain → -
agentidemp-rs
Idempotency keys for LLM agent retries. Deterministic content-derived keys (UUIDv5 or sha256-hex) so retries dedupe at the provider.
Rust ★ 0 7d agoExplain → -
lru-tokens
LRU cache where eviction is weighted by token count, not entry count. Bound a prompt cache by tokens (or any other size unit) instead of N entries. Zero deps.
Rust ★ 0 1mo agoExplain → -
prompt-hash
Deterministic cache key for an LLM prompt: normalize whitespace, hash messages, mix in model + temperature. Pairs with semantic-cache-key. Zero deps.
Rust ★ 0 1mo agoExplain → -
token-budget-py
Thread-safe shared token + USD budget for concurrent LLM tasks. Reserve/commit two-phase API for fan-out workloads. Sibling to the Rust crate token-budget-pool. Zero deps.
Python ★ 0 7d agoExplain → -
embed-key
Deterministic cache key for an embedding request: hash text + mix in provider, model, and dimensionality. So a cache survives model upgrades without false hits. Zero deps.
Rust ★ 0 1mo agoExplain → -
ragdrift-arize-bridge
Phoenix-compatible exporter: ragdrift's five drift scalars as OpenTelemetry span attributes for Arize alerting.
Python ★ 0 1mo agoExplain → -
regex-pii-rs
Regex-only PII detector for emails, phones, SSNs, credit cards, and prefixed API keys. Rust port of pii-sentry. Zero deps.
Rust ★ 0 1mo agoExplain → -
hermes-budget-skin
Drop-in USD caps + egress allowlist + audit log for the Hermes Agent runtime.
Python ★ 0 1mo agoExplain → -
step-id
Stable IDs for agent steps: deterministic hash of (run_id, step_index, kind) so events and traces share keys across reruns. Zero deps.
Rust ★ 0 1mo agoExplain → -
llm-fallback-chain
Tiny provider failover chain for LLM calls. BYO provider callables, async + sync, custom skip predicate, attempt trace.
Python ★ 0 7d agoExplain → -
llm-message-hash-py
Python port of llm-message-hash: canonical sha256 hash of LLM request structures. Per-provider presets drop noise fields. For cache keys and idempotency.
Python ★ 0 7d agoExplain → -
agent-redact
Zero-dep PII and secret scrubber for AI agent logs, traces, and audit output.
Python ★ 0 1mo agoExplain → -
llm-cache-mem
In-process LRU cache for LLM responses. TTL, sync + async, decorator + direct API, thread-safe. Zero deps.
Python ★ 0 7d agoExplain → -
tool-error-classify
Classify exceptions from LLM agent tools into a closed ErrorKind enum (USER_INPUT/AUTH/RATE_LIMITED/...) with LLM-friendly hints. Zero deps.
Python ★ 0 1mo agoExplain → -
tool-arg-coerce-py
Python port of tool-arg-coerce. Coerce LLM-generated tool args to JSON Schema types. 'true'→True, '5'→5. Zero deps.
Python ★ 0 1mo agoExplain → -
agentprompt-py
LLM prompt templates with Jinja2 syntax. Role-aware Messages builder; output ready for Anthropic / OpenAI SDKs. Python port of agentprompt-rs.
Python ★ 0 7d agoExplain → -
llm-tool-arg-default
Fill missing tool args from schema defaults before validation. Dict schema or function signature. Nested object + array recursion. Zero deps.
Python ★ 0 7d agoExplain → -
gemini-cost-py
USD cost calculator for Google Gemini API calls. Long-context tiers, thinking tokens, grounding surcharge. Zero deps. Sibling to claude-cost-py + openai-cost-py.
Python ★ 0 7d agoExplain → -
openai-cost-py
Cache-aware USD cost calculator for OpenAI API calls. Batch API discount, prompt-cache discount, alias resolution. Zero deps. Sibling to claude-cost-py.
Python ★ 0 8d agoExplain → -
claude-cost-py
Compute the USD cost of a Claude API call from its usage block. Cache-aware, Bedrock-aware (versioned IDs, inference profiles, ARNs), zero dependencies. Python port of claude-cost.
Python ★ 0 8d agoExplain → -
agent-rate-limiter-py
Per-tool call-rate enforcement for LLM agents — sliding window, thread-safe, zero dependencies
Python ★ 0 7d agoExplain → -
agent-output-parser-py
Extract JSON, code blocks, lists, and XML tags from LLM text responses — zero dependencies
Python ★ 0 7d agoExplain → -
agent-task-graph
No description.
Python ★ 0 7d agoExplain → -
llm-token-split
No description.
Python ★ 0 7d agoExplain → -
llm-role-validator
No description.
Python ★ 0 8d agoExplain → -
agent-budget-tracker
Per-run budget tracking for tokens, cost, and API calls
Python ★ 0 7d agoExplain → -
llm-context-assembler
No description.
Python ★ 0 7d agoExplain → -
agent-context-snapshot
Point-in-time snapshots of agent conversation state for debugging and checkpointing
Python ★ 0 7d agoExplain → -
llmfleet
Fleet-level batch dispatcher for LLM APIs. Pool requests across coroutines, route to provider Batch APIs, save 50% on cost without rewriting your agent loops.
Python ★ 0 7d agoExplain → -
llm-model-picker
No description.
Python ★ 0 7d agoExplain → -
agent-subtask-manager
Parent/child task hierarchy with status tracking and completion rollup
Python ★ 0 8d agoExplain → -
cachebench
Prompt-cache observability for LLM APIs. Per-call hit ratios, cost saved, regression alerts, miss-aware retry. Anthropic + OpenAI + Bedrock.
Python ★ 0 7d agoExplain → -
agentguard-rs
Network egress firewall for AI agent tools. Declarative domain allowlist; throws on violation. Optional reqwest-middleware integration.
Rust ★ 0 7d agoExplain → -
agentprompt-rs
LLM prompt templates with Jinja2 syntax. Render system/user/assistant turns into a typed message list, ready for the Anthropic or OpenAI SDK.
Rust ★ 0 7d agoExplain → -
promptver
Hash and version prompt templates so eval results, cache keys, and audit logs stay stable when templates change. Whitespace-normalized SHA-256. Zero deps.
Rust ★ 0 1mo agoExplain → -
agenttrace-rs
Cost + latency aggregation for LLM agent runs. Group calls into named runs, get totals, p50/p95, and per-model breakdowns. Composes with cachebench.
Rust ★ 0 8d agoExplain → -
agentsnap-rs
Snapshot tests for AI agent traces. Record once, replay-and-compare on every run; the agent equivalent of Jest snapshots.
Rust ★ 0 7d agoExplain → -
schema-shrink
Analyze and simplify JSON Schemas that exceed Anthropic strict-mode compiled-grammar limits. Spots nullable unions, large enums, deep nesting, and applies documented workarounds.
Rust ★ 0 1mo agoExplain → -
llm-think-tag-strip
Strip <thinking>/<think> reasoning tags from LLM output. Returns clean text + extracted thinking. Configurable tag set, markdown-style support. Zero deps.
Python ★ 0 7d agoExplain → -
llm-batch-coalesce
Request coalescing / single-flight for LLM calls. Many concurrent callers same key, one underlying call. asyncio + threading. Zero deps.
Python ★ 0 7d agoExplain →
No repos match these filters.