-
hermes-concurrent-agents
Deploy concurrent Hermes Agent workers on unified-memory GPUs (GB10, DGX Spark) for maximum total tok/s. Profile-isolated, kanban-coordinated, crash-recovering.
Python ★ 79 26d agoExplain → -
hermes-buzz-shared-profile
macOS Hermes skill for sharing one canonical writable profile across Buzz and ACP surfaces
Python ★ 50 14d agoExplain → -
llm-wiki_obsidian_hermes_r0b0tlabbra1n
Filesystem-first LLM-Wiki + Obsidian + Hermes Agent memory system. Markdown source of truth, SQLite FTS5 search, secret scanning, tier-based memory. Built for local LLM setups.
Python ★ 30 3mo agoExplain → -
DeepSeek-V4-Flash-DSpark-v026-SM121
DeepSeek-V4-Flash-DSpark optimized vLLM 0.26.0 SM121 dual-GB10 evidence (NVFP4 KV, B12X, DSpark K6)
Python ★ 29 10d agoExplain → -
deepseek-v4-flash-nvfp4-gb10-benchmark
DeepSeek-V4-Flash native Blackwell FP8 benchmark on dual DGX Spark GB10 (TP=2, MTP, RoCE). c=1 38.4 t/s, c=16 144.6 t/s aggregate.
Python ★ 13 1mo agoExplain → -
qwen36-35b-a3b-nvfp4-gb10-native-mtp
GB10 NVFP4 native MTP reproducibility pack for Qwen3.6-35B-A3B
Python ★ 13 2mo agoExplain → -
minimax-m27-nvfp4-gb10-benchmark
MiniMax M2.7 NVFP4 dual-GB10 Blackwell benchmark: vLLM FlashInfer-CUTLASS, public data, HTML canvas report, and Docker runtime.
HTML ★ 13 2mo agoExplain → -
DeepSeek-v4-Flash-DSpark-2x-DGX-Spark ⑂
No description.
Python ★ 11 24d agoExplain → -
r0b0bench
Specification for a reproducible, provenance-bound multi-lane LLM benchmark suite.
Python ★ 10 2d agoExplain → -
laguna-s-2.1-nvfp4-sm121-vllm
Native SM121 vLLM runtime and reproducible qualification for Poolside Laguna S 2.1 NVFP4
Python ★ 9 22d agoExplain → -
hermes-ansible-node-orchestration
Hermes Agent + Ansible GB10 cluster orchestration skill and runnable operator repo
Shell ★ 8 1mo agoExplain → -
nvidia-qwen-3.6-27B-sm121-nvfp4
NVIDIA Qwen3.6-27B NVFP4 on SM121 (GB10) — vLLM v0.24.0 with native NVFP4 KV cache via FlashInfer FA2 JIT. 67% more KV capacity than FP8.
Python ★ 8 25d agoExplain → -
tldraw-skill
Production-grade Hermes Agent skill for tldraw 5.2.5+ development, migration, sync, automation, and evaluation
TypeScript ★ 7 19d agoExplain → -
qwen36-35b-a3b-nvfp4-sm121-vllm
Native SM121 vLLM deployment and benchmark evidence for NVIDIA Qwen3.6-35B-A3B-NVFP4 on one GB10
Python ★ 7 29d agoExplain → -
qwen36-35b-a3b-nvfp4-fast-sm121-vllm
SM121-native Qwen3.6-35B-A3B NVFP4 Fast serving on NVIDIA GB10 with vLLM, FlashInfer B12X, FP8 KV, and MTP
HTML ★ 6 1mo agoExplain → -
gemma4-26b-a4b-nvfp4-gb10-native-cutlass
Gemma-4-26B-A4B-NVFP4 native CUTLASS profile on GB10 / Blackwell
Shell ★ 6 2mo agoExplain → -
nemotron-3.5-lightning-sm121-nvfp4
NVIDIA Nemotron 3.5 Lightning SM121 (GB10) reproducibility suite: NVFP4 MTP K=1 runtime, r0b0bench evidence, 1M-window NIAH probe
Python ★ 5 2d agoExplain → -
step37-flash-nvfp4-sm121-vllm-docker
Step 3.7 Flash NVFP4 SM121-verified TP=2 on dual GB10 | vLLM + Ray
HTML ★ 4 2mo agoExplain → -
gemma4-31b-it-nvfp4-gb10
GB10 NVFP4 reproducibility pack for Gemma-4-31B-IT
Python ★ 4 2mo agoExplain → -
diffusiongemma-26b-nvfp4-sm121-vllm
Optimized SM121 vLLM container and benchmark report for nvidia/diffusiongemma-26B-A4B-it-NVFP4
HTML ★ 3 2mo agoExplain → -
nex-n2-mini-nvfp4
NVFP4 quantized Nex-N2-mini (Qwen3.5-MoE-35B) — vLLM serving container for Blackwell SM121
Shell ★ 3 2mo agoExplain → -
nemotron3-super-120b-a12b-nvfp4-gb10-native-mtp
GB10 NVFP4 native MTP reproducibility pack for Nemotron-3-Super-120B-A12B
HTML ★ 3 2mo agoExplain → -
FastContext-1.0-4B-RL-NVFP4
NVFP4 FastContext weights, serve configs, and with-vs-without demo (standalone)
Python ★ 2 1mo agoExplain → -
gb10-comfyui-hermes-r0b0tlab
Agent Skills compliant skill for deploying ComfyUI on NVIDIA GB10 / DGX Spark with Hermes Agent integration. Optimized for sm_121 Blackwell, ARM64, 128GB unified memory.
★ 1 3mo agoExplain → -
hy3-295b-nvfp4-gb10-benchmark
Hy3-295B W4A4 NVFP4 v3 dual-GB10 benchmark and reproducibility evidence
Python ★ 1 29d agoExplain → -
fastcontext-hermes
FastContext-Hermes: Async subagent repo exploration on NVIDIA GB10
Python ★ 1 1mo agoExplain → -
vibethinker-3b-nvfp4
VibeThinker-3B NVFP4: SM121-optimized container for 3B reasoning model (2.60x throughput)
HTML ★ 1 1mo agoExplain → -
unsloth-qwen36-35b-a3b-nvfp4-ar-gb10
Sanitized GB10 AR benchmark results for unsloth/Qwen3.6-35B-A3B-NVFP4
Python ★ 1 2mo agoExplain → -
nemotron-3.5-lightning-sglang-sm121
SGLang SM121 reproducibility suite for NVIDIA Nemotron 3.5 Lightning NVFP4
Python ★ 0 1d agoExplain → -
ling-3.0-flash-nvfp4
No description.
Python ★ 0 6d agoExplain → -
inkling-small-nvfp4-gb10-sglang
Native SGLang serving for Inkling-Small NVFP4 on two NVIDIA GB10/SM121 nodes
Python ★ 0 6d agoExplain → -
xyz-aquila-mini-nvfp4-sm121-vllm
SM121 vLLM 0.25 runtime for XYZ-Aquila-mini NVFP4
Python ★ 0 16d agoExplain → -
hermes-model-hub
One dependable OpenAI-compatible endpoint for multiple explicit model pools
Python ★ 0 24d agoExplain → -
qwen3.8-max-dossier
Qwen3.8-Max launch-day dossier — a self-portrait HTML artifact built by the model it documents. Official Qwen sources only.
HTML ★ 0 25d agoExplain → -
hermes-agent ⑂
The agent that grows with you
★ 0 25d agoExplain → -
hermes-alibaba-token-plan ⑂
Hermes Agent model-provider plugin for Alibaba Cloud Token Plan (Global + China)
★ 0 25d agoExplain → -
dgx-spark-playbooks ⑂
Collection of step-by-step playbooks for setting up AI/ML workloads on NVIDIA DGX Spark devices with Blackwell architecture.
★ 0 25d agoExplain → -
reachy_mini ⑂
Reachy Mini's SDK
★ 0 25d agoExplain → -
reachy-mini-desktop-app ⑂
No description.
★ 0 25d agoExplain → -
vllm-v0250-cu130-sm121
vLLM v0.25.0 source-built for CUDA 13.0, ARM64, and NVIDIA GB10 SM121
Dockerfile ★ 0 1mo agoExplain → -
laguna-xs2-nvfp4-sm121-vllm
Laguna XS 2.1 NVFP4 (33B MoE, 3B active) on GB10 SM121 — vLLM v0.24.0 native NVFP4 serving + DFlash Phase 2
Python ★ 0 1mo agoExplain → -
agents-a1-nvfp4-sm121-vllm
SM121-native vLLM container for Agents-A1 NVFP4
Python ★ 0 1mo agoExplain → -
openhome-ability-authoring
Agent Skill for building OpenHome Abilities
Python ★ 0 1mo agoExplain → -
laguna-m1-nvfp4-sm121-vllm
Laguna M.1 NVFP4 on SM121 — dual GB10 TP=2 vLLM serve, concurrency benchmarks, GSM8K, Hermes agent micros
Python ★ 0 1mo agoExplain → -
deepseek-v4-flash-nvfp4-sm121
DeepSeek-V4-Flash NVFP4 deployment on dual GB10 (SM121) — vLLM + Ray multi-node
HTML ★ 0 2mo agoExplain → -
agentmail-Hermes-skill
AgentSkills.io compliant Hermes skill for AgentMail — give AI agents their own email inbox
Python ★ 0 3mo agoExplain →
No repos match these filters.