2-day current streak·3-day longest streak
-
spark-bench
Mixed-capability LLM benchmark for DGX Spark — 57 scenarios, 10 domains, partial-credit grading, trial statistics
HTML ★ 136 12d agoExplain → -
DeepSeek-v4-Flash-DSpark-1M-NVFP4-KV-2x-DGX-Spark
DeepSeek V4 Flash DSpark on 2x DGX Spark — NVFP4 KV cache, 1M context, speculative decoding. Based on MiaAI-Lab dual-Spark packaging.
Python ★ 6 21d agoExplain → -
Laguna-S-2.1-NVFP4-1x-DGX-Spark
poolside Laguna S-2.1 (118B/8B MoE) on ONE NVIDIA DGX Spark - validated NVFP4 recipe + day-0 benchmark: TrueScore 86.5, agentic 99.3, 19 tok/s single / 84 tok/s @ c8. Needs vLLM >= 0.25 (older stacks emit gibberish with tools).
★ 1 2d agoExplain → -
qwen-sglang-dgx-spark
Deploy Qwen3.6-35B on a DGX Spark (GB10) with SGLang v0.5.15 for long-context multi-agent serving. Includes a reproducible SGLang-vs-vLLM comparison.
Shell ★ 1 11d agoExplain → -
atlas ⑂
Pure Rust Inference Engine
★ 0 9m agoExplain → -
Qwopus3.6-27B-Q4_K_M-DGX-Spark
Qwopus 27B (Q4_K_M) on NVIDIA DGX Spark (GB10) via llama.cpp — MTP speculative decoding, 49-scenario TrueScore benchmark
Shell ★ 0 1mo agoExplain → -
specserve
Lightweight web GUI for managing vLLM model serving with speculative decoding
Python ★ 0 2mo agoExplain →
No repos match these filters.