Popular repositories Loading
-
deepseek-v4-flash-a100
deepseek-v4-flash-a100 PublicSource-only offline deployment, testing, and operations toolkit for DeepSeek-V4-Flash-0731 on 4x/8x NVIDIA A100 GPUs, pinned to a reviewed community vLLM R1 stack.
HTML 2
-
single-dgx-spark-gb10-llm
single-dgx-spark-gb10-llm PublicDual Qwen NVFP4 deployment, exact KV budgeting, MTP/DFlash2 acceleration, telemetry, and benchmarks for NVIDIA DGX Spark GB10
HTML
-
Qwen3.8-Flash-Next-NVFP4-RTX-PRO-6000-Single
Qwen3.8-Flash-Next-NVFP4-RTX-PRO-6000-Single PublicQwen3.8 Flash-Next NVFP4 on one RTX PRO 6000: Pennyroyal SGLang, native NEXTN MTP, 256K context, C8, New API and measured benchmarks
TypeScript
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.