Skip to content
View keplerzip's full-sized avatar
  • 06:16 (UTC +08:00)

Block or report keplerzip

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Popular repositories Loading

  1. llm_wiki llm_wiki Public

    2

  2. deepseek-v4-flash-a100 deepseek-v4-flash-a100 Public

    Source-only offline deployment, testing, and operations toolkit for DeepSeek-V4-Flash-0731 on 4x/8x NVIDIA A100 GPUs, pinned to a reviewed community vLLM R1 stack.

    HTML 2

  3. single-dgx-spark-gb10-llm single-dgx-spark-gb10-llm Public

    Dual Qwen NVFP4 deployment, exact KV budgeting, MTP/DFlash2 acceleration, telemetry, and benchmarks for NVIDIA DGX Spark GB10

    HTML

  4. Qwen3.8-Flash-Next-NVFP4-RTX-PRO-6000-Single Qwen3.8-Flash-Next-NVFP4-RTX-PRO-6000-Single Public

    Qwen3.8 Flash-Next NVFP4 on one RTX PRO 6000: Pennyroyal SGLang, native NEXTN MTP, 256K context, C8, New API and measured benchmarks

    TypeScript