Popular repositories Loading
-
deepseek-v4-flash-a100
deepseek-v4-flash-a100 PublicSource-only offline deployment, testing, and operations toolkit for DeepSeek-V4-Flash-0731 on 4x/8x NVIDIA A100 GPUs, pinned to a reviewed community vLLM R1 stack.
HTML 2
-
deepseek-v4.1-flash-a100
deepseek-v4.1-flash-a100 PublicOffline DeepSeek-V4.1-Flash deployment for 8x A100: DSpark k5, 256K context, local APIs and measured benchmarks.
Python 1
-
single-dgx-spark-gb10-llm
single-dgx-spark-gb10-llm PublicDual Qwen NVFP4 deployment, exact KV budgeting, MTP/DFlash2 acceleration, telemetry, and benchmarks for NVIDIA DGX Spark GB10
HTML
-
Qwen3.8-Flash-Next-NVFP4-RTX-PRO-6000-Single
Qwen3.8-Flash-Next-NVFP4-RTX-PRO-6000-Single PublicQwen3.8 Flash-Next NVFP4 on one RTX PRO 6000: Pennyroyal SGLang, native NEXTN MTP, 256K context, C8, New API and measured benchmarks
TypeScript
-
deepseek-v4-flash-vision-exp-a100
deepseek-v4-flash-vision-exp-a100 PublicSource-only A100 deployment for DeepSeek V4 Flash Vision Exp: DSpark, 256K, C32, Chat/Responses/Messages and multimodal tests.
Python
If the problem persists, check the GitHub status page or contact support.