🎯
Popular repositories Loading
-
DeepSeek-v4-Flash-DSpark-2x-DGX-Spark
DeepSeek-v4-Flash-DSpark-2x-DGX-Spark PublicDeepSeek-v4-Flash 0731 recipe for 2x DGX Sparks
-
DeepSeek-V4-Flash-Dual-DGX-Spark-1M-Context
DeepSeek-V4-Flash-Dual-DGX-Spark-1M-Context PublicDeploy DeepSeek V4 Flash (MoE reasoning model) on dual DGX Spark nodes with 1M token context, InfiniBand, and FP8 KV-cache
-
Laguna-S-2.1-DGX-Spark-RTX-6000-PRO
Laguna-S-2.1-DGX-Spark-RTX-6000-PRO PublicvLLM 0.25.1 serving stack for poolside/Laguna-S-2.1-NVFP4 with DFlash speculative decoding — DGX Spark & RTX 6000 PRO
-
GLM-5.2-NVFP4-AQLM-Triple-DGX-Sparks
GLM-5.2-NVFP4-AQLM-Triple-DGX-Sparks PublicGLM-5.2 NVFP4+AQLM on 3× DGX Spark — 380k context MTP serve stack
-
Qwen3.6-27B-NVFP4-vLLM
Qwen3.6-27B-NVFP4-vLLM PublicProduction-ready vLLM deployment wrapper for Qwen3.6-27B (NVFP4) — self-hosted OpenAI-compatible inference
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.


