🎯
Popular repositories Loading
-
-
DeepSeek-V4-Flash-Dual-DGX-Spark-1M-Context
DeepSeek-V4-Flash-Dual-DGX-Spark-1M-Context PublicDeploy DeepSeek V4 Flash (MoE reasoning model) on dual DGX Spark nodes with 1M token context, InfiniBand, and FP8 KV-cache
-
Laguna-S-2.1-DGX-Spark-RTX-6000-PRO
Laguna-S-2.1-DGX-Spark-RTX-6000-PRO PublicvLLM 0.25.1 serving stack for poolside/Laguna-S-2.1-NVFP4 with DFlash speculative decoding — DGX Spark & RTX 6000 PRO
-
Qwen3.6-27B-NVFP4-vLLM
Qwen3.6-27B-NVFP4-vLLM PublicProduction-ready vLLM deployment wrapper for Qwen3.6-27B (NVFP4) — self-hosted OpenAI-compatible inference
-
GLM-5.2-NVFP4-AQLM-Triple-DGX-Sparks
GLM-5.2-NVFP4-AQLM-Triple-DGX-Sparks PublicGLM-5.2 NVFP4+AQLM on 3× DGX Spark — ~248k context MTP serve stack
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.

