Popular repositories Loading
-
DeepSeek-v4-DSpark-Aidendle94-GB10-ServingStack
DeepSeek-v4-DSpark-Aidendle94-GB10-ServingStack PublicDocker compose serving stack for DeepSeek v4 Flash DSpark for NVIDIA Spark GB10 system using Aidendle94 image
-
glm-5.3-flash-nvidia-nvfp4-dflash-2x-dgx-sparks
glm-5.3-flash-nvidia-nvfp4-dflash-2x-dgx-sparks PublicGLM-5.3-Flash NVIDIA official NVFP4 + DFlash2 on 2x DGX Spark — TP=2 over RoCE, cudagraphs + async scheduling, verified 700K ctx (hardmode 94/100), one .env
-
glm-5.3-flash-nvfp4-2x-dgx-sparks
glm-5.3-flash-nvfp4-2x-dgx-sparks PublicGLM-5.3-Flash NVFP4 (lab quant) on 2x DGX Spark — TP=2 over RoCE, one .env, compose head/worker, download/start/stop/status/tail-log
-
EngineMetrics
EngineMetrics PublicReal-time vLLM + TensorFold performance dashboard TUI — single file, zero dependencies
-
MediaLLMProxy
MediaLLMProxy PublicOpenAI-compatible local media bridge and model-scoped reasoning/structured-output compatibility proxy.
-
qwen3.8-flash-next-2x-dgx-sparks
qwen3.8-flash-next-2x-dgx-sparks PublicGeneralized Qwen3.8-Flash-Next FP8 TP2 serving stack for 2x DGX Spark (RoCE) — one .env, compose head/worker, start/stop/status/download
If the problem persists, check the GitHub status page or contact support.