operational
Pinned Loading
-
dualmind-plus-one
dualmind-plus-one PublicDualMind MoshiPlex Edition — Autonomous AI Dialogue Showcase (Full-duplex Audio-to-Audio AI Conversational System based on PersonaPlex and Moshi)
-
comfyui-v100-sxm2-minimax-h3
comfyui-v100-sxm2-minimax-h3 PublicMulti-GPU acceleration for MiniMax H3 video generation on NVIDIA V100 (sm_70). Ulysses sequence parallelism as a drop-in ComfyUI custom node — ~19 min to ~7 min on 8x V100.
-
sglang-v100-sxm2-qwen3.8-flash-next
sglang-v100-sxm2-qwen3.8-flash-next PublicQwen3.8-Flash-Next (125B MoE, NVFP4) on 4x V100-SXM2-32GB and DeepSeek-V4.1-flash on 8x V100 — a Volta port of SGLang for agentic coding.
-
vllm-v100-sxm2-qwen3.5-397b
vllm-v100-sxm2-qwen3.5-397b PublicServe Qwen3.5-397B-A17B (AWQ) on 8x Tesla V100-SXM2-32GB (DGX-1, TP8) for agentic coding & ops — a downstream fork of 1Cat-vLLM.
Python 1
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.
