Tech2wild
GLM-5.3-Flash-NVFP4-DFlash2-2x-DGX-Spark got 191 stars in 31 days
Be the first to back @Tech2Wild
Nobody has launched a Paid coin for this dev yet. Launch one and 80% of its trading fees go to @Tech2Wild in dollars on X. Takes about a minute.
- 1Copy this line for the coin's description
Fees to @Tech2Wild via UsePaidThis exact line is how Paid knows who to pay. Don't change it.
- 2Launch it on PaidOPEN PAID →ticker idea: $DEEPSEEKV4
Paste the line into the description. Paid sets 100% of creator fees to its treasury, then pays @Tech2Wild. How Paid works
- 3Come back and check
Launched it? We'll confirm with Paid and put it on the board.
- 4Tell CT
Share the card below so people find the coin (and the dev notices).
Share this scout
Post this card on X. The link unfurls into the card automatically, so your followers see the stats.
Download imageTheir projects
★ = stars, like likes from other developers. + = new stars since our last scan.
DeepSeek V4 Flash DSpark 1M NVFP4 KV recipe for 2x DGX Spark
GLM-5.3-Flash (NVFP4) on 2x NVIDIA DGX Spark - vLLM TP2, 262K context, MTP. World-first deploy recipe: 7 day-0 bugs found and fixed, patched sm121 image, probes and full report.
Qwen3.8-Flash-Next (NVFP4) on DGX Spark in vLLM: one Spark 43.9 tok/s with our disk-backed n-gram table patch, staged gather and reduced-vocab MTP draft; TP2 SPEED 53.7; TP4 CONTEXT 9M KV pool. Launchers, patch, 40-prompt harness, KV ledger, credits. SGLang TP2 day-0 lane kept.
DeepSeek-V4.1-Flash (552B MoE, MXFP4 experts, 1M ctx) on four NVIDIA DGX Sparks with vLLM TP4: Engram-on-disk patch, sm121 kernel build, launchers, measured numbers
Working recipe to serve DeepSeek-V4-Flash across two NVIDIA DGX Spark (GB10) nodes with vLLM (TP=2, FP8 KV, MTP) over a RoCE/RDMA link — Docker image, launch scripts, RDMA/NCCL setup, and the gotchas.
Recipe: GLM-5.2 (unpruned QuantTrio Int4-Int8Mix) at 200K ctx with MTP spec decode on a 4x NVIDIA DGX Spark (GB10) cluster
Run MiniMax H3 locally: a full 15s clip with audio on a single RTX 3090. The fix is one flag.
MiniMax-M3 (428B, no pruning) at 36 tok/s on 2× NVIDIA DGX Spark — W4A16 GPTQ + NVFP4 KV + EAGLE-3 speculative decoding on vLLM. Three serving lanes: speed / balanced / long-context.
Keep discovering
Also building in Other, around the same level.
A directory of hacker forums