dgx-spark-vllm-qwen3.6-35b-a3b-dflash. High-performance Qwen3.6-35B-A3B-DFlash inference on NVIDIA DGX Spark (~50 tok/s)

github.com/ZengboJamesWang/dgx-spark-vllm-qwen3.6-35b-a3b-dflash

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.