Qwen3.6 35B
unsloth/qwen3.6-35b-a3b-nvfp4-fast
unsloth/qwen3.6-35b-a3b-gguf on HuggingFace ↗
Recipe
- Profile
- unsloth-qwen3-6-35b-a3b-nvfp4-fast-eugr
- Engine
- vLLM
- Served as
- qwen3.6-35b-a3b-nvfp4-fast
Run on your Spark
single node
shell
spark inference up unsloth-qwen3-6-35b-a3b-nvfp4-fast-eugr
Why we run it
Golden fleet target — auto-scaffolded from recipe qwen36-q4-llama.
Bench notes
bench-v2 avg 55.6 decode tok/s (2 sessions, ~50k ctx fill, tool_ok=True)