Nice use on the NVIDIA DGX Spark when Qwen/Qwen3.6-35B-A3B-FP8 is inferencing using vllm. Performance is VERY good.
0 likes 0 replies
?
Nice use on the NVIDIA DGX Spark when Qwen/Qwen3.6-35B-A3B-FP8 is inferencing using vllm. Performance is VERY good.
0 likes 0 replies