Benchmarking Qwen 3.6 35B MoE (3B active) on an RTX 3090
Qwen3.6-35B-A3BRTX 3090VulkanCUDALlama.cppperformancebenchmarkingAImachine learninglanguage models.
Author: gpjt
Date: 7/24/2026
Article Summary:
A detailed analysis of the performance of the Qwen3.6-35B-A3B model on an RTX 3090 GPU, comparing the Vulkan and CUDA versions of Llama.cpp.