Why Qwen3 VL 235B A22B Instruct pressures system RAM
Qwen3 VL 235B A22B Instruct is Mixture-of-Experts: inference activates 22B active/token, but VRAM/RAM must usually hold the full ~235B expert set for fast routing. At Q4 the weight slab is ~132.2GB before KV (~0.05GB at 8K) and ~8GB OS/runtime overhead — totaling ~140.3GB raw, rounded to a 192GB kit. Stretching toward the full 262K-token window multiplies KV far faster than weights; that is the usual “I bought enough RAM for the model but still OOM” failure on Alibaba Qwen MoE pages.
What RAM kit to buy
Shop 192GB-class capacity for Qwen3 VL 235B A22B Instruct: workstation DDR5 RDIMM/LRDIMM or multi-kit desktop builds, not a single gamer 2×16GB stick. Use our 128GB+ price hubs and RAM Finder; confirm ECC needs for your board. GPU path: Apple Mac Studio (192GB Unified Memory) or Institutional Node (8x H100 / A100) (144.2GB VRAM class) if you want weights on-device instead of system-RAM offload.
Workload notes
Qwen-family models like Qwen3 VL 235B A22B Instruct often ship strong coding/agent variants; leave RAM for tool runners and browser IDEs beside the weights. At 235B, Qwen3 VL 235B A22B Instruct sits in the large local-LLM band: Q4 on a strong GPU is realistic, FP16 usually is not on consumer cards. Release window noted as 2025/2026; always re-check the model card before buying hardware for a specific checkpoint.




