The V100 32GB SXM2 has about 41% more memory bandwidth (900 vs 640 GB/s). Batch-1 decode tracks bandwidth closely, so a model that fits both cards should decode roughly 41% faster on the V100 32GB SXM2. The V100 32GB SXM2 fits more of the catalog at Q4/8K — 43 vs 25 of 73 models (32 vs 16 GB VRAM). Largest model that fits the V100 32GB SXM2 but not the RX 9070 XT 16GB: Qwen3.6-35B-A3B-FP8 (36.0B). MSRP is $599 for the RX 9070 XT 16GB vs $10,000 for the V100 32GB SXM2. MSRP is a launch list price, not live retail — Amazon prices move; click through for the current price. Not retail-shoppable: V100 32GB SXM2 (datacenter-class; cloud rental is the realistic way to use it). In the Q4/8K head-to-head below (49 shared models), the V100 32GB SXM2 posts the higher tok/s in 46 rows vs 3. Methodology.