Issue 2026-10-03 · Industry · 开源 · 应用

Low-VRAM users hope for Qwen4 35B and similar models

A r/LocalLLaMA user joked that owners of 8GB VRAM plus 16GB RAM check the subreddit daily hoping for new models such as Qwen4 35B A3B, while still relying on Gemma 26B QAT at about 26 tokens per second.

r/LocalLLaMA3 d ago
Read original ↗