1 Post
Tag: Quantization

BY eric
Sep 19, 2026
Now It Remembers: A 27B Model With a 262K-Token Context on the Same RTX 3060
The fifth in our RTX 3060 series. The same Qwen3.8-27B we ran last month, re-quantized to under 6GB, now runs entirely on the 12GB card, about four times faster, and holds up to 262,000 tokens of context. A whole novel, on a five-year-old gaming card.
