"AI Disruption" Publication 5900 Subscriptions 20% Discount Offer Link.
Just one month after the launch of Google’s Gemma 3, a new version has already been released.
This version has been optimized with Quantization-Aware Training (QAT), significantly reducing memory requirements while maintaining high quality.
For example, after QAT optimization, the VRAM usage of Gemma 3 27B can be drastically reduced from 54GB to 14.1GB, making it fully capable of running locally on consumer-grade GPUs like the NVIDIA RTX 3090!



