Close Menu
Tech Nova Mindset – Empower Innovation and Forward Thinking

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    NASA’s new dark energy space telescope can also detect killer asteroids

    August 5, 2026

    The AI Notetaker Has Been Invited to All the Meetings

    August 5, 2026

    OK, Well, Rogue AI Agents Are Hacking Again

    August 5, 2026
    Facebook X (Twitter) Instagram
    Trending
    • NASA’s new dark energy space telescope can also detect killer asteroids
    • The AI Notetaker Has Been Invited to All the Meetings
    • OK, Well, Rogue AI Agents Are Hacking Again
    • Heat Is an Orbital Data Center’s Greatest Foe. These Tiles Dump It at the Source.
    • The White House Is Keeping Its AI Cybersecurity Framework Secret
    • How One Startup Built a (Mostly) China-Free Robot
    • The 2026 R&D Benchmark Report: Waste, AI and the Race to Market
    • Is AI making us dumber? Maybe not. But our skills are at risk
    Tech Nova Mindset – Empower Innovation and Forward Thinking
    • Home
    • Gadgets
    • Reviews
    • Tech News
    • Future Tech
    • AI & Robotics
    • How-To Guides
    • More
      • Cybersecurity
      • Startups & Innovation
    Tech Nova Mindset – Empower Innovation and Forward Thinking
    Home»Gadgets»Gemma 4 models use a training trick to slash their memory footprint
    Gadgets

    Gemma 4 models use a training trick to slash their memory footprint

    kirklandc008@gmail.comBy kirklandc008@gmail.comJune 6, 2026No Comments2 Mins Read
    Facebook Twitter Pinterest LinkedIn Tumblr Email
    The promotional graphic for the Gemma 4 QAT models.
    Share
    Facebook Twitter LinkedIn Pinterest Email

    TL;DR

    • Gemma 4 models are now available for download with quantization-aware training (QAT), which reduces the size and memory footprint of the models.
    • These open-source models retain quality better thanks to QAT compared to those that use post-training quantization (PTQ).
    • The Gemma 4 models optimized with QAT are available in five sizes: Gemma 4 E2B, Gemma 4 E4B, Gemma 4 12B, Gemma 4 26B A4B, and Gemma 4 31B.

    Following Google’s launch of the laptop-grade Gemma 4 12B model earlier this week, the company is releasing new Gemma 4 model checkpoints with quantization-aware training. Quantization is necessary to reduce the amount of memory required to run lightweight models. The standard method is post-training quantization (PTQ), which quantizes the model after training, but could result in weaker performance. The latest Gemma 4 versions use quantization-aware training (QAT) instead to reduce model quality loss and accelerate decode speed, according to Google’s blog post.

    Google says that incorporating quantization into the training process results in checkpoints with better performance than models refined with PTQ. The compressed models run on phones and laptops well thanks to a custom mobile-quantization schema. This involves using pre-calculated settings, 2-bit compression in certain parts of the model, and vocabulary list and short-term memory compression. For the user, this results in a smaller model that consumes less system memory.

    Don’t want to miss the best from Android Authority?

    There are multiple model sizes available with QAT optimization, include Gemma 4 E2B, Gemma 4 E4B, Gemma 4 12B, Gemma 4 26B A4B, and Gemma 4 31B. The smallest versions, like the text-only Gemma 4 E2B model, require less than a gigabyte of memory to run. These small Gemma 4 checkpoints without intensive resource requirements are ideal for running on phones.

    Google shared the approximate memory requirements to load the new Gemma 4 models with QAT in various sizes:

    There are four different formats of Gemma 4 QAT models available for download: unquantized QAT checkpoints, GPT-Generated Unified Format (GGUF), mobile-optimized, and Compressed Tensors. These models preserve “similar quality to bfloat16 while dramatically reducing the memory requirements to load the model,” according to Google.

    After downloading the Gemma 4 QAT model weights, users can run the checkpoints on their phones, laptops, or desktops. You can find the mobile and desktop models on Hugging Face, as well as in LM Studio.

    Thank you for being part of our community. Read our Comment Policy before posting.

    Footprint Gemma Memory models slash training trick
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    kirklandc008@gmail.com
    • Website

    Related Posts

    It’s Frighteningly Easy to Jailbreak Some Frontier AI Models

    July 29, 2026

    Malicious sites use JavaScript to build malware in browser memory

    July 25, 2026

    The OpenAI Models That Hacked Hugging Face Were ‘Active on the Internet’ for Days

    July 25, 2026
    Leave A Reply Cancel Reply

    Top Posts

    Nothing CEO says phone prices are going to keep going up

    June 12, 20267 Views

    Google DeepMind Plans to Track AGI Progress With These 10 Traits of General Intelligence

    March 21, 20263 Views

    The AirPods 4 and Lego’s brick-ified Grogu are our favorite deals this week

    October 12, 20253 Views
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram
    Latest Reviews

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Recent Posts
    • NASA’s new dark energy space telescope can also detect killer asteroids
    • The AI Notetaker Has Been Invited to All the Meetings
    • OK, Well, Rogue AI Agents Are Hacking Again
    • Heat Is an Orbital Data Center’s Greatest Foe. These Tiles Dump It at the Source.
    • The White House Is Keeping Its AI Cybersecurity Framework Secret

    NASA’s new dark energy space telescope can also detect killer asteroids

    August 5, 2026

    The AI Notetaker Has Been Invited to All the Meetings

    August 5, 2026

    OK, Well, Rogue AI Agents Are Hacking Again

    August 5, 2026

    Heat Is an Orbital Data Center’s Greatest Foe. These Tiles Dump It at the Source.

    August 5, 2026
    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Contact Us
    • Privacy Policy
    • Terms and Conditions
    • Disclaimer
    © 2026 TechNovaMindset. Designed by By Pro.

    Type above and press Enter to search. Press Esc to cancel.