Close Menu

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    Meta and Nvidia plant flag in open-weight AI race led by Chinese Labs

    August 12, 2026

    Burnham warned Iran war could hit UK growth next year

    August 12, 2026

    Did poop enable the evolution of complex animals?

    August 12, 2026
    Facebook X (Twitter) Instagram
    Addison Markets
    • Home
    • USA
    • Europe
    • Business
    • Investing
    • Tech
    • Politics
    • Contact Us
    Addison Markets
    Home»Tech»Google’s new Gemma 4 12B model is designed to run on any laptop with 16GB of RAM
    Tech

    Google’s new Gemma 4 12B model is designed to run on any laptop with 16GB of RAM

    franperez66q@protonmail.comBy franperez66q@protonmail.comJune 4, 2026No Comments2 Mins Read
    Facebook Twitter Pinterest Telegram LinkedIn Tumblr WhatsApp Email
    Share
    Facebook Twitter LinkedIn Pinterest Telegram Email




    Gemma 4 12B is almost as capable as the version with 26 billion parameters.

    Credit:
    Google

    Gemma 4 12B is almost as capable as the version with 26 billion parameters.


    Credit:

    Google

    Google says the new model is capable of complex multistep reasoning and agentic workflows that previously required the larger Gemma variants. Despite the smaller parameter count, Gemma 4 12B comes with the newly devised Multi-Token Prediction (MTP) drafters, which take advantage of unused processing cycles to calculate possible future tokens. The result is greater speed and efficiency. Google has released optional MTP versions of the other Gemma 4 models, but this is the first one to have MTP out of the box.

    Gemma 4 12B is also more efficient thanks to a new approach to multimodality. The Gemma 4 family is natively multimodal, accepting text, audio, or images as inputs. Most gen AI models—including the other Gemma 4 variants—use dedicated encoders to process non-text inputs and pass that data to the LLM. This works well enough, but it increases latency and memory usage.

    With the new mid-weight model, Google has implemented a streamlined embedding module for vision, featuring single-matrix multiplication and positional embedding, which allows the data to pass to the LLM with proper spatial awareness. This eliminates the need for a bulky middleman encoder. For audio, there’s no encoding at all. The developers worked out a method of projecting the raw audio signal into the same vectors used for text tokens.

    If you want to check out the new Gemma 4 model, it’s accessible without a download via tools like LM Studio, Google AI Edge Gallery, and more. But the whole idea with Gemma 4 12B is that you can run it locally and on your own terms. If you’ve got the RAM, the model weights are available for download immediately on Kaggle and Hugging Face. It’s just shy of 18GB.



    Source link

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    franperez66q@protonmail.com
    • Website

    Related Posts

    Meta and Nvidia plant flag in open-weight AI race led by Chinese Labs

    August 12, 2026

    Did poop enable the evolution of complex animals?

    August 12, 2026

    AI’s costly buildout complicates the Fed’s inflation fight

    August 12, 2026

    Tracking extreme heat by the hour makes climate change seem even worse

    August 12, 2026

    Tencent Q2 earnings: Gaming accelerates, AI-driven ads boost revenue

    August 12, 2026

    Less than 2.5% of Taylor Farms’ recalled lettuce went to Taco Bells

    August 12, 2026
    Leave A Reply Cancel Reply

    Top Reviews
    Editors Picks

    Meta and Nvidia plant flag in open-weight AI race led by Chinese Labs

    August 12, 2026

    Burnham warned Iran war could hit UK growth next year

    August 12, 2026

    Did poop enable the evolution of complex animals?

    August 12, 2026

    Swinney calls for Burnham rethink on Chinese factory in Highlands

    August 12, 2026
    © 2026 All right reserved
    • Privacy Policy
    • Terms & Conditions

    Type above and press Enter to search. Press Esc to cancel.