Artificial Intelligence

Google Releases Gemma 4 12B Open AI Model Designed for Laptops

MOUNTAIN VIEW, Calif.: Google releases Gemma 4 12B, a multimodal open-weight AI model designed to run locally on laptops with just 16GB of memory.

Google DeepMind announces the model under an Apache 2.0 license. Gemma 4 12B handles text, images, and native audio inputs without separate encoders. The encoder-free architecture cuts latency and memory requirements significantly versus traditional multimodal designs.

The model packs 12 billion parameters and delivers performance close to much larger systems. Google benchmarks show it approaching the Gemma 4 26B mixture-of-experts model on key tasks. This makes it the company’s first mid-sized Gemma model with native audio support.

Also read: Google Launches Two New AI Research Agents via Gemini API

Weights deploy freely on Kaggle and Hugging Face at just under 18GB total. The model runs through Hugging Face Transformers, vLLM, SGLang, MLX, llama.cpp, and LiteRT-LM. Developers can serve it as an OpenAI-compatible local API through the new litert-lm CLI.

Google also launches AI Edge Gallery and AI Edge Eloquent for macOS users. Both apps run Gemma 4 12B fully on-device, processing voice and visual inputs locally. A sandboxed Python execution loop lets users plot scientific charts inside the chat interface.

DRAM prices jumped roughly 90% in Q1 2026 as memory production redirected toward AI data centers. Micron told CNBC at CES it had effectively sold out memory capacity for 2026. A capable 16GB-memory model sidesteps both the hardware crunch and ongoing cloud inference costs.

Gemma models have now crossed 150 million total downloads since launch.

The launch reflects a broader industry pivot toward on-device AI deployment recently. Microsoft pushed Surface Laptop Ultra with RTX Spark earlier this week for local AI workloads. Apple Intelligence, Anthropic, and OpenAI all explore similar device-resident model strategies.

Shobhit Kalra

Shobhit Kalra is the Chief Sub Editor at Tea4Tech, with over 12 years of experience across digital media, digital marketing, and health technology. He is responsible for editorial review, content structuring, and quality control of articles covering software, SaaS products, and developments across the technology ecosystem. || At Tea4Tech, Shobhit oversees content accuracy, clarity, and adherence to editorial standards, ensuring published stories meet the newsroom’s guidelines for originality, sourcing, and consistency.

Recent Posts

OpenAI Launches GPT-6 Astra With Advanced Computer and AI Capabilities

New Delhi: OpenAI has launched GPT-6 Astra, its latest and most advanced AI model, with…

2 days ago

India’s Rilo Joins Adobe in AI Marketing Automation Deal

New Delhi: Adobe has acquired Rilo, an India based marketing intelligence startup, in a deal…

6 days ago

Anthropic Levels Up Claude Fable 5.1 With More Power, Less Cost

New York: Anthropic has launched Claude Fable 5.1 and Claude Mythos 5.1, its latest AI…

7 days ago

Perplexity Launches ‘Portable Computer’, a Local AI Agent Built With Nvidia

Bengaluru: Perplexity AI has launched a new product called Portable Computer, an AI agent that…

2 weeks ago

India’s Ringg Bags $10M From Peak XV for Voice AI Push

New Delhi: Bengaluru-based voice AI startup Ringg has raised $10 million from Peak XV Partners,…

2 weeks ago

Google Halo Explained: A New Way to See What Your AI Agent Is Doing

New York: Google is moving forward with Android Halo, a new system-level feature designed to…

2 weeks ago