2 Sources
[1]
IBM's open source Granite 4.0 Nano AI models can run locally directly in your browser
In an industry where model size is often seen as a proxy for intelligence, IBM is charting a different course -- one that values efficiency over enormity, and accessibility over abstraction. The 114-year-old tech giant's four new Granite 4.0 Nano models, released today, range from just 350 million
[2]
IBM releases small open-source Granite 4 models for mobile devices and browsers - SiliconANGLE
IBM releases small open-source Granite 4 models for mobile devices and browsers IBM Corp. today announced the release of Granite 4 Nano, a family of extremely small generative artificial intelligence models designed to run at the edge, on-device or in browsers. The company said the models exhibit
Share
Copy Link
IBM unveils four new open-source Granite 4.0 Nano AI models ranging from 350M to 1.5B parameters, designed to run locally on consumer hardware including laptops and web browsers. These compact models outperform competitors in benchmarks while requiring minimal computing resources.
IBM has released four new Granite 4.0 Nano AI models that fundamentally challenge the prevailing "bigger is better" philosophy in artificial intelligence
1
. The models, ranging from just 350 million to 1.5 billion parameters, represent a fraction of the size of server-bound models from companies like OpenAI, Anthropic, and Google, yet deliver competitive performance in their class2
.
Source: SiliconANGLE
The 114-year-old tech giant's approach prioritizes efficiency and accessibility over raw computational power. The smallest 350M variants can run comfortably on modern laptop CPUs with 8-16GB of RAM, while the 1.5B models typically require a GPU with at least 6-8GB of VRAM for optimal performance
1
. Remarkably, the smallest models can even run locally in web browsers, as demonstrated by Joshua Lochner, creator of Transformer.js and machine learning engineer at Hugging Face1
.The Granite 4.0 Nano family includes four distinct models now available on Hugging Face under the permissive Apache 2.0 license
1
. The lineup consists of Granite-4.0-H-1B and H-350M models featuring hybrid state space architecture (SSM), alongside standard transformer variants Granite-4.0-1B and 350M2
.The hybrid architecture represents IBM's innovative approach, combining transformer design with processing components based on the Mamba neural network architecture, which proves more hardware-efficient than traditional transformers
2
. The H-series models excel in low-latency edge environments, while the transformer variants offer broader compatibility with existing tools like llama.cpp1
.Despite their compact size, the Granite 4.0 Nano models demonstrate impressive benchmark results that rival or exceed larger competitors. According to data from David Cox, VP of AI Models at IBM Research, the Granite-4.0-H-1B scored 78.5 on IFEval instruction following benchmarks, significantly outperforming Qwen3-1.7B at 73.1 and Gemma 3-1B at 59.3
1
2
.In function calling capabilities, measured by Berkeley's Function Calling Leaderboard v3, the Granite-4.0-1B achieved a leading score of 54.8, surpassing Qwen3 at 52.2 and significantly outperforming Gemma 3 at 16.3
2
. The models also excelled in safety benchmarks, scoring over 90% on SALAD and AttaQ evaluations1
.Related Stories
IBM enters a crowded small language model market that includes competitors like Qwen3, Google's Gemma, LiquidAI's LFM2, and Mistral's dense models in the sub-2B parameter space
1
. However, while major players like OpenAI and Anthropic focus on models requiring GPU clusters, IBM targets developers seeking performant LLMs for local or constrained hardware environments2
.The models are certified under ISO 42001 for responsible AI development, a standard IBM helped pioneer, and maintain native compatibility with popular frameworks including llama.cpp, vLLM, and MLX
1
. This comprehensive compatibility ensures broad adoption potential across diverse development environments and use cases.Summarized by
Navi
[1]
03 Oct 2025•Technology

21 Oct 2024•Technology

19 Dec 2024•Technology

1
Technology

2
Science and Research

3
Technology
