homeservicesworkaboutblogfree templatescontactFree Tools →Free AI ModelsResearch LibraryROI CalculatorSavings CalculatorAI Readiness ScoreHire vs. AutomateAutomation Quote
book a 30-min call
home / free-models / granite-4-2-8b

IBM Granite 4.2 8B

IBM · United States · Apache 2.0

Commercial use: Yes — free for commercial use

Apache 2.0. IBM publishes provenance and governance documentation alongside the weights, which is unusual and useful in a procurement pack.

What this means for your business

A small, permissively licensed model from IBM that ships with documentation of what it was trained on.

Why it should matter to you

Most open models cannot tell you where their training data came from. That question is live in litigation and increasingly appears in enterprise procurement. Granite is one of the few families with an answer, and the answer is written down.

How it connects to our work

Where a client has real provenance exposure — a product that ships generated output, or a regulated buyer asking the question — we recommend a provenance-documented model even at a capability cost. It is a risk decision and it should be made on purpose.

Parameters8B
Active per token8B (dense)
Context128K
Modalitytext
Memory @ Q4_K_M~5 GB
Memory @ Q8_0~9 GB
LicenceApache 2.0
Last verified2026-09

Advantages

  • Apache 2.0 from a vendor whose enterprise paperwork your legal team already recognises
  • Documented training-data governance — a real differentiator when someone asks what it was trained on
  • Official GGUF, FP8, NVFP4 and MXFP4 builds published by IBM
  • Small enough for an 8 GB card

Disadvantages

  • Not competitive with Qwen3.8-27B on raw capability
  • Modest community fine-tune ecosystem
  • Conservative outputs; less useful for open-ended generation

Reach for it when

Regulated enterprises where training-data provenance is a procurement question and IBM as a counterparty carries weight.

Where it falls down

Anything needing frontier reasoning. This is a governance choice, not a capability choice, and it should be made deliberately.

Running it

8 GB GPU or 16 GB Mac at Q4_K_M. See the hardware sizing tables for how that maps to specific chips and cards, and the quantisation guide for what you give up at each bit width.

will it fitIBM Granite 4.2 8B against common GPUs
IBM Granite 4.2 8B at Q4_K_M
0 GB
IBM Granite 4.2 8B at Q8_0
0 GB
RTX 3060 12GB
0 GB
RTX 4060 Ti 16GB
0 GB
RTX 4090
0 GB
RTX 5090
0 GB
RTX 6000 Ada
0 GB
A100 80GB
0 GB
H100 80GB
0 GB
Weight sizes for IBM Granite 4.2 8B as recorded in this directory; GPU memory from the hardware table on the directory hub. Weights only: context adds KV cache.
where the weights fitGPU and Mac memory, checked
Q4_K_MQ8_0
RTX 3060 12GB (12 GB)
RTX 4060 Ti 16GB (16 GB)
RTX 4090 (24 GB)
RTX 5090 (32 GB)
RTX 6000 Ada (48 GB)
A100 80GB (80 GB)
H100 80GB (80 GB)
Mac, 16 GB unified
Mac, 24 GB unified
Mac, 32 GB unified
Mac, 36 GB unified
Mac, 64 GB unified
Mac, 96 GB unified
Mac, 128 GB unified
Mac, 192 GB unified
Mac, 512 GB unified
✓ fits with room for context, ~ fits with under 15% headroom, ✕ does not fit. Weights only, computed from the figures on this page.

Free tiers carrying this model

ProviderTypeThe catchLive limits
NVIDIA NIMInference hostIt is a credit grant, not a permanent free tier. When the credits run out you are on a paid plan or self-hosting the container.check →
OpenRouterAggregatorFree-pool membership changes without notice — a model you built on can stop being free, and the list above will drift. Data passes through OpenRouter and then the upstream provider, so check the privacy and routing settings before sending anything sensitive.check →

Jurisdiction: United States

Best raw capability and the deepest tooling ecosystem. For EU personal data you are relying on a transfer framework rather than on data never leaving the bloc, so check whether your DPA and your customers accept that.

Watch for: Enterprise API tiers usually promise no training on your data; consumer tiers and free tiers frequently do not. The free tier is where this bites.

Frequently asked

Can I use IBM Granite 4.2 8B commercially?

Apache 2.0. IBM publishes provenance and governance documentation alongside the weights, which is unusual and useful in a procurement pack.

What hardware do I need to run IBM Granite 4.2 8B?

8 GB GPU or 16 GB Mac at Q4_K_M. Weights alone are roughly 5 GB at Q4_K_M and 9 GB at Q8_0. Add KV cache on top of that, which grows with your context length.

What licence is IBM Granite 4.2 8B released under?

Apache 2.0. Full commercial use, modification and redistribution. Patent grant included. The most permissive licence in common use for open-weight models.

Where can I use IBM Granite 4.2 8B for free?

Free tiers carrying it include NVIDIA NIM, OpenRouter. Limits differ per provider and change often, so check each provider's own limits page. You can also self-host the weights, which has no rate limit at all.

Similar models

Wiring IBM Granite 4.2 8B into something real?

We build the evaluation harness, the failover and the cost ceilings around a model like this, so it survives contact with production.

Maps to AI agent development, AI automation development and AI strategy and consulting. Or see it working: our case studies.