homeservicesworkaboutblogfree templatescontactFree Tools →Free AI ModelsResearch LibraryROI CalculatorSavings CalculatorAI Readiness ScoreHire vs. AutomateAutomation Quote
book a 30-min call
home / free-models / hunyuan-ocr

HunyuanOCR

Tencent · China ·

Commercial use: Conditional — commercial use with a trigger

Custom Tencent licence, tagged "other". Read the repository terms before commercial use — do not assume permissive because the download is open.

What this means for your business

A model that reads scanned documents — invoices, forms, records — and turns them into structured text.

Why it should matter to you

Most businesses sitting on a paper or PDF archive assume digitising it is a bespoke project. Specialist OCR models have made the reading part close to solved. What has not changed is everything after: deciding that four spellings of a name are one customer.

How it connects to our work

Reading the page is now the easy half. On the patient registry extraction, OCR was days and identity resolution was months — and that ratio is typical rather than unusual.

From our field notesLegacy system access is the most overlooked pre-build blocker.

ParametersOCR specialist
Active per tokendense
Contextn/a
Modalityvision
Memory @ Q4_K_M~6 GB
Memory @ Q8_0~11 GB
LicenceCustom vendor licence
Last verified2026-09

Advantages

  • Purpose-built for document OCR rather than general vision, and it shows on dense scanned text
  • 660K downloads, indicating real adoption for extraction work
  • Handles multilingual and structured documents better than general VLMs

Disadvantages

  • Custom licence needs reading
  • Chinese-origin, which draws procurement questions in document work involving personal data
  • Specialist: no use outside document understanding

Reach for it when

Extraction pipelines over scanned archives, where a general vision model wastes capacity and money.

Where it falls down

Anything but documents. And check the licence before it touches a commercial pipeline — this is exactly the case where "downloadable" is doing a lot of unearned work.

Running it

12 GB GPU. See the hardware sizing tables for how that maps to specific chips and cards, and the quantisation guide for what you give up at each bit width.

will it fitHunyuanOCR against common GPUs
HunyuanOCR at Q4_K_M
0 GB
HunyuanOCR at Q8_0
0 GB
RTX 3060 12GB
0 GB
RTX 4060 Ti 16GB
0 GB
RTX 4090
0 GB
RTX 5090
0 GB
RTX 6000 Ada
0 GB
A100 80GB
0 GB
H100 80GB
0 GB
Weight sizes for HunyuanOCR as recorded in this directory; GPU memory from the hardware table on the directory hub. Weights only: context adds KV cache.
where the weights fitGPU and Mac memory, checked
Q4_K_MQ8_0
RTX 3060 12GB (12 GB)~
RTX 4060 Ti 16GB (16 GB)
RTX 4090 (24 GB)
RTX 5090 (32 GB)
RTX 6000 Ada (48 GB)
A100 80GB (80 GB)
H100 80GB (80 GB)
Mac, 16 GB unified
Mac, 24 GB unified
Mac, 32 GB unified
Mac, 36 GB unified
Mac, 64 GB unified
Mac, 96 GB unified
Mac, 128 GB unified
Mac, 192 GB unified
Mac, 512 GB unified
✓ fits with room for context, ~ fits with under 15% headroom, ✕ does not fit. Weights only, computed from the figures on this page.

Jurisdiction: China

This is the crucial split most write-ups miss: using a Chinese lab's HOSTED API sends your data to Chinese infrastructure and is a real procurement question. Downloading their OPEN WEIGHTS and running them on your own hardware, or on a Western host, sends nothing anywhere. DeepSeek and Z.ai publish under MIT; Alibaba publishes most of Qwen3 under Apache 2.0. Those are among the most permissive licences on this page.

Watch for: Several US states and a number of government bodies restrict Chinese-hosted AI services on official devices. That restriction is about the hosted service, not about the weights running in your own VPC — but expect to have to explain the difference to a security reviewer.

Frequently asked

Can I use HunyuanOCR commercially?

Custom Tencent licence, tagged "other". Read the repository terms before commercial use — do not assume permissive because the download is open.

What hardware do I need to run HunyuanOCR?

12 GB GPU. Weights alone are roughly 6 GB at Q4_K_M and 11 GB at Q8_0. Add KV cache on top of that, which grows with your context length.

What licence is HunyuanOCR released under?

A custom licence written by Tencent, not a standard open-source licence. Custom Tencent licence, tagged "other". Read the repository terms before commercial use — do not assume permissive because the download is open.

Where can I use HunyuanOCR for free?

Self-hosting the weights is the free route. 12 GB GPU.

Similar models

Wiring HunyuanOCR into something real?

We build the evaluation harness, the failover and the cost ceilings around a model like this, so it survives contact with production.

Maps to AI agent development, AI automation development and AI strategy and consulting. Or see it working: our case studies.