The KAGAMI mark КАГАМИ
kagami.bg/academy · lesson · machine-readable viewVERIFIED 2026-10-01 · UPDATED 2026-10-01
IDENTITY
module
GX10-04-212 · HunyuanOCR 1.5 (Tencent): a 1B end-to-end OCR vision-language model with text spotting (coordinates), and why its licence rules it out in the EU
series
GX10 (local AI server class: NVIDIA GB10, e.g. ASUS Ascent GX10 / DGX Spark)
level
Beginner
duration
about 45 min
prerequisites
Ability to read a licence text; shell access only for the optional licence-check command
trust_label
VERIFIED 2026-10-01 (Hugging Face model card and LICENSE file, GitHub README of the 1.5 release and of the archived 1.0 branch) · UPDATED 2026-10-01 · NOT TESTED (nothing was downloaded or run)
licence_finding
Tencent Hunyuan Community License Agreement (HunyuanOCR release date 2025-11-25). The file states that it does NOT apply in the European Union, United Kingdom and South Korea; Territory = worldwide excluding those. Use, reproduction, modification, distribution or display of the works or of their Output outside the Territory is unlicensed (section 5c; Acceptable Use Policy item 1). Bulgaria is in the EU, so the model must not be used there
versions
HunyuanOCR-1.5 (announced 2026-07-07) at the root of tencent/HunyuanOCR; HunyuanOCR-1.0 archived under v1.0/ · 1B parameters, BF16 · arXiv 2607.04884
language
human view: bg · english edition: /en/academy/gx10/ (same file name)
previous / next
04-211 dots.ocr / 04-213 Tesseract (related: 04-214 EasyOCR, 04-215 PaddleOCR, 04-216 OCR comparison)
PURPOSE

Teach the habit of reading a model licence before installing the model, using HunyuanOCR as the worked example: a small and capable OCR model whose licence excludes the EU. Summarise what the model is (document parsing, text spotting with coordinates, information extraction, text-image translation), what its documentation says about inference, and which OCR tools with coordinate output can be used in the EU instead. No speed or accuracy claims are verified here.

KEY CONCEPTS
COMMANDS / PATHS
CHECKLIST
NEXT MODULE

04-213 Tesseract · related: 04-214 EasyOCR, 04-215 PaddleOCR (classic pipeline with coordinates) · GX10 series index · offer: Quick experiment (kagami.bg/stalbata/)

SOURCES
TAGS
gx10nvidia-gb10ocrhunyuanocrtext-spottinglicence-checkeu-exclusionlocal-ai
VERIFIED · 01.10.2026 UPDATED · 01.10.2026

HunyuanOCR on GX10: a licence that excludes the EU

HunyuanOCR is a small (1 billion parameters) model from Tencent that reads documents and returns where on the page each line is — coordinates. It is technically interesting. But its licence says it does not apply in the European Union — which includes Bulgaria. This lesson shows how to read a licence before installing anything, and what to use instead.

⏱ ~45 min Beginner GX10 NVIDIA GB10 · 128 GB unified memory HunyuanOCR 1.5 · 1B · Tencent licence
HunyuanOCR (weights and code)🔒 local Tencent online demo🌐 global
⛔
Cannot be used in Bulgaria
The HunyuanOCR licence opens with the sentence: "THIS LICENSE AGREEMENT DOES NOT APPLY IN THE EUROPEAN UNION, UNITED KINGDOM AND SOUTH KOREA". The Territory is the whole world excluding those three. Bulgaria is in the EU, so do not install or use the model here — neither for yourself nor for a client. We read the text as published; for a binding decision, ask a lawyer.
🔄
UPDATED · 01.10.2026 — what changed
The lesson was rewritten. The old version only said "licence 'other' — check the terms". Now we have read the licence itself and it explicitly excludes the EU — that is the main point of the lesson. We added: the current version HunyuanOCR 1.5 (07.07.2026), the exact territory clauses, the other restrictions in the licence, what the documentation says about running it (the shared environment since 24.07.2026 and the lighter recipes), an honest "What we have not run" section and a table of alternatives for the EU. We removed: the example run code (the licence does not apply here), the unsupported claims "SOTA" and "many parallel instances on the GX10", the base-model notice that had nothing to do with this lesson, and a private use case that does not belong in a public lesson.
⚠️
What we have not run ourselves
We did not download or run the model — which is why there is no "TESTED" label. Everything here is checked against the published Hugging Face and GitHub pages as of 01.10.2026. Unchecked: how it behaves on a GB10-class machine with an Arm processor (its documentation gives CUDA 13 examples and says nothing about Arm), how well it reads Bulgarian (the language names we saw do not include Bulgarian) and how accurate its coordinates are on your documents. We give no accuracy figures.

01What you'll learn

02Before you start

💡
Why this matters for the GX10
A GB10-class machine fits a 1B-parameter model easily — which is exactly why it is tempting to run it "just to try". But permission comes from the licence, not from whether the hardware can do it. A good habit: for every new model, open the LICENSE file first.

03Steps

  1. The licence first

    On the model's page on Hugging Face the "License" field shows tencent-hunyuan-community — not Apache-2.0 and not MIT, but Tencent's own licence. The text itself is in the LICENSE file. You can also view it from the terminal:

    bash · on the machine
    curl -sL https://huggingface.co/tencent/HunyuanOCR/raw/main/LICENSE | head -n 3
    curl -sL https://huggingface.co/tencent/HunyuanOCR/raw/main/LICENSE | grep -n -i "territory"

    The first command shows the title and the sentence about the EU. The second shows every line with the word "territory" — so you find the territory clauses without reading 20 pages at once.

  2. What the licence says about the territory

    Five places in the text say the same thing:

    WhereWhat it says (meaning)
    The first lineThe agreement does not apply in the European Union, the United Kingdom and South Korea and is limited to the "Territory".
    Definition 1(l)"Territory" = the whole world, excluding the EU, the United Kingdom and South Korea.
    Section 2The right to use, copy, modify and distribute is granted for the Territory only.
    Section 5(c)Do not use, copy, modify, distribute or display the materials or their outputs outside the Territory; such use is "unlicensed and unauthorized".
    Acceptable Use Policy, item 1Do not use the model or its derivatives outside the Territory.

    Two things that are often missed: (1) the restriction applies to the output too (the recognised text, the JSON, the Markdown) — not only to the weights themselves; (2) "I only downloaded the weights from Hugging Face" changes nothing — the terms are about use, not about where it was downloaded.

    ✅
    Rule
    If any part of the use — the machine, the users or the client — is in the EU, skip the model and pick an alternative. Do not look for a "workaround".
  3. The other restrictions in the licence

    Even outside the EU the licence is not "free". In short, what else it says:

    • Do not improve another AI model with the model or its outputs (section 5b) — for example, do not train your own model on text recognised with it.
    • Acceptable Use Policy: a ban on military use, on high-stakes automated decisions (health, employment, credit, law enforcement and others) and on publishing machine-generated content without a clear note that it is machine-generated.
    • If you offer a service built on it: include a copy of the licence and a "Notice" file, state clearly who the real provider is and that Tencent is not affiliated with the service (sections 3d and 3e).
    • Over 100 million monthly active users — a separate licence from Tencent is needed (section 4).
    • Law and courts: the laws and courts of Hong Kong (section 9).

    This is our reading of the published text, not legal advice. For real company use — a lawyer.

  4. What the model is

    According to the model page (version 1.5, announced on 07.07.2026) HunyuanOCR is a "lightweight, end-to-end OCR-specialized vision-language model": you give it an image and an instruction, you get text back. One model does four things — document parsing, text spotting with coordinates, information extraction and translation of text in images.

    WhatAccording to the documentation
    Sizeabout 1 billion parameters (1.12 billion per Hugging Face), BF16 format
    Versions1.5 sits at the root of the repository; 1.0 is archived in the v1.0/ folder
    New in 1.5faster decoding (DFlash), running through llama.cpp on a laptop or an ordinary graphics card, images up to 4K, context up to 128K
    LanguagesFor 1.0 the documentation claims "over 100 languages". Bulgarian is not named in the texts we checked — ⚠️ not confirmed. Translation of text in images is described for 14 languages (German, Spanish, Turkish, Italian, Russian, French, Portuguese, Arabic, Thai, Vietnamese, Indonesian, Malay, Japanese, Korean) — Bulgarian is not among them
    RunningvLLM (OpenAI-compatible server), "transformers" (5.13.0 or newer), llama.cpp (GGUF)
    HardwareFor 1.0: Linux, an NVIDIA GPU, ~20 GB of GPU memory with vLLM. For 1.5: CUDA 13

    What "coordinates" (spotting) are: for each line of text the model returns not only the letters but also where in the picture the line is — a box. This is useful when you want to find a field in a form or place something at an exact spot in a PDF. The official client has separate task types for this (spotting_json and spotting_hunyuan) and for document parsing (doc_parse), tables, formulas and charts.

    ℹ️
    The vendor's numbers
    Tencent publishes comparison tables in which the model leads — on its own in-house text-spotting test and on OmniDocBench. These are the vendor's numbers; we have not reproduced them. OCR benchmarks are crowded with entries; a trial on your own pages decides.
  5. Does it work on GB10 — honestly: we do not know

    Since 24.07.2026 both the Hugging Face page and the GitHub README describe one shared environment for running it: CUDA 13, Python 3.12, vLLM 0.25.1 or newer and flash-attn 2.8.3 built by hand. The first 1.5 release used three separate environments that cannot be mixed; they remain in the documentation as "lighter recipes" for machines without CUDA 13 — for example vLLM 0.18.1 with CUDA 12. None of the texts says anything about Arm64 (the GB10 processor is not x86). Have we run it — no. And after the previous step the question is moot for the EU anyway.

  6. What to use instead in the EU

    If you need line coordinates, there are tools with a lesson here. We checked the licences of PaddleOCR (Apache-2.0) and dots.mocr in their own lessons; for the others, check the licence the same way as in step 1. The full comparison is in lesson 04-216.

    ToolWhat forLesson
    PaddleOCR 3.x 🔒 localApache-2.0 licence. The documented result fields include dt_polys (the line boxes). Bulgarian with lang="bg". PP-StructureV3 returns Markdown and tables04-215
    Tesseract 🔒 localClassic OCR on the CPU; Bulgarian with bul; word positions through image_to_data or TSV/hOCR output04-213
    EasyOCR 🔒 localReturns a box, the text and a confidence score for each line; Bulgarian with bg04-214
    dots.mocr 🔒 localA whole-page model: box, category and text for each element (a block, not a line). MIT plus a supplementary agreement with no EU exclusion, but with bans on sensitive personal data; Bulgarian is not named04-211

    Before you take any of them into a product — repeat step 1 for it: open the licence and check the terms. That habit is the point of this lesson.

  7. Security — the minimum

    • Tencent's online demo is a Tencent service — documents you upload there leave your machine. Do not use it for confidential documents.
    • Do not put other people's real documents into tests and lessons; use your own or invented ones.
    • Recognised amounts, dates and names are a draft — before they enter accounting, a contract or a database, a person reviews them.

04Check

Quiz

1. What is the "Territory" according to the HunyuanOCR licence?

2. A Bulgarian company wants to put HunyuanOCR into a service for clients in Varna. What follows from the licence?

3. What can you use in the EU if you need line coordinates?

4. Which other restriction is in the HunyuanOCR licence?

05What's next

06Sources

  1. HunyuanOCR on Hugging Face 🌐 global — the model page (1.5), licence field, size, ways to run it.
  2. Tencent Hunyuan Community License Agreement — the first line, definition 1(l), sections 2, 3, 4, 5, 9 and the Acceptable Use Policy.
  3. HunyuanOCR on GitHub — README of 1.5 (news of 07.07.2026, environments) · README of 1.0 (languages, spotting prompt, requirements).
  4. HunyuanOCR-1.5 (arXiv 2607.04884) · 1.0 technical report (arXiv 2511.19575).