Loading model details...
Fetching the latest models and pricing from the API.
Loading model details...
Fetching the latest models and pricing from the API.
Released in late 2025, HunyuanOCR is an open-source contribution from Tencent that outperforms many larger proprietary models. It utilizes a "Global-to-Local" architecture with a SigLIP-v2 visual encoder to handle high-resolution inputs and extreme aspect ratios (like long receipts) without splitting images artificially.
/ 1M input tokens
/ 1M output tokens
from openai import OpenAI
# Initialize the OpenAI client with Qubrid base URL
client = OpenAI(
base_url="https://qubrid.com/v1",
api_key="QUBRID_API_KEY",
)
response = client.chat.completions.create(
model="tencent/HunyuanOCR",
messages=[
{
"role": "user",
"content": [
{
"type": "text",
"text": "What is in this image? Describe the main elements."
},
{
"type": "image_url",
"image_url": {
"url": "https://cdn.britannica.com/61/93061-050-99147DCE/Statue-of-Liberty-Island-New-York-Bay.jpg"
}
}
]
}
],
max_tokens=4096,
temperature=0,
stream=False,
extra_body={
"language": "en",
"ocr_mode": "general",
}
)
print(response.choices[0].message.content)from openai import OpenAI
# Initialize the OpenAI client with Qubrid base URL
client = OpenAI(
base_url="https://qubrid.com/v1",
api_key="QUBRID_API_KEY",
)
response = client.chat.completions.create(
model="tencent/HunyuanOCR",
messages=[
{
"role": "user",
"content": [
{
"type": "text",
"text": "What is in this image? Describe the main elements."
},
{
"type": "image_url",
"image_url": {
"url": "https://cdn.britannica.com/61/93061-050-99147DCE/Statue-of-Liberty-Island-New-York-Bay.jpg"
}
}
]
}
],
max_tokens=4096,
temperature=0,
stream=False,
extra_body={
"language": "en",
"ocr_mode": "general",
}
)
print(response.choices[0].message.content)Example response
INVOICE #1042 Date: 2026-05-28 Item Qty Amount API Credits 1 $49.00 Total: $49.00
from openai import OpenAI
# Initialize the OpenAI client with Qubrid base URL
client = OpenAI(
base_url="https://qubrid.com/v1",
api_key="QUBRID_API_KEY",
)
response = client.chat.completions.create(
model="tencent/HunyuanOCR",
messages=[
{
"role": "user",
"content": [
{
"type": "text",
"text": "What is in this image? Describe the main elements."
},
{
"type": "image_url",
"image_url": {
"url": "https://cdn.britannica.com/61/93061-050-99147DCE/Statue-of-Liberty-Island-New-York-Bay.jpg"
}
}
]
}
],
max_tokens=4096,
temperature=0,
stream=False,
extra_body={
"language": "en",
"ocr_mode": "general",
}
)
print(response.choices[0].message.content)from openai import OpenAI
# Initialize the OpenAI client with Qubrid base URL
client = OpenAI(
base_url="https://qubrid.com/v1",
api_key="QUBRID_API_KEY",
)
response = client.chat.completions.create(
model="tencent/HunyuanOCR",
messages=[
{
"role": "user",
"content": [
{
"type": "text",
"text": "What is in this image? Describe the main elements."
},
{
"type": "image_url",
"image_url": {
"url": "https://cdn.britannica.com/61/93061-050-99147DCE/Statue-of-Liberty-Island-New-York-Bay.jpg"
}
}
]
}
],
max_tokens=4096,
temperature=0,
stream=False,
extra_body={
"language": "en",
"ocr_mode": "general",
}
)
print(response.choices[0].message.content)Example response
INVOICE #1042 Date: 2026-05-28 Item Qty Amount API Credits 1 $49.00 Total: $49.00
Send document or image URLs • Extracts structured text from visuals Open in Playground for streaming, files, and all parameters.
Open in Playground