Model catalog

Tencent AI Models

Browse the AI models associated with Tencent AI. Compare current and historical models by family, capabilities, context window, availability and intended use.

53 models tracked
53 Total models
33 Model families
10 Model types
48 Current / accessible
All models

Tencent AI model catalog

Tencent AI logo
AuK

AuK

Open-source text-to-speech, reference-voice generation, speech and lyric editing, emotion and timbre transformation, speech enhancement, and source separation

Type Other
Reasoning 2/10
Speed 5/10
Multimodal Audio input Media output
Status

Current open-source release

View model →

Lightweight local assistants, edge-oriented inference, long-context text processing, prototyping, quantized deployment, and low-resource instruction-following workloads.

Type Lightweight
Context 262K
Reasoning 3/10
Speed 9/10
Fine-tuning
Status

Current open-weight release

View model →

Efficient local text generation, long-context analysis, lightweight reasoning, coding assistance, and agent-oriented applications

Type Lightweight
Context 256K
Reasoning 6/10
Speed 9/10
Tool use Fine-tuning Streaming
Status

Current open-weight model; publicly available for download and local deployment

View model →

Local and self-hosted text generation, Chinese and multilingual instruction following, long-context analysis, mathematics, coding, reasoning, and lightweight agent workloads

Type General Purpose
Context 262K
Reasoning 7/10
Speed 8/10
Tool use Fine-tuning Streaming
Status

Current open-weight model; publicly available for self-hosted deployment

Input No official hosted API price published; self-hosted model weights are available
Output No official hosted API price published; self-hosted inference costs depend on infrastructure
View model →

Chinese and multilingual text generation, reasoning, coding, long-context workloads, private local inference, quantized deployment, and domain fine-tuning.

Type General Purpose
Context 262K
Reasoning 7/10
Speed 7/10
Fine-tuning
Status

Current open-weight downloadable model

Input No first-party hosted API token price identified; downloadable weights are intended for self-hosted or third-party deployment.
Output No first-party hosted API token price identified; downloadable weights are intended for self-hosted or third-party deployment.
View model →
Tencent AI logo
Hunyuan-A13B

Hunyuan-A13B

Open-weight reasoning, long-context analysis, mathematics, science, coding, agent workflows, and cost-conscious self-hosted inference

Type Reasoning
Context 262K
Reasoning 8/10
Speed 8/10
Tool use Fine-tuning Streaming
Status

Current and accessible as an open-weight model; also listed as the Tencent Cloud API model hunyuan-a13b. Tencent Cloud documentation notes an ongoing migration of Hunyuan services toward TokenHub.

Input ¥0.50 per 1 million input tokens on Tencent Cloud postpaid API
Output ¥2 per 1 million output tokens on Tencent Cloud postpaid API
View model →
Tencent AI logo
Hunyuan-Large

Hunyuan-Large

Large-scale Chinese and English text generation, reasoning, mathematics, coding, long-context analysis, research, and self-hosted experimentation

Type General Purpose
Context 256K
Reasoning 8/10
Speed 4/10
Fine-tuning Streaming
Status

Available as an open-weight release; legacy for the historical Tencent Cloud hunyuan-large API identifier

View model →
Tencent AI logo
Hunyuan-MT

Hunyuan-MT-7B

Self-hosted multilingual translation, localization, translation research, and Chinese dialect or minority-language translation

Type Other
Context 33K
Reasoning 2/10
Speed 5/10
Fine-tuning
Status

Current open-weight model

View model →
Tencent AI logo
Hunyuan3D

Hunyuan3D-2.1

Image-to-3D asset creation, game and virtual-world content, product visualization, design prototyping, and self-hosted 3D generation

Type Other
Speed 4/10
Multimodal Image input Media output
Status

Current open-weight release; self-hosted deployment

Input No official hosted token or per-request API price; model weights are available for self-hosting
Output No official hosted output price; generated 3D assets are produced through self-hosted inference
View model →
Tencent AI logo
Hunyuan3D

Hunyuan3D 2.0

Local image-to-3D asset generation, textured mesh creation, game and design prototypes, Blender workflows, and research on open 3D generative models

Type Other
Speed 6/10
Multimodal Image input Media output
Status

Available open-weight model system; superseded by the newer Hunyuan3D-2.1 release but still publicly accessible

View model →
Tencent AI logo
Hunyuan 3D

HY-3D-3.0

Text-to-3D, image-to-3D, sketch-to-3D, rapid game and e-commerce asset creation, 3D printing, and production-oriented asset prototyping.

Type Other
Speed 8/10
Multimodal Image input Media output
Status

Current and accessible through Tencent Cloud APIs; newer HY-3D-3.1 is also available as a separate model version.

Input Credit-based pricing: 25 credits for a default Professional normal textured generation; 15 credits for Geometry mode; 30 credits for LowPoly mode. Additional features such as PBR, multi-view, and custom face count consume extra credits.
Output Not token-priced; generated 3D assets are billed by generation credits.
View model →

Local English- and Chinese-language text-to-image generation, creative prototyping, ComfyUI workflows, LoRA customization, and ControlNet-based image conditioning

Type Other
Reasoning 1/10
Speed 6/10
Media output Fine-tuning
Status

Open-weight and downloadable; no official deprecation or shutdown notice found

Input No official hosted API input pricing published; intended primarily for local deployment
Output No official hosted API output pricing published; image generation uses local compute rather than a documented token price
View model →
Tencent AI logo
HunyuanImage

Hy-Image-3.0

Text-to-image generation, reference-guided image creation, marketing graphics, e-commerce content, visual ideation, and image-generation applications requiring custom aspect ratios

Type Other
Speed 7/10
Multimodal Image input Media output
Status

Current and accessible through Tencent Cloud TokenHub; also available as an open-weight HunyuanImage-3.0 release

Output Approximately 0.20 CNY per generated image in Tencent Cloud mainland China TokenHub pricing; approximately 0.032 USD per image in cited international pricing
View model →
Tencent AI logo
HunyuanOCR

HunyuanOCR-1.5

Multilingual OCR, document parsing, text spotting, table and formula extraction, structured information extraction, and local visual-document processing

Type Multimodal
Context 131K
Reasoning 3/10
Speed 8/10
Multimodal Image input Fine-tuning
Status

Current open-weight release

View model →
Tencent AI logo
Hunyuan turbos-vision

HY-Vision-Video

Video description, video question answering, video summarization, content review, scene analysis, and video metadata generation.

Type Multimodal
Context 32K
Reasoning 3/10
Speed 8/10
Multimodal Video input
Status

Online and currently listed in Tencent Cloud TokenHub; the older Hunyuan platform entry was retired on 2026-06-22, while the TokenHub model remains available.

Input CNY 3 per 1 million input tokens
Output CNY 9 per 1 million output tokens
View model →
Tencent AI logo
HunyuanVideo

HunyuanVideo

Research and production experimentation with high-quality local text-to-video generation

Type Other
Reasoning 1/10
Speed 3/10
Media output
Status

Open-source and currently accessible; original HunyuanVideo model, with HunyuanVideo-1.5 released later as a lighter successor

View model →
Tencent AI logo
HunyuanVideo

HunyuanVideo-1.5

Local text-to-video and image-to-video generation, open-source video research, creative prototyping, and developers needing a comparatively lightweight high-quality video model.

Type Other
Reasoning 1/10
Speed 6/10
Multimodal Image input Media output
Status

Current and publicly accessible open-weight model

Input No official hosted API price identified; model weights are available for local deployment under the Tencent Hunyuan Community License.
Output No official hosted API price identified; local inference costs depend on hardware and infrastructure.
View model →
Tencent AI logo
HunyuanVideo

HunyuanVideo-I2V

Image-to-video generation, reference-image animation, visual effects, creative video prototyping, and locally hosted open-weight video workflows

Type Other
Speed 4/10
Multimodal Image input Media output
Status

Available as an open-weight model and official open-source repository; no official shutdown date found

View model →
Tencent AI logo
Hunyuan Vision 1.5

HY-Vision-1.5-Thinking

Image-grounded reasoning, OCR, chart and document analysis, visual localization, educational problem solving, and multilingual visual question answering

Type Multimodal
Context 40K
Reasoning 8/10
Speed 6/10
Multimodal Image input Video input
Status

Online and currently available through Tencent Cloud TokenHub

Input CNY 3 per 1M input tokens
Output CNY 9 per 1M output tokens
View model →

Coding agents, long-context analysis, complex reasoning, productivity automation, structured workflows, and multi-step tool use

Type Reasoning
Context 256K
Reasoning 8/10
Speed 7/10
Tool use Web search Fine-tuning
Status

Current; open-weight and available through Tencent Cloud TokenHub

Input CNY 1 per 1 million input tokens; cached input CNY 0.25 per 1 million tokens on Tencent Cloud TokenHub
Output CNY 4 per 1 million output tokens on Tencent Cloud TokenHub
View model →

Automated decomposition of FBX 3D models into separate model components.

Type Other
Media output
Status

Current and API-accessible

Input Not publicly specified in the model documentation; enabling post-processing adds 20 points.
View model →

Fast text-to-3D and image-to-3D asset generation, prototyping, and automated 3D content pipelines

Type Other
Reasoning 1/10
Speed 8/10
Multimodal Image input Media output
Status

Current and available through Tencent Cloud TokenHub

Input 15–25 Tencent Cloud points per generation request; 1 point is listed as CNY 0.12
View model →

Automated conversion of existing 3D assets between common interchange and delivery formats

Type Other
Speed 8/10
Media output
Status

Current TokenHub 3D format-conversion service

Input 5 credits per call
View model →

Generating short human character animations from natural-language action descriptions

Type Other
Reasoning 1/10
Speed 6/10
Multimodal Media output
Status

Current and accessible through Tencent TokenHub

Input 10 points per generation request
View model →

Automated retopology, polygon reduction, and preparation of existing 3D meshes for games, rendering, animation, and downstream asset workflows

Type Other
Speed 7/10
Media output
Status

Current and API-accessible

View model →

Automated rigging and skinning of human or animal 3D characters for animation, games, virtual characters, and asset prototyping

Type Other
Reasoning 1/10
Speed 5/10
Media output
Status

Current and accessible through Tencent Cloud TokenHub

Input 10 Tencent Cloud credits per request
Output Included in the 10-credit per-request charge; output is a rigged 3D model file
View model →

Reference-guided texturing of existing OBJ or GLB meshes, including PBR material generation and automated asset preparation

Type Other
Reasoning 1/10
Speed 6/10
Multimodal Image input Media output
Status

Current and available through Tencent Cloud TokenHub API

Output 30 Tencent Cloud points per generation
View model →

Automated UV unwrapping and preparation of 3D assets for texturing, rendering, game development, and digital-content workflows.

Type Other
Reasoning 1/10
Speed 7/10
Media output
Status

Current and accessible through Tencent Cloud TokenHub

Output 10 points per call; Tencent Cloud states that 1 point equals CNY 0.12
View model →

Real-time and asynchronous Mandarin, English, mixed-language, and Chinese-dialect transcription, captions, subtitles, and short voice-command recognition

Type Other
Reasoning 4/10
Speed 8/10
Audio input Streaming
Status

Preview; internal-test availability

Input Usage-based Tencent Cloud ASR pricing; exact model price not verified in the reviewed official documentation
Output No separate output price; transcription is billed through the applicable Tencent Cloud ASR or large-model 2.0 pricing scheme
View model →

High-resolution text-to-image generation, reference-image creation, multi-turn image editing, posters, UI concepts, marketing assets, and product-image workflows

Type Multimodal
Context 100K
Speed 7/10
Multimodal Image input Media output
Status

Current preview model

Input 10 CNY per million tokens; reference-image generation uses the published TokenHub token rules
Output 15,000 tokens per 1K or 2K image; 20,000 tokens per 4K image, equivalent to approximately CNY 0.15 and CNY 0.20 respectively at the listed CNY 10 per million-token rate
View model →

Low-latency multilingual translation, localized content, structured translation instructions, and cost-sensitive production workflows

Type Lightweight
Context 8K
Reasoning 2/10
Speed 9/10
Streaming
Status

Current and available through Tencent Cloud TokenHub

Input ¥0.3 per 1 million tokens
Output ¥1.2 per 1 million tokens
View model →

Professional multilingual translation, localization, terminology-sensitive workflows, and context-aware business translation

Type Translation
Context 8K
Reasoning 2/10
Speed 8/10
Streaming
Status

Current and available through Tencent Cloud TokenHub

Input ¥0.5 per 1 million tokens in China; US$0.074 per 1 million tokens on Tencent Cloud international pricing
Output ¥2 per 1 million tokens in China; US$0.295 per 1 million tokens on Tencent Cloud international pricing
View model →

Professional, domain-specific and high-quality multilingual translation with contextual disambiguation and instruction following

Type Specialized
Context 8K
Reasoning 6/10
Speed 8/10
Streaming
Status

Online and currently available through Tencent Cloud TokenHub

Input 0.5 CNY per 1 million input tokens
Output 2 CNY per 1 million output tokens
View model →
Tencent AI logo
Hy-Role

Hy-Role

Chinese role-play, character simulation, fictional dialogue, AI avatars, and emotionally oriented conversational experiences

Type Other
Context 32K
Reasoning 3/10
Speed 7/10
Streaming
Status

Current and available through Tencent Cloud TokenHub

Input CNY 2.4 per 1 million input tokens
Output CNY 9.6 per 1 million output tokens
View model →

Image understanding, OCR, chart and diagram analysis, STEM visual reasoning, visual question answering, and multi-image comparison

Type Multimodal
Context 44K
Reasoning 7/10
Speed 8/10
Multimodal Image input
Status

Current and available through Tencent Cloud TokenHub

Input ¥7.5 per 1 million input tokens
Output ¥17.5 per 1 million output tokens
View model →

Generating explorable 3D environments, Gaussian-splat scenes, point clouds, and collision meshes from text or reference images

Type Other
Reasoning 1/10
Speed 2/10
Multimodal Image input Media output
Status

Current and accessible through Tencent Cloud TokenHub

Input ¥10 per 1 million tokens; reference usage is 8,000,000 tokens per generated scene
Output Approximately ¥80 per generated 3D scene at the documented reference usage
View model →

Text-to-3D, image-to-3D, multi-view reconstruction, game assets, product visualization, digital-human content, e-commerce assets and 3D printing workflows

Type Multimodal
Reasoning 1/10
Speed 7/10
Multimodal Image input Media output
Status

Current and available through Tencent Cloud HY-3D APIs and TokenHub

Input Not token-priced; text and image inputs are included in the per-generation task charge
Output 15–60 credits per generation; CNY 0.12 per credit, approximately CNY 1.80–7.20 per generation
View model →

Long-context coding agents, complex tool-use workflows, productivity automation, document analysis, game development, and scientific reasoning

Type General Purpose
Context 1M
Reasoning 8/10
Speed 6/10
Tool use Fine-tuning Streaming
Status

Preview; currently available as an open-weight model and through Tencent products, Tencent Cloud TokenHub, and OpenRouter

Input CNY 6 per 1M tokens; cached input CNY 0.3 per 1M tokens
Output CNY 18 per 1M tokens
View model →

Text-to-panorama and image-to-panorama generation for immersive environments, virtual tours, games, simulations, visualization, and 3D-world pipelines.

Type Other
Reasoning 1/10
Speed 7/10
Multimodal Image input Media output
Status

Current and accessible through Tencent Cloud TokenHub

Input USD 1.60 per million tokens; reference usage 481,250 tokens per panorama
Output Approximately USD 0.77 per generated panorama
View model →

High-volume semantic retrieval, vector search, FAQ matching, text clustering, classification, and cost- or latency-sensitive knowledge-base applications

Type Embedding
Context 33K
Reasoning 1/10
Speed 8/10
Status

Current and available through Tencent Cloud TokenHub

Input $0.07 per million text-input tokens
View model →

High-quality multilingual semantic search, retrieval-augmented generation, enterprise knowledge bases, similarity matching, and text classification.

Type Embedding
Context 32K
Reasoning 1/10
Speed 8/10
Status

Current and available through Tencent Cloud TokenHub and the Tencent Cloud embeddings API.

Input USD 0.084 per million input tokens internationally; RMB 0.6 per million input tokens on the China TokenHub pricing page.
Output Not applicable; the model is billed for text input tokens and returns embedding vectors.
View model →

Fast cross-modal image-text retrieval, multimodal semantic matching, and video search

Type Multimodal
Context 33K
Reasoning 1/10
Speed 8/10
Multimodal Image input Video input
Status

Current and available through Tencent Cloud TokenHub

Input USD 0.07 per million text-input tokens; USD 0.098 per million image-input tokens; USD 0.21 per million video-input tokens
View model →
Tencent AI logo
Kinfra-VL-Embedding

Kinfra-VL-Embedding-8b

High-precision multimodal retrieval, cross-modal image-text search, video search, and semantic matching across text, image, and video collections

Type Multimodal Embedding
Context 33K
Reasoning 1/10
Speed 6/10
Multimodal Image input Video input
Status

Current and available through Tencent Cloud TokenHub

Input USD 0.084 per million text-input tokens; USD 0.126 per million image-input tokens; USD 0.252 per million video-input tokens
View model →

Video, image, audio, and text understanding; multimedia summarization; content tagging; structural video analysis; object localization; enterprise media workflows

Type Multimodal
Context 128K
Reasoning 5/10
Speed 7/10
Multimodal Image input Audio input
Status

Currently accessible; scheduled for shutdown on October 15, 2026 at 00:00 Beijing time on Tencent Cloud TokenHub and ADP

Input ¥1.2 per million tokens
Output ¥3.5 per million tokens
View model →
Tencent AI logo
WAND-Dubbing-Clone

WAND-Dubbing-Clone-V1

Multilingual video translation, voice-preserving dubbing, subtitle translation, online courses, films, and short-form video localization.

Type Other
Reasoning 2/10
Speed 5/10
Multimodal Audio input Video input
Status

Available

Input CNY 10 per 1 million tokens; usage is resolution-dependent and billed according to video duration and token consumption.
View model →
Tencent AI logo
WAND-Dubbing-Clone

WAND-Dubbing-Clone-v2

Multilingual video localization, translated online courses, short-form video dubbing, film and media localization, and long-video voice-preserving translation

Type Other
Reasoning 2/10
Speed 6/10
Multimodal Video input Media output
Status

Current and available through Tencent Cloud TokenHub as an asynchronous AI dubbing model

Input ¥0.2061 per second for 720p or lower; ¥0.2311 per second at 1080p; ¥0.2811 per second at 2K or higher. TokenHub lists the underlying rate as ¥10 per million tokens.
View model →

Fast everyday image creation, social-media graphics, marketing materials, short-video covers, and reference-guided image generation

Type Multimodal
Reasoning 1/10
Speed 8/10
Multimodal Image input Media output
Status

Current and available through Tencent Cloud TokenHub

Input 10 CNY per million tokens; input images are free for the Flash tier
Output Reference consumption: 45,000 tokens per 1K image (0.45 CNY), 67,500 tokens per 2K image (0.675 CNY), and 100,800 tokens per 4K image (1.008 CNY)
View model →

Low-cost, high-volume text-to-image and reference-to-image generation, including e-commerce product imagery, batch visual assets, and marketing content

Type Other
Reasoning 1/10
Speed 8/10
Multimodal Image input Media output
Status

Currently available managed image-generation model

Input 10 CNY per 1 million tokens; reference pricing is approximately 0.162 CNY per 1K image, 0.18 CNY per 2K image, and 0.225 CNY per 4K image
Output Approximately 0.162 CNY per 1K image, 0.18 CNY per 2K image, and 0.225 CNY per 4K image
View model →
Tencent AI logo
WAND-Vega-Image 1.0

WAND-Vega-Image-1.0-Pro

Professional text-to-image and reference-to-image generation, brand visuals, refined product imagery, high-quality design assets, and 1K-to-4K creative production.

Type Image Generation
Speed 7/10
Multimodal Image input Media output
Status

Current and available through Tencent Cloud TokenHub

Input Reference images: first 3 images free; from the 4th image, 0.10 CNY per image
Output 1K: 0.95 CNY per image; 2K: 0.95 CNY per image; 4K: 1.71 CNY per image
View model →

Fast, cost-conscious generation of short e-commerce, advertising, social media, and reference-guided video assets

Type Lightweight
Speed 8/10
Multimodal Image input Audio input
Status

Current and accessible through Tencent Cloud TokenHub

View model →
Tencent AI logo
WAND ASR

WAND-ASR-v1

Audio and video transcription, subtitle generation, sentence-level timestamps, and short-form speech-to-text workflows

Type Other
Speed 7/10
Multimodal Audio input Video input
Status

Current and available through Tencent Cloud TokenHub

Input 10 CNY per 1 million tokens; reference usage is 50 tokens per second, approximately 0.0005 CNY per second
View model →

Multi-view and video-to-3D reconstruction, camera and depth estimation, point-map prediction, surface-normal estimation, novel-view synthesis, and 3D Gaussian Splatting workflows

Type Other
Reasoning 2/10
Speed 7/10
Multimodal Image input Video input
Status

Legacy but publicly available open-weight model; superseded by WorldMirror-2.0 in Tencent's HY-World 2.0 catalog

View model →
Tencent AI logo

WAND-Dubbing-STS-v1

Reasoning 0/10
Speed 0/10