Model catalog

01.AI Models

Browse the AI models associated with 01.AI. Compare current and historical models by family, capabilities, context window, availability and intended use.

26 models tracked
26 Total models
5 Model families
4 Model types
24 Current / accessible
All models

01.AI model catalog

Local bilingual text generation, research, fine-tuning, offline applications, and resource-conscious deployment

Type General Purpose
Context 4K
Reasoning 5/10
Speed 7/10
Fine-tuning
Status

Available open-weight model; older first-generation Yi base model

View model →

Long-document completion, local deployment, English-Chinese text generation, research, and downstream fine-tuning

Type General Purpose
Context 200K
Reasoning 4/10
Speed 6/10
Fine-tuning Streaming
Status

Available as downloadable open weights; legacy-generation model

Input No official hosted API price verified; downloadable weights
Output No official hosted API price verified; downloadable weights
View model →

Local bilingual English-Chinese chat, personal projects, academic experimentation, and fine-tuning on modest open-weight infrastructure

Type General Purpose
Context 4K
Reasoning 5/10
Speed 7/10
Fine-tuning
Status

Open-weight and downloadable; legacy historical model

Input No official hosted API price verified; self-hosted weights have no per-token provider charge
Output No official hosted API price verified; self-hosted weights have no per-token provider charge
View model →

Local text generation, code completion, mathematics, bilingual English-Chinese applications, research, and downstream fine-tuning

Type General Purpose
Context 4K
Reasoning 6/10
Speed 7/10
Fine-tuning
Status

Available as an open-weight downloadable model; older Yi-generation model with no verified first-party hosted API offering for this exact model

View model →

Long-context document processing, code generation, mathematics, bilingual English-Chinese text generation, local inference, and domain fine-tuning

Type General Purpose
Context 200K
Reasoning 5/10
Speed 6/10
Fine-tuning Streaming
Status

Available open-weight base model; legacy generation but still publicly downloadable

Input No official hosted API price; downloadable weights
Output No official hosted API price; downloadable weights
View model →

Self-hosted bilingual text generation, research, custom fine-tuning, coding experiments, and English-Chinese applications

Type General Purpose
Context 4K
Reasoning 6/10
Speed 4/10
Fine-tuning Streaming
Status

Available open-weight model; older-generation base checkpoint

Input No official hosted API price verified; self-hosted weights
Output No official hosted API price verified; self-hosted weights
View model →

Long-document analysis, English-Chinese generation, retrieval experiments, research, and self-hosted fine-tuning

Type General Purpose
Context 200K
Reasoning 7/10
Speed 4/10
Fine-tuning
Status

Available as downloadable open-weight model; no official hosted API availability or retirement date verified

View model →

Self-hosted bilingual assistants, English-Chinese dialogue, open-weight LLM research, private inference, quantization, and fine-tuning

Type General Purpose
Context 4K
Reasoning 6/10
Speed 3/10
Fine-tuning Streaming
Status

Open-weight; downloadable; legacy-generation model with no verified provider-managed hosted API availability

View model →

Long-context chat, complex text analysis, multilingual generation, prediction, and general-purpose enterprise language applications.

Type General Purpose
Context 32K
Reasoning 7/10
Speed 6/10
Streaming
Status

Available through 01.AI API documentation; current lifecycle details are not explicitly stated by the provider.

Input $3 per 1 million input tokens
Output $3 per 1 million output tokens
View model →

Cost-sensitive hosted chat, Chinese-English generation, coding, mathematics, reasoning, summarization and high-volume API workloads

Type General Purpose
Context 16K
Reasoning 8/10
Speed 9/10
Streaming
Status

Proprietary hosted API model; current public availability and lifecycle status are not clearly documented

Input $0.14 per 1 million input tokens historically reported; verify current pricing with 01.AI
Output $0.14 per 1 million output tokens historically reported; verify current pricing with 01.AI
View model →

Local text generation, self-hosted applications, experimentation, domain adaptation, and fine-tuning

Type Lightweight
Context 4K
Reasoning 5/10
Speed 7/10
Fine-tuning Streaming
Status

Available as an open-weight downloadable model for self-hosted and local inference; no hosted inference provider is currently listed on its Hugging Face model page.

View model →

Local conversational applications, instruction following, lightweight coding, bilingual English-Chinese text generation, experimentation, and self-hosted inference

Type General Purpose
Context 4K
Reasoning 4/10
Speed 7/10
Fine-tuning Streaming
Status

Open-weight; accessible for local deployment; no explicit retirement date found

View model →

Local text generation, research, fine-tuning, Chinese-English applications, and cost-sensitive self-hosted deployments

Type General Purpose
Context 4K
Reasoning 6/10
Speed 7/10
Fine-tuning
Status

Open-weight and downloadable; no official deprecation or shutdown date found

Input No official first-party hosted price; downloadable weights under Apache 2.0
Output No official first-party hosted price; downloadable weights under Apache 2.0
View model →

Self-hosted conversational assistants, local text generation, lightweight coding help, multilingual experimentation, and fine-tuning research.

Type General Purpose
Context 4K
Reasoning 6/10
Speed 8/10
Fine-tuning Streaming
Status

Available open-weight model; official weights remain accessible, with no exact provider-published deprecation or shutdown date verified.

Input No official hosted API price verified; self-hosted weights are available under Apache 2.0.
Output No official hosted API price verified; self-hosted weights are available under Apache 2.0.
View model →

Local bilingual text generation, research, domain adaptation, and fine-tuning

Type General Purpose
Context 4K
Reasoning 7/10
Speed 4/10
Fine-tuning
Status

Open-weight and currently downloadable; standard 4K-context base checkpoint

View model →

Self-hosted bilingual assistants, English-Chinese text generation, coding support, mathematics, reasoning, fine-tuning, and privacy-sensitive deployments

Type General Purpose
Context 4K
Reasoning 7/10
Speed 4/10
Fine-tuning Streaming
Status

Open-weight model remains available for self-hosted deployment; 01.AI hosted model-platform API service ended on 2026-09-03

View model →

Local code completion, code generation, multilingual programming tasks, long-context source-code analysis, and downstream fine-tuning

Type Coding
Context 131K
Reasoning 3/10
Speed 8/10
Fine-tuning Streaming
Status

Available as downloadable open-weight model; no exact first-party hosted API availability verified

Input No official first-party hosted API price identified; downloadable weights are available under Apache 2.0
Output No official first-party hosted API price identified; downloadable weights are available under Apache 2.0
View model →

Local code generation, completion, debugging, code explanation, lightweight IDE assistants, and model fine-tuning experiments

Type Coding
Context 131K
Reasoning 3/10
Speed 8/10
Fine-tuning Streaming
Status

Available open-weight model

Input No official hosted API pricing; model weights are openly available
Output No official hosted API pricing; model weights are openly available
View model →
01.AI logo
Yi-Coder

Yi-Coder-9B

Local code generation, completion, editing, repository-scale context, multilingual programming, and fine-tuning

Type Coding
Context 131K
Reasoning 6/10
Speed 7/10
Fine-tuning
Status

Available open-weight base model

Input No official hosted API price; downloadable weights for self-hosting or third-party deployment
Output No official hosted API price; downloadable weights for self-hosting or third-party deployment
View model →

Self-hosted coding assistance, code generation, debugging, code explanation, code translation, and long-context repository analysis

Type Coding
Context 131K
Reasoning 6/10
Speed 6/10
Fine-tuning Streaming
Status

Open-weight and publicly available; no verified official hosted API listing for this exact model

Input No official hosted API price verified; downloadable weights are available under Apache 2.0
Output No official hosted API price verified; self-hosting and third-party inference costs vary
View model →

Local bilingual image understanding, visual question answering, OCR-oriented image analysis, and image-to-text applications

Type Multimodal
Context 4K
Reasoning 5/10
Speed 5/10
Multimodal Image input Fine-tuning
Status

Open-weight multimodal model family; current accessibility and active maintenance are not clearly documented by 01.AI

View model →

Local bilingual image understanding, visual question answering, OCR-assisted extraction, image summarization, and lightweight multimodal experimentation

Type Multimodal
Context 4K
Reasoning 4/10
Speed 7/10
Multimodal Image input
Status

Available as an open-weight self-hosted model; no current official hosted inference provider is listed on its model page

View model →

Self-hosted bilingual image understanding, visual question answering, image text recognition, and research applications with substantial GPU capacity

Type Multimodal
Context 4K
Reasoning 6/10
Speed 3/10
Multimodal Image input
Status

Available as open-weight downloadable model; no current first-party hosted API availability verified

View model →
01.AI logo
Yi Large

Yi Large FC

Function calling, tool selection, agent orchestration, and structured workflow automation

Type Coding
Context 33K
Reasoning 7/10
Speed 7/10
Tool use Streaming
Status

Documented by 01.AI; current live availability not independently verified

Input $3 per 1 million tokens
Output $3 per 1 million tokens
View model →

Low-cost short-context chat, text generation, summarization, classification, and general language applications

Type General Purpose
Context 4K
Reasoning 6/10
Speed 8/10
Streaming
Status

Documented in the 01.AI API catalog; current operational availability should be verified through the provider

Input $0.19 per 1 million tokens
Output $0.19 per 1 million tokens
View model →
01.AI logo

Yi-9B-Chat

Identity review and disambiguation only

Reasoning 0/10
Speed 0/10
Status

Unverified identifier; no official 01.AI model record found