Seed GR-RL
Long-horizon, high-precision dexterous robot manipulation and real-world VLA policy specialization
Current research model/framework; public API availability not documented
Browse the AI models associated with ByteDance Seed. Compare current and historical models by family, capabilities, context window, availability and intended use.
Long-horizon, high-precision dexterous robot manipulation and real-world VLA policy specialization
Current research model/framework; public API availability not documented
Protein and biomolecular complex structure prediction, computational biology research, molecular design workflows, and self-hosted scientific inference
Active open-source biomolecular structure prediction project with multiple model variants, including Protenix-v1 and Protenix-v2
Multimodal agent workflows, search and information retrieval, coding agents, GUI interaction, image and video understanding, complex instruction following, and long-context business tasks.
Beta; accessible through BytePlus ModelArk as seed-1-8-251228. The separate LAS multimodal deep-thinking operator using Seed1.8 ended service on 2026-09-20.
Self-hosted general language modeling, long-context research, reasoning experiments, coding assistance, summarization, and foundation-model fine-tuning
Available open-weight model
Self-hosted long-context reasoning, coding, agentic tool use, research, question answering, summarization, and general text generation
Current open-weight model; downloadable under the Apache-2.0 license
Research, custom post-training, long-context processing, coding experiments, and self-hosted foundation-model deployment
Current open-weight downloadable foundation model
General-purpose Chinese and multilingual assistance, coding, reasoning, image and document understanding, and voice-interaction applications.
Retired; the primary doubao-1-5-pro-32k-250115 deployment was scheduled to shut down on 2026-09-21 at 14:00 China Standard Time.
Visual reasoning, image and video understanding, OCR, visual grounding, GUI-agent research, gameplay analysis, and multimodal benchmark evaluation
Retired; Volcano Engine service ended on 2026-03-31
Cross-modal semantic search, text-image retrieval, video retrieval, multimodal knowledge bases, classification, clustering, and recommendation
Current and available through Volcano Engine
Multimodal document analysis, visual question answering, coding, mathematics, general reasoning, long-context analysis and adaptive-thinking applications
Deprecated; new endpoint creation stopped on 2026-09-24; existing service scheduled for automatic migration or replacement on 2026-11-24
Deep reasoning, coding, mathematics, logical analysis, document understanding, and visual reasoning over images or videos
Deprecated; scheduled for retirement and migration to a Seed2.0 model
Cost-conscious production applications requiring long-context multimodal understanding, document and video analysis, coding assistance, tool use, GUI automation, and structured extraction.
Current; latest documented release seed-2-0-lite-260428
High-concurrency inference, batch generation, classification, extraction, summarization, and cost-sensitive multimodal workloads
Current; available through ByteDance's Volcano Engine model API
Complex multimodal reasoning, long-chain agent workflows, visual and video analysis, document understanding, scientific research support, coding, and enterprise automation.
Active and currently accessible through Volcano Engine Ark; canonical deployment version 260215
Complex agent workflows, high-value office and research tasks, long-horizon coding, document and visual analysis, video understanding, and tool-enabled productivity automation.
Current and officially released; available through ByteDance Seed, Doubao, and Volcano Engine API channels.
Fast multimodal agents, coding assistants, document and video analysis, tool-calling workflows, structured data extraction, and cost-sensitive production applications
Current and available through Volcano Engine Ark; API model identifier doubao-seed-2-1-turbo-260628
Long-form audiovisual storytelling, text-to-video, reference-based video generation, creative production, advertising, education, industrial simulation, and video editing.
Current; available through ByteDance platforms including Jimeng AI, Doubao Pro, and the Seed platform. API access through BytePlus ModelArk was announced as forthcoming.
Multimodal text-to-video and reference-based video creation, cinematic short clips, video editing and extension, multi-shot storytelling, and synchronized audio-video production.
Current official model page remains available; older generation superseded by Seedance 2.5
Full-scene audio creation, expressive voice generation, dubbing, dialogue, sound effects, ambience, advertising, games, podcasts, and multilingual audio production
Currently available through BytePlus
High-throughput code generation research, diffusion-language-model evaluation, and code-editing experiments
Experimental research preview
Embodied robotics research, long-horizon manipulation, bimanual control, dexterous object handling, and adapting robot policies to new objects and tasks
Current officially documented research model; no public hosted API, commercial pricing, or downloadable weights verified
Real-time audio-visual assistants, scene-aware guidance, live explanation, interactive learning, accessibility, and proactive multimodal collaboration
Current; fully rolled out for large-scale audio-visual full-duplex deployment
Professional image generation and editing, high-density infographics, advertising and e-commerce creative, multilingual visual content, precise spatial edits, layer separation, and multi-reference image workflows
Current and publicly documented; available through BytePlus ModelArk and ByteDance Seed services
Open-weight computer-use research, GUI grounding, browser automation prototypes, screenshot-based interface interaction, and visual action-model experimentation
Available open-weight research release; superseded by UI-TARS-2 for the provider's newer UI-TARS development line