Model catalog

SenseTime Models

Browse the AI models associated with SenseTime. Compare current and historical models by family, capabilities, context window, availability and intended use.

12 models tracked
12 Total models
10 Model families
4 Model types
12 Current / accessible
All models

SenseTime model catalog

Audio-driven digital humans, lip-sync video, singing avatars, multilingual character animation, multi-person dialogue, and long-duration talking-video generation

Type Video Generation
Reasoning 1/10
Speed 9/10
Multimodal Image input Audio input
Status

Current; ongoing project

View model →
SenseTime logo
SenseNova-MARS

SenseNova-MARS-8B

Visual question answering, high-resolution image understanding, multimodal search, image-grounded research, and tool-assisted agentic reasoning

Type Multimodal
Context 262K
Reasoning 8/10
Speed 7/10
Multimodal Image input Tool use
Status

Current open-weight research model

View model →
SenseTime logo
SenseNova-MARS

SenseNova-MARS-32B

Visual deep-search, fine-grained image understanding, multimodal agent research, and self-hosted tool-using vision-language applications

Type Multimodal
Context 262K
Reasoning 8/10
Speed 4/10
Multimodal Image input Video input
Status

Current open-weight model

View model →

Spatial intelligence research, image-text question answering, visual reasoning, 3D scene analysis, and solid-geometry problems

Type Multimodal
Context 33K
Reasoning 7/10
Speed 6/10
Multimodal Image input
Status

Current open-weight release

View model →

Native image generation, high-resolution visual creation, image editing, infographic and layout generation, visual understanding, and multimodal research or creative workflows

Type Multimodal
Reasoning 6/10
Speed 4/10
Multimodal Image input Media output
Status

Current open-weight flagship checkpoint

Input No official hosted API price verified; downloadable model weights are available
Output No official hosted API price verified; downloadable model weights are available
View model →
SenseTime logo
SenseNova-Vision

SenseNova-Vision-7B-MoT

Unified computer-vision research, detection, OCR, segmentation, depth and normal estimation, visual grounding, and multi-view geometry

Type Multimodal
Reasoning 7/10
Speed 4/10
Multimodal Image input Media output
Status

Current open-weight research model

View model →

Long-horizon multimodal agents, data analysis, deep research, complex information presentation, office automation, and tool-driven workflows

Type Multimodal
Reasoning 8/10
Speed 9/10
Multimodal Image input Video input
Status

Preview; currently available through the SenseNova Token Plan

Input Free during the public beta Token Plan; standard per-token API pricing is not publicly verified
Output Free during the public beta Token Plan; standard per-token API pricing is not publicly verified
View model →
SenseTime logo
SenseNova Seko

SekoIDX

Character-consistent image generation for multi-episode videos, motion comics, short dramas, storyboards, and cross-shot visual production

Type Image Generation
Reasoning 2/10
Multimodal Image input Media output
Status

Current; integrated into the SenseNova Seko series and Seko 2.0 platform

View model →
SenseTime logo
SenseNova U

SenseNova U1 Pro

Professional image creation, infographics, advertising, e-commerce assets, presentations, educational diagrams, storyboards, and multi-step visual delivery workflows

Type Multimodal
Reasoning 8/10
Speed 5/10
Multimodal Image input Media output
Status

Current production release; enterprise API available through whitelist access

View model →
SenseTime logo
SenseNova U1

SenseNova U1

Open-source visual understanding, image generation, image editing, infographic creation, visual reasoning, and continuous image-text workflows.

Type Multimodal
Reasoning 7/10
Speed 7/10
Multimodal Image input Media output
Status

Open-source model series; U1-8B-MoT and U1-A3B-MoT variants are available. SenseNova U1.5 is the newer successor for visual creation and editing, while U1 checkpoints remain documented and downloadable.

View model →
SenseTime logo
SenseNova U1

SenseNova U1 Fast

Fast infographic generation, dense visual explanations, charts, diagrams, presentation graphics, and information-heavy layouts

Type Lightweight
Speed 8/10
Media output
Status

Current and accessible through SenseNova Token Plan and dedicated image-generation integrations

Input Free public beta with usage quota; commercial pricing not verified
Output Free public beta with usage quota; commercial pricing not verified
View model →
SenseTime logo
SenseNova U1.5

SenseNova U1.5 Lite

Text-to-image generation, reference-image creation, visual design, posters, infographics, product imagery, and iterative image editing

Type Multimodal
Reasoning 5/10
Speed 8/10
Multimodal Image input Media output
Status

Current and available through the SenseNova Token Plan

View model →