Canopy Height Maps v2
Forest mapping, canopy-height estimation, restoration monitoring, ecological analysis, and geospatial research using satellite imagery
Current; open-source research model with gated model-weight access
Browse the AI models associated with Meta AI. Compare current and historical models by family, capabilities, context window, availability and intended use.
Forest mapping, canopy-height estimation, restoration monitoring, ecological analysis, and geospatial research using satellite imagery
Current; open-source research model with gated model-weight access
Image embeddings, dense feature extraction, image retrieval, classification, segmentation, depth estimation, object discovery, video tracking pipelines, and geospatial computer vision
Current; downloadable open-weight research model suite
Local inference, fine-tuning, private deployment, text generation, retrieval-augmented generation, research, and cost-sensitive applications
Current open-weight static model; downloadable and usable through compatible local or hosted inference deployments
High-quality open-weight research, multilingual applications, coding, reasoning, synthetic-data generation, model distillation and self-hosted deployments with substantial infrastructure
Available as an open-weight model; older generation superseded by newer Llama releases
Private local inference, mobile and edge assistants, summarization, rewriting, retrieval-supported generation, and lightweight multilingual applications
Available downloadable open-weight model; static checkpoint
Private local inference, multilingual text generation, edge applications, model adaptation, and fine-tuning
Available; static pretrained open-weight model
Visual question answering, image reasoning, chart and document understanding, image captioning, multimodal research, and self-hosted or partner-hosted AI applications.
Available open-weight model; static model trained on an offline dataset
Self-hosted visual question answering, image captioning, document analysis, visual reasoning, and multimodal assistants
Available open-weight static model
Self-hosted or hosted multilingual chat, coding assistance, long-context text generation, tool calling, synthetic data, and applications requiring open model weights
Available open-weight model; static model trained on an offline dataset
Open-weight multimodal assistants, image understanding, visual question answering, coding, multilingual applications, creative writing and long-context text processing.
Available; open-weight static checkpoint
Long-context document and code analysis, visual question answering, multimodal assistants, multilingual applications, self-hosted inference, and customized deployments
Available; open-weight model
Text and image moderation for prompts and generated responses in generative AI systems
Current and available as an open-weight model
Low-cost prompt and response safety classification, local moderation, mobile and edge deployments, and customizable LLM guardrails
Current open-weight model; downloadable subject to access approval
Self-hosted multilingual input and output moderation for LLM applications
Available open-weight safety classifier
Safety classification of mixed text-and-image prompts and text responses in multimodal LLM systems
Available open-weight multimodal safety model; Meta's current model repositories continue to list it.
Low-latency detection of prompt injections and jailbreak attempts in LLM applications, agents, retrieved documents, and other untrusted text
Current open-weight safety classifier
Multilingual prompt-injection detection, jailbreak screening, agent security, and filtering untrusted text before it reaches an LLM
Current open-weight model; gated download access
Local agents, long-running tool workflows, coding assistants, multimodal document and screenshot understanding, private on-device inference, and model customization
Current open-weight model; self-hosted and available through selected third-party hosted inference providers
Text-to-image generation, precise image editing, multi-image composition, anchored visual series, product imagery, creative assets, and grounded visual content.
Current and available through Meta Model API; also available in selected Meta AI consumer experiences
Agentic workflows, coding agents, computer-use automation, multimodal document and media analysis, long-context reasoning, tool orchestration, and web-grounded applications.
Current but superseded by Muse Spark 1.2 and Muse Spark 1.3; available on the Meta Model API Standard tier in public preview for US developers.
Long-horizon coding agents, repository-scale software engineering, multimodal code generation, debugging, refactoring, and tool-driven workflows
Available; previous version, with Muse Spark 1.3 recommended for new work
Long-horizon coding agents, software engineering, browser and computer-use workflows, tool orchestration, large repositories, document analysis, and multimodal reasoning
Current; available through Meta Model API and Muse Code
Real-time speech-to-text, live captions, voice agents, meeting transcription, call intelligence, dictation, and speaker-aware transcription
Current and available through Meta Model API
Multilingual speech representation learning, audio embeddings, low-resource language research, and custom downstream speech systems
Current; open-source research model family
Cross-modal audio-video-text retrieval, audiovisual embeddings, sound-event understanding, media indexing, and multimodal perception systems.
Current and openly available
Single-image 3D human mesh recovery, pose and shape estimation, AR/VR, robotics perception, and computer-vision research
Current research release; downloadable checkpoints with gated Hugging Face access
Single-image reconstruction of textured 3D objects from natural scenes, 3D computer-vision research, Gaussian-splat workflows, and rapid asset prototyping
Available research release; gated model checkpoints
Prompted audio separation, speech and noise isolation, instrument and vocal extraction, audiovisual sound segmentation, and audio-editing research
Current open research release; downloadable checkpoints and public demo available
Reference-free evaluation and benchmarking of text-guided audio-separation outputs
Current gated research release
Noncommercial research on expressive multilingual speech-to-speech translation, prosody transfer and voice-style preservation
Available as a gated research release
Real-time multilingual speech recognition, simultaneous translation, speech-to-text translation, and speech-to-speech translation
Open-source research model; publicly available
Multilingual automatic speech recognition, speech-to-text translation, text translation, text-to-speech translation, and speech-to-speech translation
Available open-weight research model; noncommercial research use
Open-vocabulary object detection, pixel-level image segmentation, and multi-object video tracking
Current and available; hosted through Meta Model API and available as released research checkpoints
Computational neuroscience, fMRI response prediction, brain encoding, multisensory research, and in-silico experiment design
Open-weight research release