Model catalog

Technology Innovation Institute (TII) Models

Browse the AI models associated with Technology Innovation Institute (TII). Compare current and historical models by family, capabilities, context window, availability and intended use.

39 models tracked
39 Total models
11 Model families
5 Model types
38 Current / accessible
All models

Technology Innovation Institute (TII) model catalog

Self-hosted text generation, research, quantization, and domain-specific fine-tuning

Type General Purpose
Context 2K
Reasoning 4/10
Speed 7/10
Fine-tuning
Status

Available open-weight model; older Falcon generation with newer successors

Input No official hosted API price; self-hosted model weights
Output No official hosted API price; self-hosted model weights
View model →

Research, self-hosted text generation, model fine-tuning, summarization, and general language-model experimentation

Type General Purpose
Context 2K
Reasoning 4/10
Speed 4/10
Fine-tuning Streaming
Status

Available as open weights; legacy-generation model

Input No official hosted API pricing; self-hosted open weights
Output No official hosted API pricing; self-hosted open weights
View model →

Research, self-hosted text generation, language-model evaluation, and domain-specific fine-tuning when substantial GPU infrastructure is available

Type General Purpose
Context 2K
Reasoning 6/10
Speed 3/10
Fine-tuning Streaming
Status

Available open-weight model; legacy-generation checkpoint

Input No official first-party hosted API pricing; self-hosted/open-weight distribution
Output No official first-party hosted API pricing; self-hosted/open-weight distribution
View model →

Memory-efficient local text generation, edge-device experimentation, research, and downstream fine-tuning

Type Lightweight
Context 33K
Reasoning 3/10
Speed 8/10
Fine-tuning
Status

Current open-weight model; downloadable from Hugging Face

Input No official hosted API pricing; self-hosted model weights
Output No official hosted API pricing; self-hosted model weights
View model →

Efficient local, edge, and resource-constrained English text generation; experimentation with BitNet models; full fine-tuning on the supplied prequantized revision.

Type Lightweight
Context 33K
Reasoning 3/10
Speed 8/10
Fine-tuning
Status

Current open-weight/downloadable model

View model →

Memory-efficient local text generation, edge deployment, BitNet research, continued pretraining, and custom fine-tuning

Type Lightweight
Context 33K
Reasoning 4/10
Speed 8/10
Fine-tuning Streaming
Status

Current open-weight model; publicly downloadable

View model →

Lightweight local text generation, continued pretraining, domain adaptation, language-model research, and edge-oriented deployments

Type Lightweight
Context 16K
Reasoning 3/10
Speed 8/10
Streaming
Status

Current open-weight model

View model →

Lightweight local chat, text generation, edge deployment, prototyping, and fine-tuning experiments

Type Lightweight
Context 16K
Reasoning 2/10
Speed 9/10
Fine-tuning Streaming
Status

Current; open-weight; instruction-tuned

View model →

Local multilingual text generation, long-context experimentation, research, and downstream fine-tuning

Type General Purpose
Context 131K
Reasoning 6/10
Speed 8/10
Status

Current open-weight model; available for download and self-hosted inference

View model →

Efficient local text generation, multilingual language modeling, long-context applications, model adaptation, and compact reasoning research.

Type General Purpose
Context 131K
Reasoning 6/10
Speed 8/10
Status

Current open-weight model

Input No official hosted API pricing; downloadable weights are available for self-hosted deployment.
Output No official hosted API pricing; downloadable weights are available for self-hosted deployment.
View model →

Efficient local inference, compact conversational assistants, multilingual text generation, long-context applications, and resource-constrained deployments

Type Lightweight
Context 131K
Reasoning 7/10
Speed 8/10
Streaming
Status

Current; open-weight

Input No official hosted API pricing; self-hosted/open-weight model
Output No official hosted API pricing; self-hosted/open-weight model
View model →

Efficient local or private deployment, multilingual instruction following, compact conversational applications, coding assistance, mathematics, and long-context text processing

Type Lightweight
Context 131K
Reasoning 5/10
Speed 8/10
Fine-tuning Streaming
Status

Current open-weight model; downloadable and usable for self-hosted inference

View model →

Local text generation, multilingual research, domain adaptation, fine-tuning, long-context experiments, and resource-conscious deployment.

Type General Purpose
Context 131K
Reasoning 4/10
Speed 8/10
Fine-tuning Streaming
Status

Current open-weight model; downloadable from Hugging Face and usable with supported local inference frameworks.

Input No official hosted API pricing; self-hosted model weights
Output No official hosted API pricing; self-hosted model weights
View model →

Local multilingual chat, lightweight instruction following, retrieval-augmented generation, efficient text generation, and cost-sensitive self-hosted deployments

Type General Purpose
Context 131K
Reasoning 5/10
Speed 8/10
Streaming
Status

Current open-weight model

View model →

Private or local multilingual text generation, long-context experimentation, domain adaptation, and fine-tuning from a 7B-scale open-weight foundation

Type General Purpose
Context 262K
Reasoning 6/10
Speed 8/10
Fine-tuning
Status

Current open-weight model; pretrained base checkpoint

Input No official hosted API pricing; downloadable open-weight checkpoint
Output No official hosted API pricing; deployment cost depends on self-hosted infrastructure
View model →

Self-hosted multilingual assistants, long-context text processing, coding, research, private deployments, and quantized local inference

Type General Purpose
Context 262K
Reasoning 7/10
Speed 8/10
Tool use
Status

Current open-weight model

View model →

Self-hosted multilingual text generation, long-context research, coding, RAG, foundation-model experimentation, and downstream fine-tuning

Type General Purpose
Context 262K
Reasoning 7/10
Speed 5/10
Fine-tuning
Status

Current open-weight model; available from the official Hugging Face repository

View model →

Long-context text generation, multilingual instruction following, document processing, retrieval-augmented generation, coding and self-hosted enterprise deployments

Type General Purpose
Context 262K
Reasoning 7/10
Speed 6/10
Fine-tuning Streaming
Status

Current open-weight model; publicly downloadable and usable with Transformers, vLLM and llama.cpp

Input No official hosted API price; self-hosted model weights
Output No official hosted API price; self-hosted model weights
View model →
Technology Innovation Institute (TII) logo
Falcon-H1-Arabic

Falcon-H1-Arabic

Arabic NLP, dialect-aware assistants, long-context document analysis, summarization, multilingual reasoning, and locally deployed applications

Type General Purpose
Reasoning 7/10
Speed 7/10
Status

Current; open model family

View model →

Ultra-lightweight local text generation, edge deployment, offline instruction following, rewriting, extraction, and embedded AI experiments

Type Lightweight
Context 262K
Reasoning 2/10
Speed 10/10
Fine-tuning Streaming
Status

Current open-weight model

Input No official hosted API price; downloadable weights
Output No official hosted API price; downloadable weights
View model →

Lightweight local Python generation, fill-in-the-middle completion, edge deployment, code education, and low-resource developer tools

Type Coding
Context 262K
Reasoning 2/10
Speed 9/10
Status

Current and publicly accessible open-weight model

View model →

Lightweight local text generation, instruction-following, edge devices, embedded applications, experimentation, and privacy-sensitive deployments.

Type Lightweight
Context 262K
Reasoning 2/10
Speed 9/10
Status

Current open-weight model

Input No official hosted API pricing found; model weights are available for local deployment under the Falcon-LLM License.
Output No official hosted API pricing found; local inference costs depend on hardware and deployment framework.
View model →

Lightweight local reasoning, edge deployment, offline assistants, experimentation, and compact text-generation applications

Type Reasoning
Context 262K
Reasoning 7/10
Speed 8/10
Streaming
Status

Current open-weight model; downloadable and available for self-hosted inference

View model →

Lightweight local function calling, edge automation, API argument generation, and resource-constrained assistants

Type Lightweight
Context 262K
Reasoning 2/10
Speed 9/10
Tool use Streaming
Status

Current downloadable open-weight model

View model →

Local reasoning experiments, edge deployment, offline assistants, privacy-sensitive text processing, and resource-constrained inference

Type Reasoning
Context 262K
Reasoning 4/10
Speed 9/10
Fine-tuning Streaming
Status

Available open-weight checkpoint

View model →
Technology Innovation Institute (TII) logo
Falcon-H1-Tiny-R

Falcon-H1-Tiny-R-90M

Ultra-lightweight local reasoning, embedded applications, edge deployment, experimentation, and low-memory text generation

Type Reasoning
Context 262K
Reasoning 3/10
Speed 9/10
Streaming
Status

Current open-weight model

View model →
Technology Innovation Institute (TII) logo
Falcon-H1R

Falcon-H1R-7B

Mathematical reasoning, programming, long-context analysis, local deployment, self-hosted inference, and test-time scaling

Type Reasoning
Context 262K
Reasoning 8/10
Speed 8/10
Streaming
Status

Current open-weight model

View model →
Technology Innovation Institute (TII) logo
Falcon 2

Falcon2-11B

Research, multilingual text generation, fine-tuning, quantization, and self-hosted inference

Type General Purpose
Context 8K
Reasoning 4/10
Speed 6/10
Fine-tuning Streaming
Status

Available as an open-weight pretrained base model

Input No official hosted API price; downloadable weights for self-hosted deployment
Output No official hosted API price; infrastructure-dependent
View model →

Local inference, multilingual text completion, research, continued pretraining, domain adaptation, and fine-tuning on constrained hardware

Type Lightweight
Context 4K
Reasoning 4/10
Speed 8/10
Fine-tuning Streaming
Status

Available open-weight pretrained base model

Input No official hosted API price; self-hosted/open-weight model
Output No official hosted API price; self-hosted/open-weight model
View model →

Lightweight local assistants, multilingual instruction following, extraction, classification, education, and resource-conscious deployments

Type Lightweight
Context 8K
Reasoning 4/10
Speed 8/10
Tool use Streaming
Status

Current open-weight model; downloadable from Hugging Face

Input No official first-party hosted API pricing
Output No official first-party hosted API pricing
View model →

Fine-tuning, multilingual text generation, research, compact local inference, and edge-oriented deployments

Type General Purpose
Context 8K
Reasoning 5/10
Speed 8/10
Fine-tuning Streaming
Status

Available open-weight model

Input No official hosted API price; downloadable weights for self-hosted deployment
Output No official hosted API price; downloadable weights for self-hosted deployment
View model →

Self-hosted chat assistants, multilingual text generation, lightweight coding and mathematics, local inference, and resource-conscious deployments

Type General Purpose
Context 32K
Reasoning 6/10
Speed 8/10
Tool use Fine-tuning Streaming
Status

Available as an open-weight model

View model →

Fine-tuning, multilingual text generation, code and mathematics experiments, research, and self-hosted inference

Type General Purpose
Context 33K
Reasoning 6/10
Speed 7/10
Fine-tuning
Status

Current open-weight model; downloadable and self-hostable

View model →

Self-hosted multilingual chat, instruction following, reasoning, mathematics, coding, research, and long-context text generation

Type General Purpose
Context 33K
Reasoning 7/10
Speed 7/10
Tool use Fine-tuning Streaming
Status

Available open-weight model

Input No official hosted API price; downloadable weights
Output No official hosted API price; downloadable weights
View model →

Fine-tuning, multilingual text generation, language-model research, code and mathematics experimentation, and self-hosted inference

Type General Purpose
Context 33K
Reasoning 7/10
Speed 6/10
Fine-tuning Streaming
Status

Available open-weight model

Input No official hosted API pricing; self-hosted/open-weight deployment
Output No official hosted API pricing; self-hosted/open-weight deployment
View model →

Self-hosted multilingual assistants, STEM and mathematics tasks, coding, instruction following, research, and local function-calling systems.

Type General Purpose
Context 33K
Reasoning 7/10
Speed 5/10
Tool use Fine-tuning Streaming
Status

Available open-weight model; official repository remains accessible. No official deprecation or shutdown date found.

Input No official hosted API pricing; downloadable weights are provided for self-managed inference.
Output No official hosted API pricing; downloadable weights are provided for self-managed inference.
View model →
Technology Innovation Institute (TII) logo
Falcon Mamba

Falcon Mamba 7B

Local English text generation, open-weight research, long-sequence experimentation, and memory-conscious inference

Type General Purpose
Context 8K
Reasoning 4/10
Speed 8/10
Fine-tuning
Status

Available for download; legacy or superseded by Falcon-H1-7B-Base

View model →
Technology Innovation Institute (TII) logo
Falcon Perception

Falcon OCR

Local or self-hosted document OCR, formula recognition, table extraction, receipts, invoices, papers, and layout-aware document parsing

Type Multimodal
Context 16K
Reasoning 3/10
Speed 8/10
Multimodal Image input Streaming
Status

Current open-weight release

View model →
Technology Innovation Institute (TII) logo
Falcon Perception

Falcon Perception

Natural-language object grounding, open-vocabulary detection, promptable instance segmentation, crowded-scene perception, robotics and visual inspection pipelines

Type Multimodal
Context 8K
Reasoning 2/10
Speed 8/10
Multimodal Image input Media output
Status

Current; open-weight; Apache-2.0 licensed

View model →