Aleph-Alpha-GermanWeb-Quality-Classifier-fastText
Fast German-language document quality classification and large-scale web-data filtering
Available; open-weight research model; not deployed by a Hugging Face Inference Provider
Browse the AI models associated with Aleph Alpha. Compare current and historical models by family, capabilities, context window, availability and intended use.
Fast German-language document quality classification and large-scale web-data filtering
Available; open-weight research model; not deployed by a Hugging Face Inference Provider
Fast local classification of German documents for grammar-oriented corpus filtering and data curation
Available as an open-weight model repository; not deployed by Hugging Face Inference Providers
German grammar-related text-quality filtering, dataset curation, and binary classification of web text
Available as an open-weight Hugging Face model
German web-document quality scoring, dataset filtering, ranking, and pretraining-data curation
Available open-weight model
Research on tokenizer-free language modeling, English-German text generation, multilingual NLP, open-weight deployment, and custom model adaptation
Available as open-weight research software through Hugging Face; not deployed by an inference provider
English and German text generation, instruction-following research, tokenizer-free language-model research, and self-hosted non-commercial experimentation
Available as an open-weight research model
English and German instruction following, tokenizer-free language-model research, text-compression experiments, and self-hosted open-weight deployments
Available open-weight research release
Historical multilingual text completion, language-model research, and semantic representation workflows
Legacy; current public availability is not clearly documented
Legacy multilingual text generation, conversational applications, and explainability-oriented workflows using Aleph Alpha's first Luminous generation
Legacy; current public availability not verified
Semantic search, information retrieval, query-document matching, clustering, classification, similarity scoring, and text feature extraction
Legacy; current API availability unverified
Historical multilingual text-completion research and general language-task experiments
Historical or legacy; current availability not verified
Steerable multilingual text generation, summarization, classification, question answering, and explainability-oriented enterprise workflows
Available; first-generation Luminous control model
Historical multilingual text completion, language understanding research, and compatibility work involving Aleph Alpha’s original Luminous API
Legacy generation; historical API and Playground availability documented, current public availability unverified
Zero-shot multilingual text generation, instruction following, classification, conversational prototypes, and explainability-oriented enterprise workflows.
Legacy model; documented in Aleph Alpha SDK materials, with current endpoint availability dependent on provider access and deployment status.
Vision-language research, image captioning, visual question answering, and experiments with adapter-based multimodal fine-tuning
Legacy research/demo model; publicly released checkpoint and source code remain available, but it is not documented as a current hosted commercial model
Multilingual information retrieval, semantic search, reranking, clustering, and instruction-guided text embeddings
Available as downloadable open-weight model; customer and on-premises deployment options documented
Compact vector representations for semantic search, information retrieval, reranking, clustering, and similarity-based classification
Available open-weight checkpoint; no hosted inference deployment or provider API pricing was verified
Concise multilingual generation, summarization, extraction, domain-specific text workflows, and self-hosted deployment
Available as an open-weight model; also referenced in Aleph Alpha API tooling
Multilingual text generation, classification, summarization, question answering, engineering and automotive applications, and safety-conscious research deployments
Available open-weight release; safety-aligned variant
English and German instruction following, multilingual research, tokenizer-free language-model experimentation, and self-hosted deployment
Available as downloadable open-weight research software
English and German instruction following, multilingual text generation, German-language applications, and research into tokenizer-free language models
Current open-weight research release; downloadable from Hugging Face
Research on tokenizer-free language modeling, English and German text generation, long-context experiments, and self-hosted non-commercial applications
Current publicly downloadable open-weight research model