Model Hub

Browse PQC-verified AI models, datasets, and tools

X
xingyang1/Distill-Any-Depth-Large-hf HF Unverified

Depth-EstimationTransformersSafetensorsDepth_anythingDistill-Any-DepthVision HIGH
OpenGVLab/GUI-Odyssey HF Unverified

Dataset Card for GUI Odyssey News⭐️ A new and improved version of the GUIOdyssey dataset has been released! πŸŽ‰πŸŽ‰ πŸ‘‰ Please use the latest version and refer to the updated README for the most up-to-date information. We highly recommend using the new version for all training and evaluation! Repository: https://github.com/OpenGVLab/GUI-Odyssey Latest Version of Dataset: hflqf88888/GUIOdyssey Paper: https://arxiv.org/pdf/2406.08451 Introduction GUI Odyssey is… See the full description on the dataset page: https://huggingface.co/datasets/OpenGVLab/GUI-Odyssey.

Language:enSize_categories:1K<n<10KFormat:jsonModality:imageModality:tabularModality:text
PleIAs/common_corpus HF Unverified

Common Corpus Full paper - ICLR 2026 oral Common Corpus is the largest open and permissible licensed text dataset, comprising 2.27 trillion tokens (2,267,302,720,836 tokens). It is a diverse dataset, consisting of books, newspapers, scientific articles, government and legal documents, code, and more. Common Corpus has been created by Pleias in association with several partners. Common Corpus differs from existing open datasets in that it is: Truly Open: contains only data that… See the full description on the dataset page: https://huggingface.co/datasets/PleIAs/common_corpus.

Language:enLanguage:frLanguage:deLanguage:zhLanguage:itLanguage:es
S
segmind/small-sd HF Unverified

Text-to-ImageDiffusersStable-DiffusionStable-Diffusion-DiffusersBase_model:SG161222/Realistic_Vision_V4.0_noVAEBase_model:finetune:SG161222/Realistic_Vision_V4.0_noVAE HIGH
U
unsloth/FLUX.2-klein-4B-GGUF HF Unverified

Image-To-ImageGgmlGGUFText-to-ImageUnslothImage-Editing HIGH
D
deepset/bert-large-uncased-whole-word-masking-squad2 HF Unverified

Question AnsweringTransformersPyTorchTfJAXSafetensors HIGH
M
MoritzLaurer/deberta-v3-large-zeroshot-v2.0 HF Unverified

Zero-Shot ClassificationTransformersONNXSafetensorsDeberta-V2Text Classification HIGH
J
John6666/prefect-illustrious-xl-v3-sdxl HF PQC Verified

Text-to-ImageDiffusersSafetensorsStable-DiffusionStable-Diffusion-XlAnime HIGH
D
deepset/tinyroberta-squad2 HF Unverified

Question AnsweringTransformersPyTorchSafetensorsRobertaModel-Index MEDIUM
N
nvidia/segformer-b1-finetuned-ade-512-512 HF Unverified

Image-SegmentationTransformersPyTorchTfSegformerVision MEDIUM
B
black-forest-labs/FLUX.2-small-decoder HF Unverified

Image-To-ImageDiffusersSafetensorsText-to-ImageImage-EditingFlux MEDIUM
J
John6666/amanatsu-illustrious-v11-sdxl HF PQC Verified

Text-to-ImageDiffusersSafetensorsStable-DiffusionStable-Diffusion-XlAnime HIGH
K
krea/Krea-2-Turbo HF Unverified

Text-to-ImageDiffusersSafetensorsBase_model:krea/Krea-2-RawBase_model:finetune:krea/Krea-2-RawDiffusers:Krea2Pipeline CRITICAL
jhu-clsp/ettin-pretraining-data HF Unverified

Ettin Pre-training Data Phase 1 of 3: Diverse pre-training data mixture (1.7T tokens) used to train the Ettin model suite. This dataset contains the pre-training phase data used to train all Ettin encoder and decoder models. The data is provided in MDS format ready for use with Composer and the ModernBERT training repository. πŸ“Š Data Composition Data Source Tokens (B) Percentage Description DCLM 837.2 49.1% High-quality web crawl data CC Head 356.6… See the full description on the dataset page: https://huggingface.co/datasets/jhu-clsp/ettin-pretraining-data.

Task_categories:text-GenerationTask_categories:fill-MaskTask_categories:text-ClassificationLanguage:enPretrainingLanguage-Modeling
mlfoundations/MINT-1T-HTML HF Unverified

πŸƒ MINT-1T:Scaling Open-Source Multimodal Data by 10x: A Multimodal Dataset with One Trillion Tokens πŸƒ MINT-1T is an open-source Multimodal INTerleaved dataset with 1 trillion text tokens and 3.4 billion images, a 10x scale-up from existing open-source datasets. Additionally, we include previously untapped sources such as PDFs and ArXiv papers. πŸƒ MINT-1T is designed to facilitate research in multimodal pretraining. πŸƒ MINT-1T is created by a team from the University of Washington in… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations/MINT-1T-HTML.

Task_categories:image-To-TextTask_categories:text-GenerationLanguage:enSize_categories:100M<n<1BFormat:parquetModality:text
J
John6666/obsession-illustriousxl-v10-sdxl HF PQC Verified

Text-to-ImageDiffusersSafetensorsStable-DiffusionStable-Diffusion-XlAnime HIGH
N
nvidia/LocateAnything-3B HF Unverified

Image-Text-to-TextTransformersSafetensorsLocateanythingImage-Feature-ExtractionNvidia HIGH
T
TahaDouaji/detr-doc-table-detection HF PQC Verified

Object-DetectionTransformersPyTorchONNXSafetensorsDetr MEDIUM
C
cagliostrolab/animagine-xl-3.1 HF PQC Verified

Text-to-ImageDiffusersSafetensorsStable-DiffusionStable-Diffusion-XlBase_model:cagliostrolab/animagine-Xl-3.0 CRITICAL
V
valhalla/distilbart-mnli-12-1 HF Unverified

Zero-Shot ClassificationTransformersPyTorchJAXBartText Classification HIGH
Showing 20 of 806 items (page 26 of 41)