Model Hub

Browse PQC-verified AI models, datasets, and tools

P
PekingU/rtdetr_r101vd_coco_o365 HF Unverified

Object-DetectionTransformersSafetensorsRt_detrVisionEnglish MEDIUM
N
NousResearch/Hermes-3-Llama-3.1-8B HF Ollama PQC Verified

Nous Research's Hermes 3 built on Llama 3.1. Strong function calling, structured output, and agentic capabilities.

TransformerText Generation8BInstructAgentic HIGH
L
Lykon/dreamshaper-7 HF Unverified

Text-to-ImageDiffusersSafetensorsStable-DiffusionStable-Diffusion-DiffusersArt HIGH
L
lxyuan/distilbert-base-multilingual-cased-sentiments-student HF PQC Verified

Text ClassificationTransformersPyTorchONNXSafetensorsDistilbert HIGH
mlfoundations/MINT-1T-HTML HF Unverified

πŸƒ MINT-1T:Scaling Open-Source Multimodal Data by 10x: A Multimodal Dataset with One Trillion Tokens πŸƒ MINT-1T is an open-source Multimodal INTerleaved dataset with 1 trillion text tokens and 3.4 billion images, a 10x scale-up from existing open-source datasets. Additionally, we include previously untapped sources such as PDFs and ArXiv papers. πŸƒ MINT-1T is designed to facilitate research in multimodal pretraining. πŸƒ MINT-1T is created by a team from the University of Washington in… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations/MINT-1T-HTML.

Task_categories:image-To-TextTask_categories:text-GenerationLanguage:enSize_categories:100M<n<1BFormat:parquetModality:text
J
jakeBland/wav2vec-vm-finetune HF PQC Verified

Audio-ClassificationTransformersTensorboardSafetensorsWav2vec2Generated_from_trainer HIGH
F
facebook/nllb-200-distilled-600M HF PQC Verified

TranslationTransformersPyTorchM2m_100Text2text-GenerationNllb HIGH
F
FastVideo/FastVideo-FastH3-4-step-Preview-v1-VSA-DataFree HF Unverified

Text-To-VideoDiffusersSafetensorsVideoAudioText-To-Audio-Video CRITICAL
ksolovev/fine-news HF Unverified

FineNews FineNews is a multilingual news-text dataset for language-model research. It filters and deduplicates a 2021–2025 snapshot of the INFINI-NEWS Corpus, which extracts articles from Common Crawl CC-News. At a glance Measure FineNews Input articles 852,824,802 Infini-News rows from 2021–2025 Output 392,627,654 physical rows Files 294,509 Parquet files Folders 60 publication months (2021-01 to 2025-12), then language Language folders 129… See the full description on the dataset page: https://huggingface.co/datasets/ksolovev/fine-news.

Task_categories:text-GenerationSize_categories:100M<n<1BNewsJournalismMediaCommon-Crawl
M
MoritzLaurer/mDeBERTa-v3-base-xnli-multilingual-nli-2mil7 HF Unverified

Zero-Shot ClassificationTransformersPyTorchONNXSafetensorsDeberta-V2 HIGH
A
Abiray/MiniMax-H3-GGUF HF Unverified

Image-To-VideoGGUFComfyuiText-To-VideoImage-Text-To-VideoVideo-To-Video CRITICAL
A
amunchet/rorshark-vit-base HF Unverified

Image-ClassificationTransformersTensorboardSafetensorsVitVision MEDIUM
Z
ZhengPeng7/BiRefNet HF Unverified

Image-SegmentationBirefnetSafetensorsBackground-RemovalMask-GenerationDichotomous Image Segmentation MEDIUM
P
pyannote/segmentation HF PQC Verified

Voice-Activity-DetectionPyannote-AudioPyTorchPyannotePyannote-Audio-ModelAudio MEDIUM
nyu-mll/glue HF PQC Verified

Dataset Card for GLUE Dataset Summary GLUE, the General Language Understanding Evaluation benchmark (https://gluebenchmark.com/) is a collection of resources for training, evaluating, and analyzing natural language understanding systems. Supported Tasks and Leaderboards The leaderboard for the GLUE benchmark can be found at this address. It comprises the following tasks: ax A manually-curated evaluation dataset for fine-grained analysis of system… See the full description on the dataset page: https://huggingface.co/datasets/nyu-mll/glue.

Task_categories:text-ClassificationTask_ids:acceptability-ClassificationTask_ids:natural-Language-InferenceTask_ids:semantic-Similarity-ScoringTask_ids:sentiment-ClassificationTask_ids:text-Scoring
L
LocalAI-io/privacy-filter-nemotron-GGUF HF Unverified

Token ClassificationGGUFPrivacy-Filter.cppLlama-CppLocalaiPii HIGH
H
hustvl/yolos-small HF PQC Verified

Object-DetectionTransformersPyTorchSafetensorsYolosVision MEDIUM
C
cross-encoder/nli-deberta-v3-base HF Unverified

Zero-Shot ClassificationSentence-TransformersPyTorchONNXSafetensorsDeberta-V2 HIGH
T
timm/resnet18.fb_swsl_ig1b_ft_in1k HF Unverified

Image-ClassificationTimmPyTorchSafetensorsTransformers MEDIUM
SwayStar123/preprocessed_commoncatalog-cc-by HF Unverified

I also seperately provide just the prompts in prompts.json keys are the image_id, and the values are the captions generated Captions generated by moondream: vikhyatk/moondream2 Latents generated by SDXL VAE: madebyollin/sdxl-vae-fp16-fix Embeddings generated by SigLIP: hf-hub:timm/ViT-SO400M-14-SigLIP-384 Original dataset: common-canvas/commoncatalog-cc-by Latents f32 and embeddings are f16 bytes Compute cost: 16x3090 for 3 day. Approximately.

Language:enSize_categories:10M<n<100MFormat:parquetModality:textLibrary:datasetsLibrary:dask
Showing 20 of 921 items (page 16 of 47)