Model Hub

Browse PQC-verified AI models, datasets, and tools

T
timm/convnext_tiny.in12k_ft_in1k HF Unverified

Image-ClassificationTimmPyTorchSafetensorsTransformers MEDIUM
D
depth-anything/DA3-GIANT HF Unverified

Depth-EstimationDepth-Anything-3SafetensorsComputer-VisionMonocular-DepthMulti-View-Geometry HIGH
O
openai/privacy-filter HF Unverified

Token ClassificationTransformersONNXSafetensorsOpenai_privacy_filterTransformers.js HIGH
F
facebook/mms-tts-hat HF PQC Verified

Text-To-SpeechTransformersPyTorchSafetensorsVitsText-To-Audio MEDIUM
C
cross-encoder/nli-deberta-v3-base HF Unverified

Zero-Shot ClassificationSentence-TransformersPyTorchONNXSafetensorsDeberta-V2 HIGH
H
Helsinki-NLP/opus-mt-it-en HF Unverified

TranslationTransformersPyTorchTfMarianText2text-Generation MEDIUM
nyu-mll/glue HF PQC Verified

Dataset Card for GLUE Dataset Summary GLUE, the General Language Understanding Evaluation benchmark (https://gluebenchmark.com/) is a collection of resources for training, evaluating, and analyzing natural language understanding systems. Supported Tasks and Leaderboards The leaderboard for the GLUE benchmark can be found at this address. It comprises the following tasks: ax A manually-curated evaluation dataset for fine-grained analysis of system… See the full description on the dataset page: https://huggingface.co/datasets/nyu-mll/glue.

Task_categories:text-ClassificationTask_ids:acceptability-ClassificationTask_ids:natural-Language-InferenceTask_ids:semantic-Similarity-ScoringTask_ids:sentiment-ClassificationTask_ids:text-Scoring
D
Davlan/xlm-roberta-large-ner-hrl HF Unverified

Token ClassificationTransformersPyTorchTfSafetensorsXlm-Roberta HIGH
O
obi/deid_roberta_i2b2 HF Unverified

Token ClassificationTransformersPyTorchSafetensorsRobertaDeidentification HIGH
mteb/results HF Unverified

Size_categories:1M<n<10MFormat:parquetFormat:optimized-ParquetModality:textLibrary:datasetsLibrary:dask
C
CompVis/stable-diffusion-v1-4 HF PQC Verified

Text-to-ImageDiffusersSafetensorsStable-DiffusionStable-Diffusion-DiffusersDiffusers:StableDiffusionPipeline CRITICAL
L
lightx2v/Qwen-Image-Lightning HF PQC Verified

Text-to-ImageDiffusersQwen-ImageDistillationLoRALora CRITICAL
Q
Qwen/Qwen3-TTS-12Hz-1.7B-VoiceDesign HF PQC Verified

Text-To-SpeechQwen-TtsSafetensorsQwen3_ttsAudioTts HIGH
M
microsoft/llmlingua-2-bert-base-multilingual-cased-meetingbank HF Unverified

Token ClassificationTransformersSafetensorsBERT MEDIUM
P
PekingU/rtdetr_r101vd_coco_o365 HF Unverified

Object-DetectionTransformersSafetensorsRt_detrVisionEnglish MEDIUM
F
facebook/seamless-m4t-v2-large HF Unverified

Speech RecognitionTransformersSafetensorsSeamless_m4t_v2Feature ExtractionAudio-To-Audio HIGH
HuggingFaceFW/FineWeb HF PQC Verified

15T token dataset of cleaned English web data. Deduplicated and filtered from CommonCrawl, outperforms C4 and RefinedWeb for LLM pretraining.

DatasetPretrainingEnglish15T tokens CRITICAL
P
playgroundai/playground-v2.5-1024px-aesthetic HF PQC Verified

Text-to-ImageDiffusersSafetensorsPlaygroundDiffusers:StableDiffusionXLPipeline CRITICAL
anon8231489123/ShareGPT_Vicuna_unfiltered HF Unverified

Further cleaning done. Please look through the dataset and ensure that I didn't miss anything. Update: Confirmed working method for training the model: https://huggingface.co/AlekseyKorshuk/vicuna-7b/discussions/4#64346c08ef6d5abefe42c12c Two choices: Removes instances of "I'm sorry, but": https://huggingface.co/datasets/anon8231489123/ShareGPT_Vicuna_unfiltered/blob/main/ShareGPT_V3_unfiltered_cleaned_split_no_imsorry.json Has instances of "I'm sorry, but":… See the full description on the dataset page: https://huggingface.co/datasets/anon8231489123/ShareGPT_Vicuna_unfiltered.

Language:en
HuggingFaceFW/fineweb-edu HF PQC Verified

📚 FineWeb-Edu 1.3 trillion tokens of the finest educational data the 🌐 web has to offer Paper: https://arxiv.org/abs/2406.17557 What is it? 📚 FineWeb-Edu dataset consists of 1.3T tokens and 5.4T tokens (FineWeb-Edu-score-2) of educational web pages filtered from 🍷 FineWeb dataset. This is the 1.3 trillion version. To enhance FineWeb's quality, we developed an educational quality classifier using annotations generated by LLama3-70B-Instruct. We then… See the full description on the dataset page: https://huggingface.co/datasets/HuggingFaceFW/fineweb-edu.

Task_categories:text-GenerationLanguage:enSize_categories:1B<n<10BFormat:parquetModality:tabularModality:text
Showing 20 of 802 items (page 19 of 41)