Model Hub

Browse PQC-verified AI models, datasets, and tools

ad1t7a/10Kh-RealOmin-OpenData HF Unverified

Boasting over 10,000 hours of cumulative data and 1 million+ clips, it ranks as the largest open-source embodied intelligence dataset in the industry. Compared with other datasets, it has the following advantages: Ample Data Volume & Strong Generalization Each skill is supported by sufficient data, collected from over 3,000 households and nearly 10,000 distinct fine-grained targets. It avoids simple repetitions and ensures robust generalization. Authentic Scenarios & Focused… See the full description on the dataset page: https://huggingface.co/datasets/ad1t7a/10Kh-RealOmin-OpenData.

Task_categories:roboticsTask_categories:reinforcement-LearningLanguage:enLanguage:zhSize_categories:n>1TModality:video
allenai/winogrande HF Unverified

Dataset Card for "winogrande" Dataset Summary WinoGrande is a new collection of 44k problems, inspired by Winograd Schema Challenge (Levesque, Davis, and Morgenstern 2011), but adjusted to improve the scale and robustness against the dataset-specific bias. Formulated as a fill-in-a-blank task with binary options, the goal is to choose the right option for a given sentence which requires commonsense reasoning. Supported Tasks and Leaderboards More Information… See the full description on the dataset page: https://huggingface.co/datasets/allenai/winogrande.

Language:enSize_categories:10K<n<100KFormat:parquetModality:textLibrary:datasetsLibrary:pandas
D
diffusers/stable-diffusion-xl-1.0-inpainting-0.1 HF PQC Verified

Text-to-ImageDiffusersSafetensorsStable-Diffusion-XlStable-Diffusion-Xl-DiffusersInpainting CRITICAL
F
facebook/nllb-200-distilled-1.3B HF Unverified

TranslationTransformersPyTorchM2m_100Text2text-GenerationNllb HIGH
W
webAI-Official/TwIL-LM3 HF Unverified

Text GenerationTransformersSafetensorsGGUFSmollm3Formal-Logic HIGH
O
openai/privacy-filter HF Unverified

Token ClassificationTransformersONNXSafetensorsOpenai_privacy_filterTransformers.js HIGH
M
MCG-NJU/videomae-base HF Unverified

Video-ClassificationTransformersPyTorchSafetensorsVideomaePretraining MEDIUM
Q
Qwen/Qwen3-TTS-12Hz-1.7B-VoiceDesign HF PQC Verified

Text-To-SpeechQwen-TtsSafetensorsQwen3_ttsAudioTts HIGH
Z
ZhengPeng7/BiRefNet_lite HF Unverified

Image-SegmentationBirefnetSafetensorsBackground-RemovalMask-GenerationDichotomous Image Segmentation MEDIUM
HuggingFaceFW/FineWeb HF PQC Verified

15T token dataset of cleaned English web data. Deduplicated and filtered from CommonCrawl, outperforms C4 and RefinedWeb for LLM pretraining.

DatasetPretrainingEnglish15T tokens CRITICAL
A
aufklarer/WeSpeaker-ResNet34-LM-MLX HF Unverified

Audio-ClassificationMlxSafetensorsWespeaker-Resnet34-LmSpeaker-EmbeddingSpeaker-Verification MEDIUM
F
facebook/mms-lid-126 HF Unverified

Audio-ClassificationTransformersPyTorchSafetensorsWav2vec2Mms HIGH
nvidia/SAGE-10k HF Unverified

SAGE-10k SAGE-10k is a large-scale interactive indoor scene dataset featuring realistic layouts, generated by the agentic-driven pipeline introduced in "SAGE: Scalable Agentic 3D Scene Generation for Embodied AI". The dataset contains 10,000 diverse scenes spanning 50 room types and styles, along with 565K uniquely generated 3D objects. 🔑 Key Features SAGE-10k integrates a wide variety of scenes, and particularly, preserves small items for… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/SAGE-10k.

Task_categories:text-To-3dLanguage:enSize_categories:10K<n<100KScene-GenerationInteractive-ScenesEmbodied-AI
H
Helsinki-NLP/opus-mt-es-en HF Unverified

TranslationTransformersPyTorchTfMarianText2text-Generation MEDIUM
F
facebook/detr-resnet-50 HF PQC Verified

Object-DetectionTransformersPyTorchSafetensorsDetrVision MEDIUM
S
SulphurAI/Sulphur-2-base HF Unverified

Text-To-VideoDiffusersGGUFConversational CRITICAL
google/IFEval HF Unverified

Dataset Card for IFEval Dataset Summary This dataset contains the prompts used in the Instruction-Following Eval (IFEval) benchmark for large language models. It contains around 500 "verifiable instructions" such as "write in more than 400 words" and "mention the keyword of AI at least 3 times" which can be verified by heuristics. To load the dataset, run: from datasets import load_dataset ifeval = load_dataset("google/IFEval") Supported Tasks and… See the full description on the dataset page: https://huggingface.co/datasets/google/IFEval.

Task_categories:text-GenerationLanguage:enSize_categories:n<1KFormat:jsonModality:textLibrary:datasets
H
handy-computer/canary-180m-flash-gguf HF Unverified

Speech RecognitionTranscribe.cppGGUFAsrSpeech-To-TextCanary HIGH
L
LocalAI-io/privacy-filter-multilingual-GGUF HF Unverified

Token ClassificationGGUFPrivacy-Filter.cppLlama-CppLocalaiPii HIGH
M
microsoft/VibeVoice-Realtime-0.5B HF PQC Verified

Text-To-SpeechTransformersSafetensorsVibevoice_streamingRealtime TTSStreaming text input HIGH
Showing 20 of 924 items (page 23 of 47)