Model Hub

Browse PQC-verified AI models, datasets, and tools

D
distilbert/distilbert-base-cased-distilled-squad HF Unverified

Question AnsweringTransformersPyTorchTfRustSafetensors HIGH
CohereLabs/xP3x HF Unverified

Dataset Card for xP3x Dataset Summary xP3x (Crosslingual Public Pool of Prompts eXtended) is a collection of prompts & datasets across 277 languages & 16 NLP tasks. It contains all of xP3 + much more! It is used for training future contenders of mT0 & BLOOMZ at project Aya @Cohere Labs 🧡 Creation: The dataset can be recreated using instructions available here together with the file in this repository named xp3x_create.py. We provide this version to save processing… See the full description on the dataset page: https://huggingface.co/datasets/CohereLabs/xP3x.

Task_categories:otherAnnotations_creators:expert-GeneratedAnnotations_creators:crowdsourcedMultilinguality:multilingualLanguage:afLanguage:ar
InternRobotics/InternData-A1 HF Unverified

InternData-A1 InternData-A1 is a hybrid synthetic-real manipulation dataset containing over 630k trajectories and 7,433 hours across 4 embodiments, 18 skills, 70 tasks, and 227 scenes, covering rigid, articulated, deformable, and fluid-object manipulation. Your browser does not support the video tag. Your browser does not support the video tag.… See the full description on the dataset page: https://huggingface.co/datasets/InternRobotics/InternData-A1.

Task_categories:otherTask_categories:roboticsLanguage:enSize_categories:n>1TModality:3dModality:image
M
microsoft/table-transformer-structure-recognition-v1.1-all HF PQC Verified

Object-DetectionTransformersSafetensorsTable-Transformer MEDIUM
Kazimir-ai/text-to-image-prompts HF Unverified

The dataset of the most popular text-to-image prompts. Dataset Details Dataset Description Curated by: kazimir.ai Funded by [optional]: [More Information Needed] Shared by [optional]: https://kazimir.ai License: apache-2.0 Dataset Sources [optional] Repository: [More Information Needed] Paper [optional]: [More Information Needed] Demo [optional]: [More Information Needed] Uses Free to use. Dataset Structure CSV file… See the full description on the dataset page: https://huggingface.co/datasets/Kazimir-ai/text-to-image-prompts.

Language:enSize_categories:10K<n<100KFormat:csvModality:textLibrary:datasetsLibrary:pandas
mvp-lab/LLaVA-OneVision-2-Data HF Unverified

LLaVA-OneVision-2-Data Training data for the LLaVA-OneVision-2 multimodal model family, covering large-scale video and spatial reasoning corpora used in mid-training. Dataset Composition Subset Format Description mid_training_video/60s_rest/ WebDataset (.tar) 10,809 shards of ~60s video clips mid_training_video/caption_v0/split_30s.jsonl JSONL Captions for 30-second video clips mid_training_video/caption_v0/split_60s.jsonl JSONL Captions for… See the full description on the dataset page: https://huggingface.co/datasets/mvp-lab/LLaVA-OneVision-2-Data.

Task_categories:video-Text-To-TextTask_categories:visual-Question-AnsweringTask_categories:image-Text-To-TextLanguage:enSize_categories:n<1KFormat:parquet
U
unsloth/LTX-2.3-GGUF HF Unverified

Image-To-VideoGgmlGGUFUnslothText-To-VideoVideo-To-Video CRITICAL
Q
Qwen/Qwen-Image HF PQC Verified

Text-to-ImageDiffusersSafetensorsDiffusers:QwenImagePipelineEnglishChinese CRITICAL
J
jameslahm/yolov10s HF Unverified

Object-DetectionYolov10SafetensorsComputer-VisionPytorch_model_hub_mixin MEDIUM
T
tencent/HunyuanImage-3.0 HF Unverified

Text-to-ImageTransformersSafetensorsHunyuan_image_3_moeText GenerationCustom_code CRITICAL
N
nphSi/Z-Image-Lora HF Unverified

Text-to-ImageDiffusersLoraSafetensorsZ-ImageBase_model:Tongyi-MAI/Z-Image CRITICAL
B
black-forest-labs/FLUX.2-dev HF PQC Verified

Image-To-ImageDiffusersSafetensorsImage GenerationImage-EditingFlux CRITICAL
WINGNUS/ACL-OCL HF Unverified

Dataset Card for ACL Anthology Corpus This repository provides full-text and metadata to the ACL anthology collection (80k articles/posters as of September 2022) also including .pdf files and grobid extractions of the pdfs. How is this different from what ACL anthology provides and what already exists? We provide pdfs, full-text, references and other details extracted by grobid from the PDFs while ACL Anthology only provides abstracts. There exists a similar corpus… See the full description on the dataset page: https://huggingface.co/datasets/WINGNUS/ACL-OCL.

Task_categories:token-ClassificationLanguage_creators:foundMultilinguality:monolingualSource_datasets:originalLanguage:enSize_categories:10K<n<100K
B
ByteDance/SDXL-Lightning HF Unverified

Text-to-ImageDiffusersStable-Diffusion CRITICAL
P
PramaLLC/BEN2 HF Unverified

Image-SegmentationBen2ONNXSafetensorsBEN2Background-Remove HIGH
GokuScraper/seedance-2-prompts-datasets HF Unverified

🎞️ Seedance-2-prompts-datasets 🎞️ The ultimate Seedance-2 video prompt dataset (50GB+). 8100+ video generation prompts with full metadata and preview frames. Truly open source: No login, no ads, no redirection. Just pure data for AI video creators. This project is a massive collection of prompts used for Bytedance's Seedance 2.0 and the resulting generated videos. The entire dataset exceeds 50GB and contains 8100+ videos, all structured into a comprehensive dataset. Due… See the full description on the dataset page: https://huggingface.co/datasets/GokuScraper/seedance-2-prompts-datasets.

Task_categories:text-To-VideoLanguage:enLanguage:zhSize_categories:1K<n<10KModality:imageModality:video
J
John6666/nova-furry-xl-il-v120-sdxl HF Unverified

Text-to-ImageDiffusersSafetensorsStable-DiffusionStable-Diffusion-XlNot-For-All-Audiences HIGH
M
myshell-ai/MeloTTS-Japanese HF Unverified

Text-To-SpeechTransformersKorean MEDIUM
SwayStar123/preprocessed_commoncatalog-cc-by HF Unverified

I also seperately provide just the prompts in prompts.json keys are the image_id, and the values are the captions generated Captions generated by moondream: vikhyatk/moondream2 Latents generated by SDXL VAE: madebyollin/sdxl-vae-fp16-fix Embeddings generated by SigLIP: hf-hub:timm/ViT-SO400M-14-SigLIP-384 Original dataset: common-canvas/commoncatalog-cc-by Latents f32 and embeddings are f16 bytes Compute cost: 16x3090 for 3 day. Approximately.

Language:enSize_categories:10M<n<100MFormat:parquetModality:textLibrary:datasetsLibrary:dask
jsulz/FIFA23 HF Unverified

About this dataset Context The datasets provided include the players data for the Career Mode from FIFA 15 to FIFA 23. The data allows multiple comparisons for the same players across the last 9 versions of the video game. Some ideas of possible analysis: Historical comparison between Messi and Ronaldo (what skill attributes changed the most during time - compared to real-life stats); Ideal budget to create a competitive team (at the level of top n teams in Europe) and… See the full description on the dataset page: https://huggingface.co/datasets/jsulz/FIFA23.

Task_categories:tabular-ClassificationTask_categories:tabular-RegressionLanguage:enSize_categories:10M<n<100MModality:tabularTabular
Showing 20 of 805 items (page 24 of 41)