Hugging Face Datasets2026 · Table · Parquet
YijiaFan/UMM-Reflection-SFT-DataUMM-Reflection SFT Data The reflection-SFT data of UMM-Reflection (Learning Native Reflection in Unified Models). It trains UMM-Reflection-BAGEL-SFT. Research use only, non-commercial. The rows are derived from datasets with different licenses, some of them non-commercial. Each row records its source dataset and license in source_dataset and source_license, and each row follows the terms of its so
Hugging Face Datasets2026 · dataset
Erotic Image PromptsDataset Card for Erotic Image Prompts Dataset Description Dataset Summary Large language models (LLMs) are surprisingly bad at creatively inventing new things, even though they are masters of hallucination. Asking an LLM — even an abliterated one — to produce a list of random erotic prompts therefore yields a rather boring, narrow-minded result. This is easy to overcome: give the model some inspir
Hugging Face Datasets2026 · Image
ML-Intern-lab/Qwen-Image-2.1-rewriter-distillQwen-Image-2.1-rewriter-distill A distillation corpus for the Qwen-Image-2.1 text-to-image prompt rewriter. Qwen/Qwen-Image-2.1-PE-T2I is a 9B thinking model that turns a short image request in any language into one long English paragraph plus an aspect ratio. It needs a ~1,700-word system prompt, thinking mode, and up to 16,256 new tokens to emit a single JSON object. This dataset records what it
Hugging Face Datasets2026 · Text
VCL Image PromptsVCL Image Prompts Copy-paste visual prompt recipes for ChatGPT, Gemini, Grok, and Microsoft AI (MAI Playground) chats — by Vibe Coder's Life. Slash tokens like /xray and /blueprint are memorable shorthand, not official commands. If a chat ignores the /, paste the expanded prompt in the recipe. No model ships a real /xray API. This dataset mirrors the open-source GitHub cookbook. People sell lists
Hugging Face Datasets2026 · Image
CADBench Extended Multimodal DatasetDataset Card Dataset Description CADBench Extended Multimodal Dataset is an independently produced public extension for multimodal CAD reconstruction research. It contains 100 CAD samples with clean and perturbed meshes, STEP/STL/OBJ/GLB representations, single-view and four-view renders, PBR images, bilingual descriptions, prompt variants, QA, geometry metadata, grading signals, and manually revi
Hugging Face Datasets2026 · Image
3D Animation Style3D Animation Style A curated collection of 54 high-resolution synthetic images by SOLRICKS, designed around a warm, cinematic 3D animation aesthetic. The dataset combines richly lit environments, fantasy interiors, stylized animal characters and expressive human characters. Dataset details Images: 54 PNG files Resolution: 50 images at 1254×1254 and 4 images at 1536×1024 Captions: English, manually
Hugging Face Datasets2026 · Table · Parquet
Curated Danbooru 2026 Dataset (AVIF / Parquet)Curated Danbooru Streaming Dataset A large-scale, high-performance curated dataset of ~330,000 (330K) high-quality anime illustrations designed for training Diffusion Transformers (DiT), Latent Diffusion Models (LDM), and text-to-image generative models focused on the anime domain. This dataset is focused on specific curated characters and high-ranking artists using knowledge base lists (character
Hugging Face Datasets2026 · dataset
danbooru2025Danbooru2026: A Large-Scale Crowdsourced and Tagged Anime Illustration Dataset [WIP] Dataset Description Danbooru2026 is a large-scale anime illustration dataset containing over 10 million community-annotated images. It is intended for research and development in anime-style image generation, image classification, multimodal learning, and related tasks. Danbooru is a long-running image board known
Hugging Face Datasets2026 · dataset
Pruna Skills doc examplesPruna Skills — doc examples Generated media and sidecars for PrunaAI/pruna-skills documentation. Each file under examples/ matches a skill demo in docs/EXAMPLES.md. PNG/MP3/MP4 outputs have a sibling .meta.json with the exact prompt, model, and inputs. Layout examples/ p-image-advanced.png p-image-advanced.meta.json quickstart-knight-still.png quickstart-knight-clip.mp4 chain-monarch-clip.mp4 … Re
Hugging Face Datasets2026 · dataset
Chinese Adult Image Caption Dataset本数据集是一个面向成人内容研究与模型训练的高质量、细粒度图像标注数据集。数据中的人物均为中国成年人,内容涵盖多种成人图像类型、场景、人物特征、姿势、服饰、环境及视觉细节,具有较强的内容多样性。 所有文本描述与相关标签均由人工完成,并经过人工检查与复核,以尽可能保证标注内容的准确性、一致性和完整性。标注内容较为详细,可用于需要细粒度成人视觉语义理解的相关研究和模型训练任务。 本数据集可用于: 文生图模型的微调与 LoRA 训练; 图像描述、图像理解及视觉语言模型训练; 多模态大模型微调,例如 Qwen-VL 等模型; 成人内容分类、识别、审核与安全研究; NSFW 图像描述生成及细粒度标签学习。 当前 Hugging Face 仓库仅公开少量示例数据。完整数据集目前约包含 500 张经过人工精细标注的图像,并仍在持续扩充和标注中。 如需获取完整数据集、了解价格、定制标注格式,或需要人工标注其
Hugging Face Datasets2026 · Image
Multi-Reference Instruction-Based Image Editing DatasetMulti-Reference Instruction-Based Image Editing Dataset Overview This dataset contains 20,000 high-resolution image pairs and multi-modal instructions designed for training advanced image-to-image editing models. It combines two complementary example types: 10,000 reference-grounded edits, where structural or stylistic changes are driven by up to three provided visual reference images, and 10,000
Hugging Face Datasets2026 · Image
SVG Generation Benchmark (Static)Rapidata Static SVG Generation Benchmark Built by Rapidata. This dataset contains 1,355,161 human responses, collected with the Rapidata Python SDK, comparing how well 30 frontier LLMs generate static SVGs from text prompts. Each row is a head-to-head comparison between two models' renders of the same prompt, scored by human annotators on one of three questions (Preference, Coherence, Alignment).
Hugging Face Datasets2026 · Image
SVG Generation Benchmark (Static)Rapidata Static SVG Generation Benchmark Built by Rapidata. This dataset contains 1,918,367 human responses, collected with the Rapidata Python SDK, comparing how well 42 frontier LLMs generate static SVGs from text prompts. Each row is a head-to-head comparison between two models' renders of the same prompt, scored by human annotators on one of three questions (Preference, Coherence, Alignment).
Hugging Face Datasets2026 · Image
SkingSking: Minecraft Skin 3D Rendering & Dataset Preprocessing Pipeline This repository is utilized for training and fine-tuning the Sking Model (Hugging Face Model). Along with the dataset assets, it contains the complete automated pipeline for dual-layer 3D voxel rendering, multi-view layout baking, automatic skin format conversion (Alex to Steve), voxel edge texture consistency resolution, and a RE
Hugging Face Datasets2026 · Table · Parquet
zlab-princeton/i1-captionsi1: A Simple and Fully Open Recipe for Strong Text-to-Image Models Boya Zeng, Tianze Luo, Shu Pu, Jucheng Shen, Taiming Lu, Gabriel Sarch, Zhuang Liu Princeton University [arXiv][code][model][project page] 1. Overview This dataset contains all captions used in our controlled experiments and the final training of the i1 model. Detailed instructions for downloading the corresponding images and match
Hugging Face Datasets2026 · Image
MONETDataset Card for MONET MONET (Massive, Open, Non-redundant and Enriched Text-to-image dataset) is a large-scale, curated image-text dataset designed for training text-to-image (T2I) systems. It contains 103.8 million high-quality image-text pairs distilled from 2.9 billion raw pairs across nine heterogeneous open sources (6 real and 3 synthetic) through successive stages of safety filtering, domai
Hugging Face Datasets2026 · Image
apple/DFNDR-12M-bf16Dataset Card for DFNDR-12M-BFloat16 This dataset contains synthetic captions, embeddings, and metadata for DFNDR-12M. The metadata has been generated using pretrained image-text models on DFN-12M, a uniformly sampled subset of 12.8M samples from DFN-2B. For details on how to use the metadata, please visit our ml-mobileclip repository. For code to generate multi-modal reinforced datasets at large s
Hugging Face Datasets2026 · Image
apple/DFNDR-12MDataset Card for DFNDR-12M This dataset contains synthetic captions, embeddings, and metadata for DFNDR-12M. The metadata has been generated using pretrained image-text models on DFN-12M, a uniformly sampled subset of 12.8M samples from DFN-2B. For details on how to use the metadata, please visit our ml-mobileclip repository. For code to generate multi-modal reinforced datasets at large scale see
Hugging Face Datasets2026 · dataset
DFNDR-2BDataset Card for DFNDR-2B This dataset contains synthetic captions, embeddings, and metadata for DFNDR-2B. The metadata has been generated using pretrained image-text models on DFN-2B, a 2B filtered subset of DataComp-12B. For details on how to use the metadata, please visit our ml-mobileclip repository. For code to generate multi-modal reinforced datasets at large scale see ml-mobileclip-dr repos
Hugging Face Datasets2026 · Image · gated
OCRGenBenchOCRGenBench: A Comprehensive Benchmark for Evaluating OCR Generative Capabilities 🔐 Dataset Access This dataset is gated. To download OCRGenBench, please submit an access request: 👉 Apply for Access — click the Access button on the dataset page Applications are automatically approved. You will receive an email confirmation once access is granted. 📖 Overview OCRGenBench is the most comprehensive be
Hugging Face Datasets2026 · dataset · gated
iis-esslingen/SearchADSearchAD Dataset Main Project Page You can find more information about the SearchAD Dataset on its official project page: https://iis-esslingen.github.io/searchad/ SearchAD Benchmark The official SearchAD rare image retrieval competition including the leaderboard can be found here Dataset Overview The SearchAD dataset is a large-scale autonomous driving datasets, specifically targeting rare and sa
Hugging Face Datasets2026 · Image
MMEB-V3MMEB-V3: Measuring the Performance Gaps of Omni-Modality Embedding Models 🌐 Website | GitHub | 🏆 Leaderboard | 📖 MMEB-V3 Paper | 📖 MMEB-V2 Paper | 📖 MMEB-V1 Paper | 🤗 Models Introduction MMEB-V3 is a comprehensive benchmark for evaluating omni-modality embedding models across text, image, video, audio, visual-document, and agent-centric retrieval scenarios. Building upon MMEB-V1 and MMEB-V2, MMEB-
Hugging Face Datasets2026 · dataset
GilgameshYX/WeatherSyntheticWeatherSynthetic Driving Scene Dataset WeatherSynthetic is a synthetic dataset featuring rich intrinsic map annotations, specifically designed for autonomous driving scenarios under diverse weather and lighting conditions. This dataset was introduced in our paper: "IntrinsicWeather: Controllable Weather Editing in Intrinsic Space". We hope it proves beneficial for future research in the field. Dat
Hugging Face Datasets2026 · Text
CSFM-ImageNet1K-CaptionCSFM-ImageNet1K-Caption Dataset Project Page | Paper | Code This repository contains dataset associated with the paper "Better Source Better Flow: Learning Condition-Dependent Source Distribution for Flow Matching". This dataset is used for training and evaluating Condition-dependent Source Flow Matching (CSFM), a framework that learns condition-dependent source distributions for flow matching. We
Hugging Face Datasets2026 · Image
ma-xu/fine-t2iFine-T2I: An Open, Large-Scale, and Diverse Dataset for High-Quality T2I Fine-Tuning [arxiv] by Xu Ma, Yitian Zhang, Qihua Dong, Yun Fu Northeastern Univeristy Please see our [Dataset Explore] to view detailed samples (loading is slow, be patient). 🆕 What's New [2026.02.20]: Fine-T2I reaches the #1 spot among Hugging Face Datasets Trending list ⭐️⭐️⭐️ [2026.02.16]: Fine-T2I tops the Hugging Face D