Hugging Face Datasets2026 · dataset
EndermanPC/ffhq512-processed-ptFFHQ-512 Processed Chunks (PyTorch Tensors) This dataset contains 7,000 images sourced from Ryan-sjtu/ffhq512-caption, preprocessed and saved as chunked PyTorch tensor (.pt) files. Preprocessing Details Images are preprocessed using a standardized pipeline: Color Format: Converted to RGB. Resize: Resized to (512, 512) using BICUBIC interpolation. Normalization: Normalized to the range [-1.0, 1.0]
Hugging Face Datasets2026 · Image
OmniTaskonomy Recipe DataOmniTaskonomy Recipe Data Paired image-to-image (I2I) and image-to-text (I2T) tasks for the R1–R6 training recipes and gradient analysis in OmniTaskonomy. Each of the six subsets has train and val splits. One row contains both objectives for the same task instance. from datasets import load_dataset data = load_dataset("Wakals/OmniTaskonomy_Recipe_Data", "jigsaw", split="train", streaming=True) sam
Hugging Face Datasets2026 · Image
OmniTaskonomyOmniTaskonomy OmniTaskonomy groups visual tasks into Recognition, Reconstruction, and Reorganization. Each task and modality has its own split, named family__task__i2i or family__task__i2t. The i2t and i2i configs group the evaluation and training splits, respectively. This release contains 9,444 I2T evaluation samples across 25 tasks and 350,000 I2I training samples across 7 tasks. I2T rows have
Hugging Face Datasets2026 · Image
dm-bench 0.1.0dm-bench 0.1.0 Torn-document reassembly benchmark. Synthetic, seeded, licence-clean pages are torn into non-overlapping fragments; a solver must place every fragment back on the page with a rigid pose. Full benchmark card, metrics and baseline: docs/BENCHMARK.md. Code: Arittra-Bag/Dataset-Maker. Contents tier val pages test-dev pages test pages easy 39 13 48 medium 37 16 50 hard 33 19 52 puzzles/<
Hugging Face Datasets2026 · Table · Parquet
YijiaFan/UMM-Reflection-SFT-DataUMM-Reflection SFT Data The reflection-SFT data of UMM-Reflection (Learning Native Reflection in Unified Models). It trains UMM-Reflection-BAGEL-SFT. Research use only, non-commercial. The rows are derived from datasets with different licenses, some of them non-commercial. Each row records its source dataset and license in source_dataset and source_license, and each row follows the terms of its so
Hugging Face Datasets2026 · Image
LME-BenchLME-Bench Long-horizon Multi-turn image Editing benchmark, introduced in MT-OPSD: On-Policy Self-Distillation for Multi-Turn Image Editing. Existing multi-turn editing benchmarks stop at five turns. LME-Bench has 100 ten-turn sessions in which every instruction is applied to the previous turn's output, to measure whether an editor keeps following instructions and keeps its images intact over long
Hugging Face Datasets2026 · Image
ESP32-CAM Autonomous Line-Following Car DatasetLighting-related limitation Most right-turn samples were recorded under lower-light conditions than those found at the official track. As a result, the model may have unintentionally associated darker images, contrast levels, shadows, or camera exposure settings with right-turn commands. Under brighter lighting, the floor may produce different reflections and the tape may have a different contrast
Hugging Face Datasets2026 · Image
Component-wise Translated CT ProjectionsComponent-wise translated CT projections Synthetic chest radiographs derived from 21,887 chest CT volumes of CT-RATE, released as the translated outputs of the pipeline described in Anatomy-Decomposed Chest Computed Tomography (CT) Projections as Scalable Supervision for Bone Suppression in Chest Radiographs Angaitkar, Kumar, Satia, Rao, Mittal, Tadepalli, Putha — arXiv:2609.24937 (2026; under rev
Hugging Face Datasets2026 · Image
3D Animation Style3D Animation Style A curated collection of 54 high-resolution synthetic images by SOLRICKS, designed around a warm, cinematic 3D animation aesthetic. The dataset combines richly lit environments, fantasy interiors, stylized animal characters and expressive human characters. Dataset details Images: 54 PNG files Resolution: 50 images at 1254×1254 and 4 images at 1536×1024 Captions: English, manually
Hugging Face Datasets2026 · dataset
inclusionAI/ConceptEdit-12MConceptEdit: Unlocking the Potential of Image Editing via Concept Scaling and Dense Supervision ConceptEdit-12M is a large-scale image editing dataset. Each sample is stored as a triplet: a source image, an edited image, a JSON metadata file describing the edit instruction, edit category, relative image paths, and VQA-style quality checks. The dataset is packaged as multiple .ta
Hugging Face Datasets2026 · dataset
danbooru2025Danbooru2026: A Large-Scale Crowdsourced and Tagged Anime Illustration Dataset [WIP] Dataset Description Danbooru2026 is a large-scale anime illustration dataset containing over 10 million community-annotated images. It is intended for research and development in anime-style image generation, image classification, multimodal learning, and related tasks. Danbooru is a long-running image board known
Hugging Face Datasets2026 · dataset
Identity Preservation Augmentation Dataset for Image EditingIdentity Preservation Augmentation Dataset for Image Editing Overview This dataset contains algorithmically generated image pairs designed to teach diffusion-based image editing models pixel-level identity preservation — the ability to keep unchanged regions of an image exactly intact while applying targeted edits. Every example consists of a reference image, a target image, and a short natural-la
Hugging Face Datasets2026 · Image
Multi-Reference Instruction-Based Image Editing DatasetMulti-Reference Instruction-Based Image Editing Dataset Overview This dataset contains 20,000 high-resolution image pairs and multi-modal instructions designed for training advanced image-to-image editing models. It combines two complementary example types: 10,000 reference-grounded edits, where structural or stylistic changes are driven by up to three provided visual reference images, and 10,000
Hugging Face Datasets2026 · Image
MV2 DatasetMV2 Dataset MV2 is a multi-view and multi-vehicle urban driving dataset designed for research in novel view synthesis, neural rendering, 3D reconstruction, cross-view scene understanding, and autonomous-driving perception. The dataset contains synchronized image sequences captured from multiple viewpoints, including ground vehicles and aerial views, along with camera parameters required for geomet
Hugging Face Datasets2026 · Image · gated
Dress-EDDress-ED Instruction-Guided Editing for Virtual Try-On and Try-Off Overview Dress-ED is a large-scale garment-editing dataset for instruction-guided virtual try-on and clothing manipulation research. Given a person image and a garment image, each example asks a model to apply a specific textual edit to the garment while keeping it worn on the person. The dataset provides three complementary instru
Hugging Face Datasets2026 · Image
MODEST: Multi-Optics DOF Stereo DatasetMODEST Multi-Optics DOF Stereo Dataset A large-scale, ultra-high resolution, computational photography dataset for depth-of-field rendering, defocus deblurring, stereo vision, and optical scene understanding Project Website • Paper Overview MODEST is the first high-resolution stereo DSLR dataset that systematically captures variations in focal length and aperture using professional camera optics.
Hugging Face Datasets2026 · dataset · gated
MV-FashionMV-Fashion: Towards Enabling Virtual Try-On and Size Estimation with Multi-View Paired Data CVPR 2026 Highlight Hunor Laczkó • Libang Jia • Loc-Phat Truong • Diego Hernández • Sergio Escalera • Jordi Gonzàlez • Meysam Madadi ⏳ Reduced availability during August Our university administration is unavailable for the month of August. You can still submit your request as normal, but please allow until
Hugging Face Datasets2026 · dataset
ChengYou305/DF3DV-1KDF3DV-1K: A Large-Scale Dataset and Benchmark for Distractor-Free Novel View Synthesis Project Page | Paper | GitHub DF3DV-1K is a large-scale real-world dataset comprising 1,048 scenes, each providing clean and cluttered image sets for benchmarking distractor-free radiance fields. In total, the dataset contains 89,924 images captured using consumer cameras to mimic casual capture, spanning 128 di
Hugging Face Datasets2026 · Image
SkingSking: Minecraft Skin 3D Rendering & Dataset Preprocessing Pipeline This repository is utilized for training and fine-tuning the Sking Model (Hugging Face Model). Along with the dataset assets, it contains the complete automated pipeline for dual-layer 3D voxel rendering, multi-view layout baking, automatic skin format conversion (Alex to Steve), voxel edge texture consistency resolution, and a RE
Hugging Face Datasets2026 · Image
CM-EVS — Coverage-Curated Panoramic RGB-D DatasetCM-EVS: A Coverage-Curated Panoramic RGB-D Dataset for Indoor Scene Understanding CM-EVS is a curated panoramic RGB-D dataset built under a single principle: maximize the geometric coverage of a 3D scene with the fewest equirectangular (ERP) frames possible. The release is structured as one redistributable Blender indoor data archive plus four license-aware adapter packages that regenerate matched
Hugging Face Datasets2026 · Image · gated
OCRGenBenchOCRGenBench: A Comprehensive Benchmark for Evaluating OCR Generative Capabilities 🔐 Dataset Access This dataset is gated. To download OCRGenBench, please submit an access request: 👉 Apply for Access — click the Access button on the dataset page Applications are automatically approved. You will receive an email confirmation once access is granted. 📖 Overview OCRGenBench is the most comprehensive be
Hugging Face Datasets2026 · Image
AnimalLiftAnimalLift Official dataset for AnimalLift · SIGGRAPH Asia 2026 AnimalLift: Reconstructing Animatable 3D Animals from a Single Image by Learning Canonical Shape, Texture, and Fur Maps Chunyi Sun¹ · Ruyi Zha¹ · Weijian Deng¹ · Junlin Han² · Dylan Campbell¹ · Stephen Gould¹ ¹ Australian National University ² University of Oxford Code · Model & Checkpoints · Dataset Files Overview · Download ·
Hugging Face Datasets2026 · dataset · gated
iis-esslingen/SearchADSearchAD Dataset Main Project Page You can find more information about the SearchAD Dataset on its official project page: https://iis-esslingen.github.io/searchad/ SearchAD Benchmark The official SearchAD rare image retrieval competition including the leaderboard can be found here Dataset Overview The SearchAD dataset is a large-scale autonomous driving datasets, specifically targeting rare and sa
Hugging Face Datasets2026 · dataset
RealX3D: A Physically-Degraded 3D Benchmark for Multi-view Visual Restoration and ReconstructionRealX3D: A Physically-Degraded 3D Benchmark for Multi-view Visual Restoration and Reconstruction RealX3D is a real-world benchmark dataset for multi-view 3D reconstruction under challenging capture conditions. It provides multi-view RGB images (both processed JPEG and Sony RAW), COLMAP sparse reconstructions, and high-precision 3D ground-truth geometry (point clouds, meshes, and rendered depth map
Hugging Face Datasets2026 · dataset
GilgameshYX/WeatherSyntheticWeatherSynthetic Driving Scene Dataset WeatherSynthetic is a synthetic dataset featuring rich intrinsic map annotations, specifically designed for autonomous driving scenarios under diverse weather and lighting conditions. This dataset was introduced in our paper: "IntrinsicWeather: Controllable Weather Editing in Intrinsic Space". We hope it proves beneficial for future research in the field. Dat