Home › Datasets › task › Novel View Synthesis
Novel View Synthesis datasets
archive 2025-07-28
46 datasets carry the task tag "Novel View Synthesis" (the task itself: Novel View Synthesis), ordered by the archive's paper count. Page 1 of 1: 46 shown of 46. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Novel View Synthesis datasets 1–46 of 46
NeRF (Neural Radiance Fields)
Neural Radiance Fields (NeRF) is a method for synthesizing novel views of complex scenes by optimizing an underlying continuous volumetric scene function using a sparse set of input views.
3,892 papers · 1 benchmark
KITTI (Karlsruhe Institute of Technology and Toyota Technological Institute) is one of the most popular datasets for use in mobile robotics and autonomous driving.
3,661 papers · 137 benchmarks
ShapeNet is a large scale repository for 3D CAD models developed by researchers from Stanford University, Princeton University and the Toyota Technological Institute at Chicago, USA.
1,947 papers · 13 benchmarks
SYNTHIA (SYNTHetic Collection of Imagery and Annotations)
The SYNTHIA dataset is a synthetic dataset that consists of 9400 multi-viewpoint photo-realistic frames rendered from a virtual city and comes with pixel-level semantic annotations for 13 classes.
538 papers · 10 benchmarks
Mip-NeRF 360 (Unbounded Anti-Aliased Neural Radiance Fields)
Mip-NeRF 360 is an extension to the Mip-NeRF that uses a non-linear parameterization, online distillation, and a novel distortion-based regularize to overcome the challenge of unbounded scenes.
347 papers · 1 benchmark
LLFF (Local Light Field Fusion)
Local Light Field Fusion (LLFF) is a practical and robust deep learning solution for capturing and rendering novel views of complex real-world scenes for virtual exploration.
340 papers · 4 benchmarks
KITTI-360 is a large-scale dataset that contains rich sensory information and full annotations.
246 papers · 7 benchmarks
BlendedMVS is a novel large-scale dataset, to provide sufficient training ground truth for learning-based MVS.
105 papers · 0 benchmarks
OmniObject3D is a large vocabulary 3D object dataset with massive high-quality real-scanned 3D objects.
70 papers · 0 benchmarks
We present a benchmark for image-based 3D reconstruction.
55 papers · 2 benchmarks
ACID (Aerial Coastline Imagery Dataset)
ACID consists of thousands of aerial drone videos of different coastline and nature scenes on YouTube.
41 papers · 1 benchmark
This dataset focus on two blur types: camera motion blur and defocus blur.
34 papers · 0 benchmarks
We introduce our new dataset, Spaces, to provide a more challenging shared dataset for future view synthesis research.
32 papers · 0 benchmarks
ScanNet++ (ScanNet++: A High-Fidelity Dataset of 3D Indoor Scenes)
ScanNet++ is a large scale dataset with 450+ 3D indoor scenes containing sub-millimeter resolution laser scans, registered 33-megapixel DSLR images, and commodity RGB-D streams from iPhone.
25 papers · 5 benchmarks
The PhotoShape dataset consists of photorealistic, relightable, 3D shapes produced by the work proposed in the work of Park et al.
22 papers · 1 benchmark
The shiny folder contains 8 scenes with challenging view-dependent effects used in our paper.
17 papers · 0 benchmarks
We introduce Stanford-ORB, a new real-world 3D Object inverse Rendering Benchmark.
16 papers · 3 benchmarks
RTMV is a large-scale synthetic dataset for novel view synthesis consisting of ∼300k images rendered from nearly 2000 complex scenes using high-quality ray tracing at high resolution (1600 × 1600 pixels).
15 papers · 1 benchmark
RealEstate10K is a large dataset of camera poses corresponding to 10 million frames derived from about 80,000 video clips, gathered from about 10,000 YouTube videos.
14 papers · 1 benchmark
This is a gun detection dataset with 51K annotated gun images for gun detection and other 51K cropped gun chip images for gun classification collected from a few different sources.
9 papers · 6 benchmarks
X3D is a dataset containing 15 scenes and covering 4 applications for X-ray 3D reconstruction.
8 papers · 2 benchmarks
OMMO is a new benchmark for several outdoor NeRF-based tasks, such as novel view synthesis, surface reconstruction, and multi-modal NeRF.
7 papers · 0 benchmarks
HDR-GS (HDR-GS: Efficient High Dynamic Range Novel View Synthesis at 1000x Speed via Gaussian Splatting)
This is dataset for high dynamic range novel view synthesis.
5 papers · 1 benchmark
NERDS 360 (NeRF for Reconstruction, Decomposition and Scene Synthesis of 360° outdoor scenes)
We present a large-scale dataset for 3D urban scene understanding.
5 papers · 0 benchmarks
RefRef (RefRef: A Synthetic Dataset and Benchmark for Reconstructing Refractive and Reflective Objects)
RefRef is a synthetic dataset and benchmark designed for the task of reconstructing scenes with complex refractive and reflective objects.
5 papers · 1 benchmark
SWORD ('Scenes with occluded regions' dataset)
The new dataset contains around 1,500 train videos and 290 test videos, with 50 frames per video on average.
4 papers · 1 benchmark
BLEFF (Blender Forward Facing Dataset)
Synthetic (Blender) Dataset for forward facing scenes Toe vaualte NVS quality and camera parameter accuracy.
3 papers · 1 benchmark
UASOL (A large-scale high-resolution outdoor stereo dataset)
The UASOL an RGB-D stereo dataset, that contains 160902 frames, filmed at 33 different scenes, each with between 2 k and 10 k frames.
3 papers · 1 benchmark
VBR (VBR: A Vision Benchmark in Rome)
This dataset presents a vision and perception research dataset collected in Rome, featuring RGB data, 3D point clouds, IMU, and GPS data.
3 papers · 0 benchmarks
This is the dataset for the CGF 2021 paper "DONeRF: Towards Real-Time Rendering of Compact Neural Radiance Fields using Depth Oracle Networks".
2 papers · 1 benchmark
The Deep Blending Dataset comprises 19 diverse scenes, offering comprehensive resources for free-viewpoint image-based rendering (IBR).
2 papers · 1 benchmark
A high-quality captured dataset for object relighting.
2 papers · 0 benchmarks
A high-quality synthetic dataset for object relighting.
2 papers · 0 benchmarks
SILVR (A Synthetic Immersive Large-Volume Plenoptic Dataset)
We present SILVR, a dataset of light field images for six-degrees-of-freedom navigation in large fully-immersive volumes.
2 papers · 0 benchmarks
SPARF is a large-scale ShapeNet-based synthetic dataset for novel view synthesis consisting of ~17 million images rendered from nearly 40,000 shapes at high resolution (400×400 pixels).
2 papers · 0 benchmarks
iFF (Intrinsic Forward Facing)
Real-world dataset on forward facing scenes with different camera intrinisc parameters.
2 papers · 1 benchmark
A real-world low-light camera motion blur dataset for evaluating deblurring radiance fields methods.
1 paper · 0 benchmarks
The first large-scale dataset for training and evaluating novel-view synthesis from blurred images.
1 paper · 0 benchmarks
LLNeRF Dataset is a real-world dataset as a benchmark for model learning and evaluation.
1 paper · 0 benchmarks
Replay is a collection of multi-view, multi-modal videos of humans interacting socially.
1 paper · 0 benchmarks
This synthetic event dataset is used in Robust e-NeRF to study the collective effect of camera speed profile, contrast threshold variation and refractory period on the quality of NeRF reconstruction from a moving event camera.
1 paper · 0 benchmarks
Synthetic dataset comprising three different environments for multi-camera dynamic novel view synthesis for soccer.
1 paper · 0 benchmarks
Dataset of paired thermal and RGB images comprising ten diverse scenes—six indoor and four outdoor scenes— for 3D scene reconstruction and novel view synthesis (e.g.
1 paper · 0 benchmarks
HuSc3D (Human Sculpture dataset for 3D object reconstruction)
HuSc3D is a novel dataset specifically designed for rigorous benchmarking of 3D reconstruction models under realistic acquisition challenges.
0 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.