Papers › SIDOD: A Synthetic Image Dataset for 3D Object Pose Recognition with Distractors

SIDOD: A Synthetic Image Dataset for 3D Object Pose Recognition with Distractors

12 Aug 2020arXiv:2008.05955archive 2025-07-28

Mona Jalal, Josef Spjut, Ben Boudaoud, Margrit Betke

We present a new, publicly-available image dataset generated by the NVIDIA Deep Learning Data Synthesizer intended for use in object detection, pose estimation, and tracking applications. This dataset contains 144k stereo image pairs that synthetically combine 18 camera viewpoints of three photorealistic virtual environments with up to 10 objects (chosen randomly from the 21 object models of the YCB dataset [1]) and flying distractors. Object and camera pose, scene lighting, and quantity of objects and distractors were randomized. Each provided view includes RGB, depth, segmentation, and surface normal images, all pixel level. We describe our approach for domain randomization and provide insight into the decisions that produced the dataset.

PaperPDF

In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.

Code

No code repository is listed for this paper in the archive or in Syntology's graph.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

ObjectObject DetectionPose Estimationobject-detection

Datasets

Introduced by this paper, per the archive.

SIDOD

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Methods

AttentionLinear LayerMulti-Head AttentionSoftmaxSynthesizer

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections