Browse State-of-the-Art › 3D Generation
3D Generation
144 papers with code · 1 benchmark · 5 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
1 leaderboard table shown for this task, 1 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| E.T. the Exceptional Trajectories (4 rows) | DIRECTOR C | E.T. the Exceptional Trajectories: Text-to-camera-trajectory... | code | — | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
5 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
1 subtask in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 144 papers with code (430 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
1 Sep 2023 5 repositories listed Syntology ran 13 of 20 samples · 7 unverified · 7 pointer-only (licence)We introduce Point-Bind, a 3D multi-modality model aligning point clouds with 2D image, language, audio, and video.
-
5 Jun 2024 4 repositories listed Syntology ran 6 of 10 samples · 4 unverified · 5 pointer-only (licence)Besides the generative capabilities of diffusion priors, motivated by the unique time-symmetry properties of rectified flow models, a variant of our method can additionally perform image inversion.
-
31 Aug 2023 4 repositories listed Syntology ran 2 of 2 samples · 0 unverifiedWe introduce MVDream, a diffusion model that is able to generate consistent multi-view images from a given text prompt.
-
23 May 2024 3 repositories listed Syntology ran 4 of 6 samples · 2 unverified · 6 pointer-only (licence)We present a novel generative 3D modeling system, coined CraftsMan, which can generate high-fidelity 3D geometries with highly varied shapes, regular mesh topologies, and detailed surfaces, and, notably, allows for…
-
10 Feb 2025 2 repositories listed Syntology ran 1 of 6 samples · 5 unverifiedSpecifically, we propose: 1) A large-scale rectified flow transformer for 3D shape generation, achieving state-of-the-art fidelity through training on extensive, high-quality data.
-
2 Dec 2024 2 repositories listed Syntology ran 4 of 10 samples · 6 unverifiedWe introduce a novel 3D generation method for versatile and high-quality 3D asset creation.
-
11 Nov 2024 2 repositories listed Syntology ran 0 of 10 samples · 10 unverifiedFor flexibility, we distill scale-conditioned part-aware 3D features for 3D part segmentation at multiple granularities.
-
10 Jun 2024 2 repositories listed Syntology ran 7 of 7 samples · 0 unverified · 7 pointer-only (licence)Recent 3D large reconstruction models (LRMs) can generate high-quality 3D content in sub-seconds by integrating multi-view diffusion models with scalable multi-view reconstructors.
-
27 Mar 2024 2 repositories listedWe tackle the challenge of efficiently reconstructing a 3D asset from a single image at millisecond speed.
-
15 Mar 2024 2 repositories listed Syntology ran 4 of 5 samples · 1 unverifiedBuilding upon our MVControl architecture, we employ a unique hybrid diffusion guidance method to direct the optimization process.
-
13 Dec 2023 2 repositories listedWe decouple domain-related guidance from the conditional guidance used in classifier-free guidance mechanisms to preserve open-world control guidance and unconditional guidance from the pre-trained model.
-
29 Nov 2023 2 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedWe justify that the refined 3D geometric priors aid in the 3D-aware capability of 2D diffusion priors, which in turn provides superior guidance for the refinement of 3D geometric priors.
-
7 Sep 2023 2 repositories listed Syntology ran 8 of 11 samples · 3 unverifiedIn this paper, we present a novel diffusion model called that generates multiview-consistent images from a single-view image.
-
28 Jul 2023 2 repositories listed Syntology ran 4 of 7 samples · 3 unverifiedVPP leverages structured voxel representation in the proposed Voxel Semantic Generator and the sparsity of unstructured point representation in the Point Upsampler, enabling efficient generation of multi-category…
-
30 May 2023 2 repositories listedThe recent advancements in image-text diffusion models have stimulated research interest in large-scale 3D generative models.
-
ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score Distillation25 May 2023 2 repositories listed Syntology ran 2 of 5 samples · 3 unverified · 3 pointer-only (licence)In comparison, VSD works well with various CFG weights as ancestral sampling from diffusion models and simultaneously improves the diversity and sample quality with a common CFG weight (i.
-
14 Nov 2022 2 repositories listed Syntology ran 2 of 4 samples · 2 unverifiedThis unique combination of text and shape guidance allows for increased control over the generation process.
-
12 Oct 2022 2 repositories listedTo advance 3D DDMs and make them useful for digital artists, we require (i) high generation quality, (ii) flexibility for manipulation and applications such as conditional synthesis and shape interpolation, and (iii)…
-
16 Jul 2025 1 repository listed Syntology ran 2 of 8 samples · 6 unverified · 8 pointer-only (licence)3D modeling is moving from virtual to physical.
-
2 Jun 2025 1 repository listed Syntology ran 1 of 8 samples · 7 unverifiedBuilding upon the 3D-aware discrete tokens, we innovatively construct a large-scale continuous training dataset named 3D-Alpaca, encompassing generation, comprehension, and editing, thus providing rich resources for…
-
26 May 2025 1 repository listedTo construct the weather particles, we first reconstruct a 3D scene using the edited images and then introduce a dynamic 4D Gaussian field to generate snowflakes, raindrops and fog in the scene.
-
23 May 2025 1 repository listed Syntology ran 2 of 10 samples · 8 unverifiedGenerating high-resolution 3D shapes using volumetric representations such as Signed Distance Functions (SDFs) presents substantial computational and memory challenges.
-
8 May 2025 1 repository listedOur experiments show that LegoGPT produces stable, diverse, and aesthetically pleasing LEGO designs that align closely with the input text prompts.
-
7 May 2025 1 repository listedThis streamlined framework demonstrates the effectiveness of S3D in generating high-quality 3D models from sketch inputs.
-
7 May 2025 1 repository listed Syntology ran 2 of 3 samples · 1 unverifiedRecent text-to-3D models can render high-quality assets, yet they still stumble on objects with complex attributes.
-
22 Apr 2025 1 repository listed Syntology ran 4 of 12 samples · 8 unverifiedWe discovered that these limitations stem from ambiguities in the 2D diffusion predictions during 3D avatar distillation, specifically: i) the avatar's appearance and geometry is underconstrained by the text input, and…
-
10 Apr 2025 1 repository listedBy using a perceptual loss, we effectively differentiate between positive and negative samples, leveraging the visual inconsistencies to improve 3D generation quality.
-
9 Apr 2025 1 repository listedWe release our enhanced dataset of approximately 500, 000 curated 3D models to facilitate further research on various downstream tasks in 3D computer vision.
-
5 Apr 2025 1 repository listedWith the ability of 4D and video generation, Video4DGen offers a powerful tool for applications in virtual reality, animation, and beyond.
-
3 Apr 2025 1 repository listedRecent advances in zero-shot text-to-3D generation have revolutionized 3D content creation by enabling direct synthesis from textual descriptions.
Syntology lines on 19 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections