Browse State-of-the-Art › Text to 3D
Text to 3D
102 papers with code · 1 benchmark · 5 datasets archive 2025-07-28
Task involves generating 3D objects based on the text prompt provided to the system.
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
1 leaderboard table shown for this task, 1 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| T³Bench (6 rows) | ProlificDreamer | 0/1 Deep Neural Networks via Block Coordinate Descent | — | — | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
5 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 102 papers with code (314 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
5 Jun 2024 4 repositories listed Syntology ran 6 of 10 samples · 4 unverified · 5 pointer-only (licence)Besides the generative capabilities of diffusion priors, motivated by the unique time-symmetry properties of rectified flow models, a variant of our method can additionally perform image inversion.
-
29 Sep 2022 4 repositories listed Syntology ran 10 of 11 samples · 1 unverifiedUsing this loss in a DeepDream-like procedure, we optimize a randomly-initialized 3D model (a Neural Radiance Field, or NeRF) via gradient descent such that its 2D renderings from random angles achieve a low loss.
-
29 May 2024 3 repositories listedWe present Cephalo, a series of multimodal vision large language models (V-LLMs) designed for materials science applications, integrating visual and linguistic data for enhanced understanding.
-
24 Mar 2023 3 repositories listed Syntology ran 7 of 18 samples · 11 unverifiedKey to Fantasia3D is the disentangled modeling and learning of geometry and appearance.
-
26 Nov 2024 2 repositories listed Syntology ran 0 of 6 samples · 6 unverifiedGenerating high-fidelity 3D content from text prompts remains a significant challenge in computer vision due to the limited size, diversity, and annotation depth of the existing datasets.
-
11 Apr 2024 2 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)Scalable annotation approaches are crucial for constructing extensive 3D-text datasets, facilitating a broader range of applications.
-
15 Mar 2024 2 repositories listed Syntology ran 4 of 5 samples · 1 unverifiedBuilding upon our MVControl architecture, we employ a unique hybrid diffusion guidance method to direct the optimization process.
-
29 Nov 2023 2 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedWe justify that the refined 3D geometric priors aid in the 3D-aware capability of 2D diffusion priors, which in turn provides superior guidance for the refinement of 3D geometric priors.
-
7 Sep 2023 2 repositories listed Syntology ran 8 of 11 samples · 3 unverifiedIn this paper, we present a novel diffusion model called that generates multiview-consistent images from a single-view image.
-
ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score Distillation25 May 2023 2 repositories listed Syntology ran 2 of 5 samples · 3 unverified · 3 pointer-only (licence)In comparison, VSD works well with various CFG weights as ancestral sampling from diffusion models and simultaneously improves the diversity and sample quality with a common CFG weight (i.
-
24 Mar 2023 2 repositories listedIn this work, we investigate the problem of creating high-fidelity 3D content from only a single image.
-
14 Nov 2022 2 repositories listed Syntology ran 2 of 4 samples · 2 unverifiedThis unique combination of text and shape guidance allows for increased control over the generation process.
-
8 May 2025 1 repository listedOur experiments show that LegoGPT produces stable, diverse, and aesthetically pleasing LEGO designs that align closely with the input text prompts.
-
7 May 2025 1 repository listed Syntology ran 2 of 3 samples · 1 unverifiedRecent text-to-3D models can render high-quality assets, yet they still stumble on objects with complex attributes.
-
25 Apr 2025 1 repository listedWe also implement the pipeline to run real-time on an autonomous vehicle and demonstrate that our approach can be used for object-goal navigation on previously unseen real-world environments.
-
24 Apr 2025 1 repository listedAs for texture branch, we use RGB images as input to obtain the textured mesh.
-
22 Apr 2025 1 repository listed Syntology ran 4 of 12 samples · 8 unverifiedWe discovered that these limitations stem from ambiguities in the 2D diffusion predictions during 3D avatar distillation, specifically: i) the avatar's appearance and geometry is underconstrained by the text input, and…
-
3 Apr 2025 1 repository listedRecent advances in zero-shot text-to-3D generation have revolutionized 3D content creation by enabling direct synthesis from textual descriptions.
-
27 Mar 2025 1 repository listedAiming at overcoming the data shortage, we propose a novel training scheme, termed as Progressive Rendering Distillation (PRD), eliminating the need for 3D ground-truths by distilling multi-view diffusion models and…
-
12 Mar 2025 1 repository listed Syntology ran 0 of 9 samples · 9 unverifiedTo overcome this, we introduce RewardSDS, a novel approach that weights noise samples based on alignment scores from a reward model, producing a weighted SDS loss.
-
3 Mar 2025 1 repository listedThe normal maps are then used to reconstruct a 3D mesh, and the multi-view images provide texture mapping, resulting in a complete 3D model.
-
24 Feb 2025 1 repository listedThis database encompasses 969 validated 3D assets generated from 170 prompts via 6 popular text-to-3D asset generation models, and corresponding subjective quality ratings for these assets from the perspectives of…
-
16 Dec 2024 1 repository listedWe propose StrandHead, a novel text to 3D head avatar generation method capable of generating disentangled 3D hair with strand representation.
-
16 Dec 2024 1 repository listedIn this paper, we present BlenderLLM, a novel framework for training LLMs specifically for CAD tasks leveraging a self-improvement methodology.
-
11 Dec 2024 1 repository listed Syntology ran 5 of 6 samples · 1 unverifiedWe introduce Generate Any Scene, a framework that systematically enumerates scene graphs representing a vast array of visual scenes, spanning realistic to imaginative compositions.
-
9 Dec 2024 1 repository listed Syntology ran 1 of 18 samples · 17 unverifiedWe design a lightweight 3D texture field to synthesize visual and tactile textures, guided by 2D diffusion model priors on both visual and tactile domains.
-
3 Dec 2024 1 repository listedT3DEM is the most crucial step in determining the quality of Emo3D generation and encompasses three key challenges: Expression Diversity, Emotion-Content Consistency, and Expression Fluidity.
-
24 Oct 2024 1 repository listed Syntology ran 0 of 7 samples · 7 unverifiedMulti-view image diffusion models have significantly advanced open-domain 3D object generation.
-
11 Oct 2024 1 repository listed Syntology ran 7 of 8 samples · 1 unverifiedThe creation of complex 3D scenes tailored to user specifications has been a tedious and challenging task with traditional 3D modeling tools.
-
25 Sep 2024 1 repository listedThe core of this framework lies in Skeleton-guided Score Distillation and Hybrid 3D Gaussian Avatar representation.
Syntology lines on 17 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections