Browse State-of-the-Art › Code Generation › Papers, page 8
Code Generation
Papers archive 2025-07-28
archive papers tagged: 1,697 · with a code link: 745 · where Syntology ran a sample: 280 (238 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (280 of 1,697 tagged: 238 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument)
Page 8 of 17: papers 701 to 800 of 1,697, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
14 Jan 2022 1 repository listed Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
27 Aug 2021 1 repository listed
-
1 Aug 2021 1 repository listed
-
1 Aug 2021 1 repository listed
-
18 Jul 2021 1 repository listed
-
19 Jun 2021 1 repository listed
-
15 Jun 2021 1 repository listed Syntology 12 ran (of which 1 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 1 pointer-only (licence)
-
8 Jun 2021 1 repository listed
-
1 Jun 2021 1 repository listed Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
18 May 2021 1 repository listed
-
17 May 2021 1 repository listed
-
27 Apr 2021 1 repository listed
-
8 Apr 2021 1 repository listed
-
6 Apr 2021 1 repository listed
-
29 Mar 2021 1 repository listed
-
9 Mar 2021 1 repository listed
-
5 Mar 2021 1 repository listed Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
21 Feb 2021 1 repository listed
-
21 Jan 2021 1 repository listed
-
18 Jan 2021 1 repository listed
-
13 Jan 2021 1 repository listed
-
1 Jan 2021 1 repository listed
-
1 Jan 2021 1 repository listed
-
1 Jan 2021 1 repository listed
-
7 Dec 2020 1 repository listed
-
22 Sep 2020 1 repository listed
-
22 Sep 2020 1 repository listed
-
16 Sep 2020 1 repository listed
-
25 Jun 2020 1 repository listed
-
18 Jun 2020 1 repository listed Syntology 3 ran (of which 2 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
12 May 2020 1 repository listed
-
14 Jan 2020 1 repository listed
-
26 Jun 2019 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
4 Jun 2019 1 repository listed
-
16 Apr 2019 1 repository listed
-
14 Mar 2019 1 repository listed
-
3 Mar 2019 1 repository listed
-
15 Nov 2018 1 repository listed
-
14 Nov 2018 1 repository listed
-
29 Aug 2018 1 repository listed
-
27 Jun 2018 1 repository listed
-
22 Dec 2017 1 repository listed
-
8 Nov 2017 1 repository listed
-
25 Apr 2017 1 repository listed
-
1 Jun 2016 1 repository listed
-
CUDA-L1: Improving CUDA Optimization via Contrastive Reinforcement Learning18 Jul 2025 0 repositories listed
-
Towards Formal Verification of LLM-Generated Code from Natural Language Prompts17 Jul 2025 0 repositories listed
-
MERA Code: A Unified Framework for Evaluating Code Generation Across Tasks16 Jul 2025 0 repositories listed
-
Scaling Up RL: Unlocking Diverse Reasoning in LLMs via Prolonged Training16 Jul 2025 0 repositories listed
-
CodeJudgeBench: Benchmarking LLM-as-a-Judge for Coding Tasks14 Jul 2025 0 repositories listed
-
Turning the Tide: Repository-based Code Reflection14 Jul 2025 0 repositories listed
-
Multilingual Multimodal Software Developer for Code Generation11 Jul 2025 0 repositories listed
-
OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique11 Jul 2025 0 repositories listed
-
Automating MD simulations for Proteins using Large language Models: NAMD-Agent10 Jul 2025 0 repositories listed
-
ArtifactsBench: Bridging the Visual-Interactive Gap in LLM Code Generation Evaluation7 Jul 2025 0 repositories listed
-
Learn Globally, Speak Locally: Bridging the Gaps in Multilingual Reasoning7 Jul 2025 0 repositories listed
-
CORE: Benchmarking LLMs Code Reasoning Capabilities through Static Analysis Tasks3 Jul 2025 0 repositories listed
-
LLM-based Realistic Safety-Critical Driving Video Generation2 Jul 2025 0 repositories listed
-
A Large Language Model-Empowered Agent for Reliable and Robust Structural Analysis27 Jun 2025 0 repositories listed
-
Where to find Grokking in LLM Pretraining? Monitor Memorization-to-Generalization without Test26 Jun 2025 0 repositories listed
-
SACL: Understanding and Combating Textual Bias in Code Retrieval with Semantic-Augmented Reranking and Localization25 Jun 2025 0 repositories listed
-
SV-LLM: An Agentic Approach for SoC Security Verification using Large Language Models25 Jun 2025 0 repositories listed
-
QHackBench: Benchmarking Large Language Models for Quantum Code Generation Using PennyLane Hackathon Challenges24 Jun 2025 0 repositories listed
-
Why Do Open-Source LLMs Struggle with Data Analysis? A Systematic Empirical Study24 Jun 2025 0 repositories listed
-
Steering Conceptual Bias via Transformer Latent-Subspace Activation23 Jun 2025 0 repositories listed
-
The Debugging Decay Index: Rethinking Debugging Strategies for Code LLMs23 Jun 2025 0 repositories listed
-
Use Property-Based Testing to Bridge LLM Code Generation and Validation23 Jun 2025 0 repositories listed
-
RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation22 Jun 2025 0 repositories listed
-
Massive Supervised Fine-tuning Experiments Reveal How Data, Layer, and Training Factors Shape LLM Alignment Quality17 Jun 2025 0 repositories listed
-
A Technical Study into Small Reasoning Language Models16 Jun 2025 0 repositories listed
-
FrontendBench: A Benchmark for Evaluating LLMs on Front-End Development via Automatic Evaluation16 Jun 2025 0 repositories listed
-
How Does LLM Reasoning Work for Code? A Survey and a Call to Action16 Jun 2025 0 repositories listed
-
Structured Program Synthesis using LLMs: Results and Insights from the IPARC Challenge15 Jun 2025 0 repositories listed
-
The Safety Reminder: A Soft Prompt to Reactivate Delayed Safety Awareness in Vision-Language Models15 Jun 2025 0 repositories listed
-
code_transformed: The Influence of Large Language Models on Code13 Jun 2025 0 repositories listed
-
ReVeal: Self-Evolving Code Agents via Iterative Generation-Verification13 Jun 2025 0 repositories listed
-
LLM-as-a-Judge for Reference-less Automatic Code Validation and Refinement for Natural Language to Bash in IT Automation12 Jun 2025 0 repositories listed
-
Prompt Variability Effects On LLM Code Generation11 Jun 2025 0 repositories listed
-
Reasoning as a Resource: Optimizing Fast and Slow Thinking in Code Generation Models11 Jun 2025 0 repositories listed
-
Understanding Software Engineering Agents Through the Lens of Traceability: An Empirical Study10 Jun 2025 0 repositories listed
-
Edit Flows: Flow Matching with Edit Operations10 Jun 2025 0 repositories listed
-
Exploring the Capabilities of the Frontier Large Language Models for Nuclear Energy Research10 Jun 2025 0 repositories listed
-
Technical Report for Argoverse2 Scenario Mining Challenges on Iterative Error Correction and Spatially-Aware Prompting10 Jun 2025 0 repositories listed
-
Worst-Case Symbolic Constraints Analysis and Generalisation with Large Language Models9 Jun 2025 0 repositories listed
-
Repeton: Structured Bug Repair with ReAct-Guided Patch-and-Test Cycles9 Jun 2025 0 repositories listed
-
ProtocolLLM: RTL Benchmark for SystemVerilog Generation of Communication Protocols9 Jun 2025 0 repositories listed
-
VeriLoC: Line-of-Code Level Prediction of Hardware Design Quality from Verilog Code8 Jun 2025 0 repositories listed
-
Can LLMs Generate Reliable Test Case Generators? A Study on Competition-Level Programming Problems7 Jun 2025 0 repositories listed
-
Can Theoretical Physics Research Benefit from Language Agents?6 Jun 2025 0 repositories listed
-
CP-Bench: Evaluating Large Language Models for Constraint Modelling6 Jun 2025 0 repositories listed
-
SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code6 Jun 2025 0 repositories listed
-
Demonstrations of Integrity Attacks in Multi-Agent Systems5 Jun 2025 0 repositories listed
-
hdl2v: A Code Translation Dataset for Enhanced LLM Verilog Generation5 Jun 2025 0 repositories listed
-
ScaleRTL: Scaling LLMs with Reasoning Data and Test-Time Compute for Accurate RTL Code Generation5 Jun 2025 0 repositories listed
-
CETBench: A Novel Dataset constructed via Transformations over Programs for Benchmarking LLMs for Code-Equivalence Checking4 Jun 2025 0 repositories listed
-
From Virtual Agents to Robot Teams: A Multi-Robot Framework Evaluation in High-Stakes Healthcare Context4 Jun 2025 0 repositories listed
-
Generating Automotive Code: Large Language Models for Software Development and Verification in Safety-Critical Systems4 Jun 2025 0 repositories listed
-
VisCoder: Fine-Tuning LLMs for Executable Python Visualization Code Generation4 Jun 2025 0 repositories listed
-
How do Pre-Trained Models Support Software Engineering? An Empirical Study in Hugging Face3 Jun 2025 0 repositories listed
Syntology lines on 7 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.