Browse State-of-the-Art › Code Generation › Papers, page 11
Code Generation
Papers archive 2025-07-28
archive papers tagged: 1,697 · with a code link: 745 · where Syntology ran a sample: 280 (238 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (280 of 1,697 tagged: 238 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument)
Page 11 of 17: papers 1,001 to 1,100 of 1,697, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Hidden Darkness in LLM-Generated Designs: Exploring Dark Patterns in Ecommerce Web Components Generated by LLMs19 Feb 2025 0 repositories listed
-
Boost, Disentangle, and Customize: A Robust System2-to-System1 Pipeline for Code Generation18 Feb 2025 0 repositories listed
-
EquiBench: Benchmarking Large Language Models' Understanding of Program Semantics via Equivalence Checking18 Feb 2025 0 repositories listed
-
GSCE: A Prompt Framework with Enhanced Reasoning for Reliable LLM-driven Drone Control18 Feb 2025 0 repositories listed
-
Sens-Merging: Sensitivity-Guided Parameter Balancing for Merging Large Language Models18 Feb 2025 0 repositories listed
-
The Role of GitHub Copilot on Software Development: A Perspective on Productivity, Security, Best Practices and Future Directions18 Feb 2025 0 repositories listed
-
LLM4EFFI: Leveraging Large Language Models to Enhance Code Efficiency and Correctness17 Feb 2025 0 repositories listed
-
UnitCoder: Scalable Iterative Code Synthesis with Unit Test Guidance17 Feb 2025 0 repositories listed
-
An Interpretable Automated Mechanism Design Framework with Large Language Models16 Feb 2025 0 repositories listed
-
Diversified Sampling Improves Scaling LLM inference16 Feb 2025 0 repositories listed
-
Performance Review on LLM for solving leetcode problems16 Feb 2025 0 repositories listed
-
Automated Visualization Code Synthesis via Multi-Path Reasoning and Feedback-Driven Optimization16 Feb 2025 0 repositories listed
-
1bit-Merging: Dynamic Quantized Merging for Large Language Models15 Feb 2025 0 repositories listed
-
3D-Grounded Vision-Language Framework for Robotic Task Planning: Automated Prompt Synthesis and Supervised Reasoning13 Feb 2025 0 repositories listed
-
CRANE: Reasoning with constrained LLM generation13 Feb 2025 0 repositories listed
-
RefineCoder: Iterative Improving of Large Language Models via Adaptive Critique Refinement for Code Generation13 Feb 2025 0 repositories listed
-
Enhancing LLM Character-Level Manipulation via Divide and Conquer12 Feb 2025 0 repositories listed
-
From PowerPoint UI Sketches to Web-Based Applications: Pattern-Driven Code Generation for GIS Dashboard Development Using Knowledge-Augmented LLMs, Context-Aware Visual Prompting, and the React Framework12 Feb 2025 0 repositories listed
-
Verifying LLM-Generated Code in the Context of Software Verification with Ada/SPARK11 Feb 2025 0 repositories listed
-
Can LLMs Replace Human Evaluators? An Empirical Study of LLM-as-a-Judge in Software Engineering10 Feb 2025 0 repositories listed
-
Cardiverse: Harnessing LLMs for Novel Card Game Prototyping10 Feb 2025 0 repositories listed
-
LessLeak-Bench: A First Investigation of Data Leakage in LLMs Across 83 Software Engineering Benchmarks10 Feb 2025 0 repositories listed
-
SnipGen: A Mining Repository Framework for Evaluating LLMs for Code10 Feb 2025 0 repositories listed
-
What I cannot execute, I do not understand: Training and Evaluating LLMs on Program Execution Traces10 Feb 2025 0 repositories listed
-
Benchmarking Prompt Engineering Techniques for Secure Code Generation with GPT Models9 Feb 2025 0 repositories listed
-
Mitigating Sensitive Information Leakage in LLMs4Code through Machine Unlearning9 Feb 2025 0 repositories listed
-
Proving the Coding Interview: A Benchmark for Formally Verified Code Generation8 Feb 2025 0 repositories listed
-
Optimistic Gradient Learning with Hessian Corrections for High-Dimensional Black-Box Optimization7 Feb 2025 0 repositories listed
-
Refining Integration-by-Parts Reduction of Feynman Integrals with Machine Learning7 Feb 2025 0 repositories listed
-
Large Language Model Guided Self-Debugging Code Generation5 Feb 2025 0 repositories listed
-
LLMs can be easily Confused by Instructional Distractions5 Feb 2025 0 repositories listed
-
Path Planning for Masked Diffusion Model Sampling5 Feb 2025 0 repositories listed
-
Teaching Language Models to Critique via Reinforcement Learning5 Feb 2025 0 repositories listed
-
Can LLMs Maintain Fundamental Abilities under KV Cache Compression?4 Feb 2025 0 repositories listed
-
Analysis of Student-LLM Interaction in a Software Engineering Project3 Feb 2025 0 repositories listed
-
Next Steps in LLM-Supported Java Verification3 Feb 2025 0 repositories listed
-
PlotGen: Multi-Agent LLM-based Scientific Data Visualization via Multimodal Feedback3 Feb 2025 0 repositories listed
-
Process-Supervised Reinforcement Learning for Code Generation3 Feb 2025 0 repositories listed
-
SE Arena: An Interactive Platform for Evaluating Foundation Models in Software Engineering3 Feb 2025 0 repositories listed
-
Security and Quality in LLM-Generated Code: A Multi-Language, Multi-Model Analysis3 Feb 2025 0 repositories listed
-
Toward Neurosymbolic Program Comprehension3 Feb 2025 0 repositories listed
-
Analysis of LLMs vs Human Experts in Requirements Engineering31 Jan 2025 0 repositories listed
-
Importing Phantoms: Measuring LLM Package Hallucination Vulnerabilities31 Jan 2025 0 repositories listed
-
Towards Adaptive Self-Improvement for Smarter Energy Systems31 Jan 2025 0 repositories listed
-
Enhancing Large Language Model Efficiencyvia Symbolic Compression: A Formal Approach Towards Interpretability30 Jan 2025 0 repositories listed
-
Statistical multi-metric evaluation and visualization of LLM system predictive performance30 Jan 2025 0 repositories listed
-
GLLM: Self-Corrective G-Code Generation using Large Language Models with User Feedback29 Jan 2025 0 repositories listed
-
Programming by Examples Meets Historical Linguistics: A Large Language Model Based Approach to Sound Law Induction27 Jan 2025 0 repositories listed
-
Advancing Generative Artificial Intelligence and Large Language Models for Demand Side Management with Internet of Electric Vehicles26 Jan 2025 0 repositories listed
-
24 Jan 2025 0 repositories listed
-
Chain of Grounded Objectives: Bridging Process and Goal-oriented Prompting for Code Generation23 Jan 2025 0 repositories listed
-
Pseudocode-Injection Magic: Enabling LLMs to Tackle Graph Computational Tasks23 Jan 2025 0 repositories listed
-
Revisit Self-Debugging with Self-Generated Tests for Code Generation22 Jan 2025 0 repositories listed
-
Consolidating TinyML Lifecycle with Large Language Models: Reality, Illusion, or Opportunity?20 Jan 2025 0 repositories listed
-
20 Jan 2025 0 repositories listed
-
Towards Advancing Code Generation with Large Language Models: A Research Roadmap20 Jan 2025 0 repositories listed
-
SOP-Agent: Empower General Purpose AI Agent with Domain-Specific SOPs16 Jan 2025 0 repositories listed
-
Leveraging Metamemory Mechanisms for Enhanced Data-Free Code Generation in LLMs14 Jan 2025 0 repositories listed
-
The Invisible Hand: Unveiling Provider Bias in Large Language Models for Code Generation14 Jan 2025 0 repositories listed
-
Evaluating Agent-based Program Repair at Google13 Jan 2025 0 repositories listed
-
Guided Code Generation with LLMs: A Multi-Agent Framework for Complex Code Tasks11 Jan 2025 0 repositories listed
-
BioAgents: Democratizing Bioinformatics Analysis with Multi-Agent Systems10 Jan 2025 0 repositories listed
-
Dafny as Verification-Aware Intermediate Language for Code Generation10 Jan 2025 0 repositories listed
-
Deriving Coding-Specific Sub-Models from LLMs using Resource-Efficient Pruning9 Jan 2025 0 repositories listed
-
Do Code LLMs Understand Design Patterns?8 Jan 2025 0 repositories listed
-
8 Jan 2025 0 repositories listed Syntology 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 15 harvested samples) · 14 pointer-only (licence)
-
Robotic Programmer: Video Instructed Policy Code Generation for Robotic Manipulation8 Jan 2025 0 repositories listed
-
The Future of AI: Exploring the Potential of Large Concept Models8 Jan 2025 0 repositories listed
-
ChronoLLM: A Framework for Customizing Large Language Model for Digital Twins generalization based on PyChrono7 Jan 2025 0 repositories listed
-
Practical Design and Benchmarking of Generative AI Applications for Surgical Billing and Coding7 Jan 2025 0 repositories listed
-
CodeVision: Detecting LLM-Generated Code Using 2D Token Probability Maps and Vision Models6 Jan 2025 0 repositories listed
-
RTLSquad: Multi-Agent Based Interpretable RTL Design6 Jan 2025 0 repositories listed
-
Cracks in The Stack: Hidden Vulnerabilities and Licensing Risks in LLM Pre-Training Datasets5 Jan 2025 0 repositories listed
-
ToolHop: A Query-Driven Benchmark for Evaluating Large Language Models in Multi-Hop Tool Use5 Jan 2025 0 repositories listed
-
A Survey on Large Language Models with some Insights on their Capabilities and Limitations3 Jan 2025 0 repositories listed
-
CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings2 Jan 2025 0 repositories listed
-
Dynamic Scaling of Unit Tests for Code Reward Modeling2 Jan 2025 0 repositories listed
-
Beyond Text: Implementing Multimodal Large Language Model-Powered Multi-Agent Systems Using a No-Code Platform1 Jan 2025 0 repositories listed
-
Enabling New HDLs with Agents31 Dec 2024 0 repositories listed
-
SecBench: A Comprehensive Multi-Dimensional Benchmarking Dataset for LLMs in Cybersecurity30 Dec 2024 0 repositories listed
-
Thinking Before Running! Efficient Code Generation with Thorough Exploration and Optimal Refinement30 Dec 2024 0 repositories listed
-
Enhancing Code LLMs with Reinforcement Learning in Code Generation: A Survey29 Dec 2024 0 repositories listed
-
SocRATES: Towards Automated Scenario-based Testing of Social Navigation Algorithms27 Dec 2024 0 repositories listed
-
SketchFill: Sketch-Guided Code Generation for Imputing Derived Missing Values26 Dec 2024 0 repositories listed
-
How Propense Are Large Language Models at Producing Code Smells? A Benchmarking Study25 Dec 2024 0 repositories listed
-
Renaissance of Literate Programming in the Era of LLMs: Enhancing LLM-Based Code Generation in Large-Scale Projects25 Dec 2024 0 repositories listed
-
M-Ped: Multi-Prompt Ensemble Decoding for Large Language Models24 Dec 2024 0 repositories listed
-
Cypress Copilot: Development of an AI Assistant for Boosting Productivity and Transforming Web Application Testing23 Dec 2024 0 repositories listed
-
Emerging Security Challenges of Large Language Models23 Dec 2024 0 repositories listed
-
AIGCodeSet: A New Annotated Dataset for AI Generated Code Detection21 Dec 2024 0 repositories listed
-
Beyond the Sum: Unlocking AI Agents Potential Through Market Forces19 Dec 2024 0 repositories listed
-
Energy consumption of code small language models serving with runtime engines and execution providers19 Dec 2024 0 repositories listed
-
Helping LLMs Improve Code Generation Using Feedback from Testing and Static Analysis19 Dec 2024 0 repositories listed
-
HPC-Coder-V2: Studying Code LLMs Across Low-Resource Parallel Languages19 Dec 2024 0 repositories listed
-
Tree-of-Code: A Tree-Structured Exploring Framework for End-to-End Code Generation and Execution in Complex Task Handling19 Dec 2024 0 repositories listed
-
Why We Build Local Large Language Models: An Observational Analysis from 35 Japanese and Multilingual LLMs19 Dec 2024 0 repositories listed
-
18 Dec 2024 0 repositories listed
-
Channel Merging: Preserving Specialization for Merged Experts18 Dec 2024 0 repositories listed
-
GenX: Mastering Code and Test Generation with Execution Feedback18 Dec 2024 0 repositories listed
-
Syzygy: Dual Code-Test C to (safe) Rust Translation using LLMs and Dynamic Analysis18 Dec 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.