Browse State-of-the-Art › Code Classification
Code Classification
15 papers with code · 0 benchmarks · 8 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
8 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
15 shown of 15 papers with code (39 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
11 Jun 2022 2 repositories listedDistribution shift has been a longstanding challenge for the reliable deployment of deep learning (DL) models due to unexpected accuracy degradation.
-
6 Apr 2020 2 repositories listedcode2vec is a recently released embedding approach that uses the proxy task of method name prediction to map Java methods to feature vectors.
-
12 Mar 2025 1 repository listedLLMs perform exceptionally well in the CASTLE dataset when identifying vulnerabilities in small code snippets.
-
10 Jan 2024 1 repository listedResearchers have investigated the potential of leveraging pre-trained language models, such as CodeBERT, to enhance source code-related tasks.
-
7 Dec 2023 1 repository listed Syntology ran 19 of 29 samples · 10 unverifiedWe propose a graph-filter-based self-attention (GFSA) to learn a general yet effective one, whose complexity, however, is slightly larger than that of the original self-attention mechanism.
-
29 Aug 2023 1 repository listedTissue phenotyping is a fundamental computational pathology (CPath) task in learning objective characterizations of histopathologic biomarkers in anatomic pathology.
-
23 May 2023 1 repository listedThe effectiveness of the proposed method is verified on two program understanding tasks including code clone detection and code classification, and it outperforms current state-of-the-arts by large margins.
-
8 May 2023 1 repository listedThese findings show that early layers can be used to obtain better results using the same resources, as well as to reduce resource usage during fine-tuning and inference.
-
7 May 2023 1 repository listedIn this study, we propose to represent AST as a heterogeneous directed hypergraph (HDHG) and process the graph by heterogeneous directed hypergraph neural network (HDHGN) for code classification.
-
6 Oct 2022 1 repository listedData augmentation has been a popular approach to supplement training data in domains such as computer vision and NLP.
-
6 Oct 2022 1 repository listedGraph neural network (GNN)-based graph learning has been popular in natural language and programming language processing, particularly in text and source code classification.
-
22 Mar 2022 1 repository listedHowever, currently, a comprehensive and systematic study on evaluating different program representation techniques across diverse tasks is still missed.
-
25 Jan 2022 1 repository listedA range of applications for automatic machine learning need the generation process to be controllable.
-
25 May 2021 1 repository listed Syntology ran 0 of 9 samples · 9 unverifiedIn addition to its large scale, CodeNet has a rich set of high-quality annotations to benchmark and help accelerate research in AI techniques for a variety of critical coding tasks, including code similarity and…
-
21 Sep 2018 1 repository listedDetermining the programming language of a source code file has been considered in the research community; it has been shown that Machine Learning (ML) and Natural Language Processing (NLP) algorithms can be effective in…
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections