Browse State-of-the-Art › Language Modelling › Papers, page 96
Language Modelling
Papers archive 2025-07-28
archive papers tagged: 17,610 · with a code link: 7,012 · where Syntology ran a sample: 2,428 (2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,428 of 17,610 tagged: 2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument)
Page 96 of 177: papers 9,501 to 9,600 of 17,610, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
A Small Claims Court for the NLP: Judging Legal Text Classification Strategies With Small Datasets9 Sep 2024 0 repositories listed
-
DeepFM-Crispr: Prediction of CRISPR On-Target Effects via Deep Learning9 Sep 2024 0 repositories listed
-
Doppelgänger's Watch: A Split Objective Approach to Large Language Models9 Sep 2024 0 repositories listed
-
Ethereum Fraud Detection via Joint Transaction Language Model and Graph Representation Learning9 Sep 2024 0 repositories listed
-
Evidence from fMRI Supports a Two-Phase Abstraction Process in Language Models9 Sep 2024 0 repositories listed
-
FairHome: A Fair Housing and Fair Lending Dataset9 Sep 2024 0 repositories listed
-
LegiLM: A Fine-Tuned Legal Language Model for Data Compliance9 Sep 2024 0 repositories listed
-
MLLM-LLaVA-FL: Multimodal Large Language Model Assisted Federated Learning9 Sep 2024 0 repositories listed
-
Regression with Large Language Models for Materials and Molecular Property Prediction9 Sep 2024 0 repositories listed
-
SongCreator: Lyrics-based Universal Song Generation9 Sep 2024 0 repositories listed
-
A Hetero-functional Graph Resilience Analysis for Convergent Systems-of-Systems8 Sep 2024 0 repositories listed
-
ELMS: Elasticized Large Language Models On Mobile Devices8 Sep 2024 0 repositories listed
-
Interactive Machine Teaching by Labeling Rules and Instances8 Sep 2024 0 repositories listed
-
STLLM-DF: A Spatial-Temporal Large Language Model with Diffusion for Enhanced Multi-Mode Traffic System Forecasting8 Sep 2024 0 repositories listed
-
Achieving Peak Performance for Large Language Models: A Systematic Review7 Sep 2024 0 repositories listed
-
Reward Guidance for Reinforcement Learning Tasks Based on Large Language Models: The LMGT Framework7 Sep 2024 0 repositories listed
-
MuAP: Multi-step Adaptive Prompt Learning for Vision-Language Model with Missing Modality7 Sep 2024 0 repositories listed
-
7 Sep 2024 0 repositories listed
-
VidLPRO: A Video-Language Pre-training Framework for Robotic and Laparoscopic Surgery7 Sep 2024 0 repositories listed
-
Confidential Computing on NVIDIA Hopper GPUs: A Performance Benchmark Study6 Sep 2024 0 repositories listed
-
Customizing Large Language Model Generation Style using Parameter-Efficient Finetuning6 Sep 2024 0 repositories listed
-
GALLa: Graph Aligned Large Language Models for Improved Source Code Understanding6 Sep 2024 0 repositories listed
-
How Does Code Pretraining Affect Language Model Task Performance?6 Sep 2024 0 repositories listed
-
Retrieval Augmented Generation-Based Incident Resolution Recommendation System for IT Support6 Sep 2024 0 repositories listed
-
Using Large Language Models to Generate Authentic Multi-agent Knowledge Work Datasets6 Sep 2024 0 repositories listed
-
A Fused Large Language Model for Predicting Startup Success5 Sep 2024 0 repositories listed
-
Bypassing DARCY Defense: Indistinguishable Universal Adversarial Triggers5 Sep 2024 0 repositories listed
-
5 Sep 2024 0 repositories listed
-
N-gram Prediction and Word Difference Representations for Language Modeling5 Sep 2024 0 repositories listed
-
A Medical Multimodal Large Language Model for Pediatric Pneumonia4 Sep 2024 0 repositories listed
-
Accelerating Large Language Model Training with Hybrid GPU-based Compression4 Sep 2024 0 repositories listed
-
Creating Domain-Specific Translation Memories for Machine Translation Fine-tuning: The TRENCARD Bilingual Cardiology Corpus4 Sep 2024 0 repositories listed
-
Exploring Sentiment Dynamics and Predictive Behaviors in Cryptocurrency Discussions by Few-Shot Learning with Large Language Models4 Sep 2024 0 repositories listed
-
Historical German Text Normalization Using Type- and Token-Based Language Modeling4 Sep 2024 0 repositories listed
-
Irrelevant Alternatives Bias Large Language Model Hiring Decisions4 Sep 2024 0 repositories listed
-
ISO: Overlap of Computation and Communication within Seqenence For LLM Inference4 Sep 2024 0 repositories listed
-
Masked Diffusion Models are Secretly Time-Agnostic Masked Models and Exploit Inaccurate Categorical Sampling4 Sep 2024 0 repositories listed
-
Oddballness: universal anomaly detection with language models4 Sep 2024 0 repositories listed
-
Pre-training data selection for biomedical domain adaptation using journal impact metrics4 Sep 2024 0 repositories listed
-
Scaling Laws for Economic Productivity: Experimental Evidence in LLM-Assisted Translation4 Sep 2024 0 repositories listed
-
4 Sep 2024 0 repositories listed Syntology 4 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Standing on the Shoulders of Giants: Reprogramming Visual-Language Model for General Deepfake Detection4 Sep 2024 0 repositories listed
-
An Implementation of Werewolf Agent That does not Truly Trust LLMs3 Sep 2024 0 repositories listed
-
Approximating mutual information of high-dimensional variables using learned representations3 Sep 2024 0 repositories listed
-
Dynamic Motion Synthesis: Masked Audio-Text Conditioned Spatio-Temporal Transformers3 Sep 2024 0 repositories listed
-
Foundations of Large Language Model Compression -- Part 1: Weight Quantization3 Sep 2024 0 repositories listed
-
LASP: Surveying the State-of-the-Art in Large Language Model-Assisted AI Planning3 Sep 2024 0 repositories listed
-
SmileyLlama: Modifying Large Language Models for Directed Chemical Space Exploration3 Sep 2024 0 repositories listed
-
Therapy as an NLP Task: Psychologists' Comparison of LLMs and Human Peers in CBT3 Sep 2024 0 repositories listed
-
VSLLaVA: a pipeline of large multimodal foundation model for industrial vibration signal analysis3 Sep 2024 0 repositories listed
-
A Perspective on Literary Metaphor in the Context of Generative AI2 Sep 2024 0 repositories listed
-
Balancing Performance and Efficiency: A Multimodal Large Language Model Pruning Method based Image Text Interaction2 Sep 2024 0 repositories listed
-
DPDEdit: Detail-Preserved Diffusion Models for Multimodal Fashion Image Editing2 Sep 2024 0 repositories listed
-
2 Sep 2024 0 repositories listed Syntology 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Grounding Language Models in Autonomous Loco-manipulation Tasks2 Sep 2024 0 repositories listed
-
Imitating Language via Scalable Inverse Reinforcement Learning2 Sep 2024 0 repositories listed
-
LATEX-GCL: Large Language Models (LLMs)-Based Data Augmentation for Text-Attributed Graph Contrastive Learning2 Sep 2024 0 repositories listed
-
Revisiting SMoE Language Models by Evaluating Inefficiencies with Task Specific Expert Pruning2 Sep 2024 0 repositories listed
-
User-Specific Dialogue Generation with User Profile-Aware Pre-Training Model and Parameter-Efficient Fine-Tuning2 Sep 2024 0 repositories listed
-
Comparing Discrete and Continuous Space LLMs for Speech Recognition1 Sep 2024 0 repositories listed
-
Multimodal Multi-turn Conversation Stance Detection: A Challenge Dataset and Effective Model1 Sep 2024 0 repositories listed
-
The Dark Side of Human Feedback: Poisoning Large Language Models via User Inputs1 Sep 2024 0 repositories listed
-
From Prediction to Application: Language Model-based Code Knowledge Tracing with Domain Adaptive Pre-Training and Automatic Feedback System with Pedagogical Prompting for Comprehensive Programming Education31 Aug 2024 0 repositories listed
-
Testing and Evaluation of Large Language Models: Correctness, Non-Toxicity, and Fairness31 Aug 2024 0 repositories listed
-
Forget to Flourish: Leveraging Machine-Unlearning on Pretrained Language Models for Privacy Leakage30 Aug 2024 0 repositories listed
-
InkubaLM: A small language model for low-resource African languages30 Aug 2024 0 repositories listed
-
Joint Estimation and Prediction of City-wide Delivery Demand: A Large Language Model Empowered Graph-based Learning Approach30 Aug 2024 0 repositories listed
-
Language-guided Scale-aware MedSegmentor for Lesion Segmentation in Medical Imaging30 Aug 2024 0 repositories listed
-
Novel-WD: Exploring acquisition of Novel World Knowledge in LLMs Using Prefix-Tuning30 Aug 2024 0 repositories listed
-
OrthoDoc: Multimodal Large Language Model for Assisting Diagnosis in Computed Tomography30 Aug 2024 0 repositories listed
-
Retrieval-Augmented Natural Language Reasoning for Explainable Visual Question Answering30 Aug 2024 0 repositories listed
-
Speaker Tagging Correction With Non-Autoregressive Language Models30 Aug 2024 0 repositories listed
-
A Gradient Analysis Framework for Rewarding Good and Penalizing Bad Examples in Language Models29 Aug 2024 0 repositories listed
-
Benchmarking Japanese Speech Recognition on ASR-LLM Setups with Multi-Pass Augmented Generative Error Correction29 Aug 2024 0 repositories listed
-
ChatSUMO: Large Language Model for Automating Traffic Scenario Generation in Simulation of Urban MObility29 Aug 2024 0 repositories listed
-
Critic-CoT: Boosting the reasoning abilities of large language model via Chain-of-thoughts Critic29 Aug 2024 0 repositories listed
-
DriveGenVLM: Real-world Video Generation for Vision Language Model based Autonomous Driving29 Aug 2024 0 repositories listed
-
Logic Contrastive Reasoning with Lightweight Large Language Model for Math Word Problems29 Aug 2024 0 repositories listed
-
Rethinking Sparse Lexical Representations for Image Retrieval in the Age of Rising Multi-Modal Large Language Models29 Aug 2024 0 repositories listed
-
SynDL: A Large-Scale Synthetic Test Collection for Passage Retrieval29 Aug 2024 0 repositories listed
-
VLM-KD: Knowledge Distillation from VLM for Long-Tail Visual Recognition29 Aug 2024 0 repositories listed
-
Boosting Lossless Speculative Decoding via Feature Sampling and Partial Alignment Distillation28 Aug 2024 0 repositories listed
-
Drop the beat! Freestyler for Accompaniment Conditioned Rapping Voice Generation28 Aug 2024 0 repositories listed
-
Kangaroo: A Powerful Video-Language Model Supporting Long-context Video Input28 Aug 2024 0 repositories listed
-
LM-PUB-QUIZ: A Comprehensive Framework for Zero-Shot Evaluation of Relational Knowledge in Language Models28 Aug 2024 0 repositories listed
-
28 Aug 2024 0 repositories listed
-
AAVENUE: Detecting LLM Biases on NLU Tasks in AAVE via a Novel Benchmark27 Aug 2024 0 repositories listed
-
Awes, Laws, and Flaws From Today's LLM Research27 Aug 2024 0 repositories listed
-
BaichuanSEED: Sharing the Potential of ExtensivE Data Collection and Deduplication by Introducing a Competitive Large Language Model Baseline27 Aug 2024 0 repositories listed
-
Parameter-Efficient Quantized Mixture-of-Experts Meets Vision-Language Instruction Tuning for Semiconductor Electron Micrograph Analysis27 Aug 2024 0 repositories listed
-
Snap and Diagnose: An Advanced Multimodal Retrieval System for Identifying Plant Diseases in the Wild27 Aug 2024 0 repositories listed
-
Unifying Multitrack Music Arrangement via Reconstruction Fine-Tuning and Efficient Tokenization27 Aug 2024 0 repositories listed
-
Investigating Language-Specific Calibration For Pruning Multilingual Large Language Models26 Aug 2024 0 repositories listed
-
Large Language Model for Patent Concept Generation26 Aug 2024 0 repositories listed
-
Predictability and Causality in Spanish and English Natural Language Generation26 Aug 2024 0 repositories listed
-
Reprogramming Foundational Large Language Models(LLMs) for Enterprise Adoption for Spatio-Temporal Forecasting Applications: Unveiling a New Era in Copilot-Guided Cross-Modal Time Series Representation Learning26 Aug 2024 0 repositories listed
-
Enhancing SQL Query Generation with Neurosymbolic Reasoning25 Aug 2024 0 repositories listed
-
LLMs are Superior Feedback Providers: Bootstrapping Reasoning for Lie Detection with Self-Generated Feedback25 Aug 2024 0 repositories listed
-
SPICED: Syntactical Bug and Trojan Pattern Identification in A/MS Circuits using LLM-Enhanced Detection25 Aug 2024 0 repositories listed
-
StockTime: A Time Series Specialized Large Language Model Architecture for Stock Price Prediction25 Aug 2024 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.