Browse State-of-the-Art › Language Modeling › Papers, page 85
Language Modeling
Papers archive 2025-07-28
archive papers tagged: 14,182 · with a code link: 5,620 · where Syntology ran a sample: 1,894 (1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,894 of 14,182 tagged: 1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument)
Page 85 of 142: papers 8,401 to 8,500 of 14,182, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
On the Power of Convolution Augmented Transformer8 Jul 2024 0 repositories listed
-
Using Grammar Masking to Ensure Syntactic Validity in LLM-based Modeling Tasks8 Jul 2024 0 repositories listed
-
Variational Best-of-N Alignment8 Jul 2024 0 repositories listed
-
Biomedical Nested NER with Large Language Model and UMLS Heuristics7 Jul 2024 0 repositories listed
-
Large Language Model as an Assignment Evaluator: Insights, Feedback, and Challenges in a 1000+ Student Course7 Jul 2024 0 repositories listed
-
Aligning Model Evaluations with Human Preferences: Mitigating Token Count Bias in Language Model Assessments5 Jul 2024 0 repositories listed
-
Efficient Controlled Language Generation with Low-Rank Autoregressive Reward Models5 Jul 2024 0 repositories listed
-
Dude: Dual Distribution-Aware Context Prompt Learning For Large Vision-Language Model5 Jul 2024 0 repositories listed
-
EventChat: Implementation and user-centric evaluation of a large language model-driven conversational recommender system for exploring leisure events in an SME context5 Jul 2024 0 repositories listed
-
Romanization Encoding For Multilingual ASR5 Jul 2024 0 repositories listed
-
5 Jul 2024 0 repositories listed
-
Semi-supervised Learning for Code-Switching ASR with Large Language Model Filter5 Jul 2024 0 repositories listed
-
Speculative Speech Recognition by Audio-Prefixed Low-Rank Adaptation of Language Models5 Jul 2024 0 repositories listed
-
Spontaneous Reward Hacking in Iterative Self-Refinement5 Jul 2024 0 repositories listed
-
Testing learning hypotheses using neural networks by manipulating learning data5 Jul 2024 0 repositories listed
-
Towards Context-aware Support for Color Vision Deficiency: An Approach Integrating LLM and AR5 Jul 2024 0 repositories listed
-
ConText at WASSA 2024 Empathy and Personality Shared Task: History-Dependent Embedding Utterance Representations for Empathy and Emotion Prediction in Conversations4 Jul 2024 0 repositories listed
-
Chain-of-Thought Augmentation with Logit Contrast for Enhanced Reasoning in Language Models4 Jul 2024 0 repositories listed
-
Improving Self Consistency in LLMs through Probabilistic Tokenization4 Jul 2024 0 repositories listed
-
MAPO: Boosting Large Language Model Performance with Model-Adaptive Prompt Optimization4 Jul 2024 0 repositories listed
-
Narrow Transformer: StarCoder-Based Java-LM For Desktop4 Jul 2024 0 repositories listed
-
On the Effectiveness of Acoustic BPE in Decoder-Only TTS4 Jul 2024 0 repositories listed
-
Unlocking the Potential of Model Merging for Low-Resource Languages4 Jul 2024 0 repositories listed
-
CogErgLLM: Exploring Large Language Model Systems Design Perspective Using Cognitive Ergonomics3 Jul 2024 0 repositories listed
-
Efficient Training of Language Models with Compact and Consistent Next Token Distributions3 Jul 2024 0 repositories listed
-
Images Speak Louder than Words: Understanding and Mitigating Bias in Vision-Language Model from a Causal Mediation Perspective3 Jul 2024 0 repositories listed
-
Large Language Model Agents for Improving Engagement with Behavior Change Interventions: Application to Digital Mindfulness3 Jul 2024 0 repositories listed
-
Learning to Reduce: Towards Improving Performance of Large Language Models on Structured Data3 Jul 2024 0 repositories listed
-
LLMcap: Large Language Model for Unsupervised PCAP Failure Detection3 Jul 2024 0 repositories listed
-
MLKD-BERT: Multi-level Knowledge Distillation for Pre-trained Language Models3 Jul 2024 0 repositories listed
-
3 Jul 2024 0 repositories listed Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Raw Text is All you Need: Knowledge-intensive Multi-turn Instruction Tuning for Large Language Model3 Jul 2024 0 repositories listed
-
RDBE: Reasoning Distillation-Based Evaluation Enhances Automatic Essay Scoring3 Jul 2024 0 repositories listed
-
SAFT: Towards Out-of-Distribution Generalization in Fine-Tuning3 Jul 2024 0 repositories listed
-
Supporting Cross-language Cross-project Bug Localization Using Pre-trained Language Models3 Jul 2024 0 repositories listed
-
Accompanied Singing Voice Synthesis with Fully Text-controlled Melody2 Jul 2024 0 repositories listed
-
An End-to-End Speech Summarization Using Large Language Model2 Jul 2024 0 repositories listed
-
Assessing the Effectiveness of GPT-4o in Climate Change Evidence Synthesis and Systematic Assessments: Preliminary Insights2 Jul 2024 0 repositories listed
-
LLMs Plagiarize: Ensuring Responsible Sourcing of Large Language Model Training Data Through Knowledge Graph Comparison2 Jul 2024 0 repositories listed
-
Fake News Detection and Manipulation Reasoning via Large Vision-Language Models2 Jul 2024 0 repositories listed
-
Investigating the Effects of Large-Scale Pseudo-Stereo Data and Different Speech Foundation Model on Dialogue Generative Spoken Language Model2 Jul 2024 0 repositories listed
-
Lightweight Large Language Model for Medication Enquiry: Med-Pal2 Jul 2024 0 repositories listed
-
Multi-Modal Video Dialog State Tracking in the Wild2 Jul 2024 0 repositories listed
-
PromptIntern: Saving Inference Costs by Internalizing Recurrent Prompt during Large Language Model Fine-tuning2 Jul 2024 0 repositories listed
-
SeqMate: A Novel Large Language Model Pipeline for Automating RNA Sequencing2 Jul 2024 0 repositories listed
-
Synthetic Multimodal Question Generation2 Jul 2024 0 repositories listed
-
Why do LLaVA Vision-Language Models Reply to Images in English?2 Jul 2024 0 repositories listed
-
1 Jul 2024 0 repositories listed
-
ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities1 Jul 2024 0 repositories listed
-
ESALE: Enhancing Code-Summary Alignment Learning for Source Code Summarization1 Jul 2024 0 repositories listed
-
First Place Solution of 2023 Global Artificial Intelligence Technology Innovation Competition Track 11 Jul 2024 0 repositories listed
-
FoldGPT: Simple and Effective Large Language Model Compression Scheme1 Jul 2024 0 repositories listed
-
Image-to-Text Logic Jailbreak: Your Imagination can Help You Do Anything1 Jul 2024 0 repositories listed
-
Large Language Model Enhanced Knowledge Representation Learning: A Survey1 Jul 2024 0 repositories listed
-
Needle in the Haystack for Memory Based Large Language Models1 Jul 2024 0 repositories listed
-
Optimization of Retrieval-Augmented Generation Context with Outlier Detection1 Jul 2024 0 repositories listed
-
Memory³: Language Modeling with Explicit Memory1 Jul 2024 0 repositories listed
-
Tokenize the World into Object-level Knowledge to Address Long-tail Events in Autonomous Driving1 Jul 2024 0 repositories listed
-
Characterizing Stereotypical Bias from Privacy-preserving Pre-Training30 Jun 2024 0 repositories listed
-
Explaining Chest X-ray Pathology Models using Textual Concepts30 Jun 2024 0 repositories listed
-
Scaling Technology Acceptance Analysis with Large Language Model (LLM) Annotation Systems30 Jun 2024 0 repositories listed
-
A Study on Effect of Reference Knowledge Choice in Generating Technical Content Relevant to SAPPhIRE Model Using Large Language Model29 Jun 2024 0 repositories listed
-
Answering real-world clinical questions using large language model based systems29 Jun 2024 0 repositories listed
-
Financial Knowledge Large Language Model29 Jun 2024 0 repositories listed
-
Open-Source Conversational AI with SpeechBrain 1.029 Jun 2024 0 repositories listed
-
Potential Renovation of Information Search Process with the Power of Large Language Model for Healthcare29 Jun 2024 0 repositories listed
-
BESTOW: Efficient and Streamable Speech Language Model with the Best of Two Worlds in GPT and T528 Jun 2024 0 repositories listed
-
Designing and Evaluating Multi-Chatbot Interface for Human-AI Communication: Preliminary Findings from a Persuasion Task28 Jun 2024 0 repositories listed
-
Investigating the Timescales of Language Processing with EEG and Language Models28 Jun 2024 0 repositories listed
-
Simulating Financial Market via Large Language Model based Agents28 Jun 2024 0 repositories listed
-
The Qiyas Benchmark: Measuring ChatGPT Mathematical and Language Understanding in Arabic28 Jun 2024 0 repositories listed
-
Adaptive Draft-Verification for Efficient Large Language Model Decoding27 Jun 2024 0 repositories listed
-
LICO: Large Language Models for In-Context Molecular Optimization27 Jun 2024 0 repositories listed
-
LoPT: Low-Rank Prompt Tuning for Parameter Efficient Language Models27 Jun 2024 0 repositories listed
-
Meta Large Language Model Compiler: Foundation Models of Compiler Optimization27 Jun 2024 0 repositories listed
-
PathAlign: A vision-language model for whole slide images in histopathology27 Jun 2024 0 repositories listed
-
xTower: A Multilingual LLM for Explaining and Correcting Translation Errors27 Jun 2024 0 repositories listed
-
Zero-shot Composed Image Retrieval Considering Query-target Relationship Leveraging Masked Image-text Pairs27 Jun 2024 0 repositories listed
-
Explicit Diversity Conditions for Effective Question Answer Generation with Large Language Models26 Jun 2024 0 repositories listed
-
Llamipa: An Incremental Discourse Parser26 Jun 2024 0 repositories listed
-
MammothModa: Multi-Modal Large Language Model26 Jun 2024 0 repositories listed
-
MUMU: Bootstrapping Multimodal Image Generation from Text-to-Image Data26 Jun 2024 0 repositories listed
-
Octo-planner: On-device Language Model for Planner-Action Agents26 Jun 2024 0 repositories listed
-
PharmaGPT: Domain-Specific Large Language Models for Bio-Pharmaceutical and Chemistry26 Jun 2024 0 repositories listed
-
Towards Large Language Model Aided Program Refinement26 Jun 2024 0 repositories listed
-
A Comprehensive Solution to Connect Speech Encoder and Large Language Model for ASR25 Jun 2024 0 repositories listed
-
AG-LSEC: Audio Grounded Lexical Speaker Error Correction25 Jun 2024 0 repositories listed
-
Beyond Demographics: Aligning Role-playing LLM-based Agents Using Human Belief Networks25 Jun 2024 0 repositories listed
-
Discrete Diffusion Language Model for Long Text Summarization25 Jun 2024 0 repositories listed
-
Find Parent then Label Children: A Two-stage Taxonomy Completion Method with Pre-trained Language Model25 Jun 2024 0 repositories listed
-
High Fidelity Text-to-Speech Via Discrete Tokens Using Token Transducer and Group Masked Language Model25 Jun 2024 0 repositories listed
-
Improving Robustness of LLM-based Speech Synthesis by Learning Monotonic Alignment25 Jun 2024 0 repositories listed
-
LABOR-LLM: Language-Based Occupational Representations with Large Language Models25 Jun 2024 0 repositories listed
-
MoE-CT: A Novel Approach For Large Language Models Training With Resistance To Catastrophic Forgetting25 Jun 2024 0 repositories listed
-
Semi-supervised classification of dental conditions in panoramic radiographs using large language model and instance segmentation: A real-world dataset evaluation25 Jun 2024 0 repositories listed
-
Understanding Language Model Circuits through Knowledge Editing25 Jun 2024 0 repositories listed
-
Is your benchmark truly adversarial? AdvScore: Evaluating Human-Grounded Adversarialness24 Jun 2024 0 repositories listed
-
AnnotatedTables: A Large Tabular Dataset with Language Model Annotations24 Jun 2024 0 repositories listed
-
Classification of Geological Borehole Descriptions Using a Domain Adapted Large Language Model24 Jun 2024 0 repositories listed
-
Context-augmented Retrieval: A Novel Framework for Fast Information Retrieval based Response Generation using Large Language Model24 Jun 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.