Browse State-of-the-Art › Language Modelling › Papers, page 101
Language Modelling
Papers archive 2025-07-28
archive papers tagged: 17,610 · with a code link: 7,012 · where Syntology ran a sample: 2,428 (2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,428 of 17,610 tagged: 2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument)
Page 101 of 177: papers 10,001 to 10,100 of 17,610, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
5 Jul 2024 0 repositories listed
-
Semi-supervised Learning for Code-Switching ASR with Large Language Model Filter5 Jul 2024 0 repositories listed
-
Speculative Speech Recognition by Audio-Prefixed Low-Rank Adaptation of Language Models5 Jul 2024 0 repositories listed
-
Spontaneous Reward Hacking in Iterative Self-Refinement5 Jul 2024 0 repositories listed
-
Statistical investigations into the geometry and homology of random programs5 Jul 2024 0 repositories listed
-
Testing learning hypotheses using neural networks by manipulating learning data5 Jul 2024 0 repositories listed
-
Towards Context-aware Support for Color Vision Deficiency: An Approach Integrating LLM and AR5 Jul 2024 0 repositories listed
-
ConText at WASSA 2024 Empathy and Personality Shared Task: History-Dependent Embedding Utterance Representations for Empathy and Emotion Prediction in Conversations4 Jul 2024 0 repositories listed
-
Chain-of-Thought Augmentation with Logit Contrast for Enhanced Reasoning in Language Models4 Jul 2024 0 repositories listed
-
Diff-Restorer: Unleashing Visual Prompts for Diffusion-based Universal Image Restoration4 Jul 2024 0 repositories listed
-
Generative Technology for Human Emotion Recognition: A Scope Review4 Jul 2024 0 repositories listed
-
Improving Self Consistency in LLMs through Probabilistic Tokenization4 Jul 2024 0 repositories listed
-
MAPO: Boosting Large Language Model Performance with Model-Adaptive Prompt Optimization4 Jul 2024 0 repositories listed
-
MRIR: Integrating Multimodal Insights for Diffusion-based Realistic Image Restoration4 Jul 2024 0 repositories listed
-
Narrow Transformer: StarCoder-Based Java-LM For Desktop4 Jul 2024 0 repositories listed
-
On the Effectiveness of Acoustic BPE in Decoder-Only TTS4 Jul 2024 0 repositories listed
-
Unlocking the Potential of Model Merging for Low-Resource Languages4 Jul 2024 0 repositories listed
-
Align and Aggregate: Compositional Reasoning with Video Alignment and Answer Aggregation for Video Question-Answering3 Jul 2024 0 repositories listed
-
CogErgLLM: Exploring Large Language Model Systems Design Perspective Using Cognitive Ergonomics3 Jul 2024 0 repositories listed
-
Croppable Knowledge Graph Embedding3 Jul 2024 0 repositories listed
-
Efficient Training of Language Models with Compact and Consistent Next Token Distributions3 Jul 2024 0 repositories listed
-
Images Speak Louder than Words: Understanding and Mitigating Bias in Vision-Language Model from a Causal Mediation Perspective3 Jul 2024 0 repositories listed
-
Large Language Model Agents for Improving Engagement with Behavior Change Interventions: Application to Digital Mindfulness3 Jul 2024 0 repositories listed
-
Learning to Reduce: Towards Improving Performance of Large Language Models on Structured Data3 Jul 2024 0 repositories listed
-
LLMcap: Large Language Model for Unsupervised PCAP Failure Detection3 Jul 2024 0 repositories listed
-
MLKD-BERT: Multi-level Knowledge Distillation for Pre-trained Language Models3 Jul 2024 0 repositories listed
-
Model-Enhanced LLM-Driven VUI Testing of VPA Apps3 Jul 2024 0 repositories listed
-
3 Jul 2024 0 repositories listed Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Raw Text is All you Need: Knowledge-intensive Multi-turn Instruction Tuning for Large Language Model3 Jul 2024 0 repositories listed
-
RDBE: Reasoning Distillation-Based Evaluation Enhances Automatic Essay Scoring3 Jul 2024 0 repositories listed
-
SAFT: Towards Out-of-Distribution Generalization in Fine-Tuning3 Jul 2024 0 repositories listed
-
Supporting Cross-language Cross-project Bug Localization Using Pre-trained Language Models3 Jul 2024 0 repositories listed
-
Accompanied Singing Voice Synthesis with Fully Text-controlled Melody2 Jul 2024 0 repositories listed
-
An End-to-End Speech Summarization Using Large Language Model2 Jul 2024 0 repositories listed
-
Assessing the Effectiveness of GPT-4o in Climate Change Evidence Synthesis and Systematic Assessments: Preliminary Insights2 Jul 2024 0 repositories listed
-
Cost-Effective Proxy Reward Model Construction with On-Policy and Active Learning2 Jul 2024 0 repositories listed
-
D-Rax: Domain-specific Radiologic assistant leveraging multi-modal data and eXpert model predictions2 Jul 2024 0 repositories listed
-
LLMs Plagiarize: Ensuring Responsible Sourcing of Large Language Model Training Data Through Knowledge Graph Comparison2 Jul 2024 0 repositories listed
-
Fake News Detection and Manipulation Reasoning via Large Vision-Language Models2 Jul 2024 0 repositories listed
-
FineCLIPER: Multi-modal Fine-grained CLIP for Dynamic Facial Expression Recognition with AdaptERs2 Jul 2024 0 repositories listed
-
Investigating the Effects of Large-Scale Pseudo-Stereo Data and Different Speech Foundation Model on Dialogue Generative Spoken Language Model2 Jul 2024 0 repositories listed
-
Lightweight Large Language Model for Medication Enquiry: Med-Pal2 Jul 2024 0 repositories listed
-
Multi-Modal Video Dialog State Tracking in the Wild2 Jul 2024 0 repositories listed
-
PromptIntern: Saving Inference Costs by Internalizing Recurrent Prompt during Large Language Model Fine-tuning2 Jul 2024 0 repositories listed
-
SeqMate: A Novel Large Language Model Pipeline for Automating RNA Sequencing2 Jul 2024 0 repositories listed
-
Synthetic Multimodal Question Generation2 Jul 2024 0 repositories listed
-
Why do LLaVA Vision-Language Models Reply to Images in English?2 Jul 2024 0 repositories listed
-
1 Jul 2024 0 repositories listed
-
Bridging the Gap: Transfer Learning from English PLMs to Malaysian English1 Jul 2024 0 repositories listed
-
ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities1 Jul 2024 0 repositories listed
-
ESALE: Enhancing Code-Summary Alignment Learning for Source Code Summarization1 Jul 2024 0 repositories listed
-
First Place Solution of 2023 Global Artificial Intelligence Technology Innovation Competition Track 11 Jul 2024 0 repositories listed
-
FoldGPT: Simple and Effective Large Language Model Compression Scheme1 Jul 2024 0 repositories listed
-
Image-to-Text Logic Jailbreak: Your Imagination can Help You Do Anything1 Jul 2024 0 repositories listed
-
Large Language Model Enhanced Knowledge Representation Learning: A Survey1 Jul 2024 0 repositories listed
-
Needle in the Haystack for Memory Based Large Language Models1 Jul 2024 0 repositories listed
-
Optimization of Retrieval-Augmented Generation Context with Outlier Detection1 Jul 2024 0 repositories listed
-
Explaining Length Bias in LLM-Based Preference Evaluations1 Jul 2024 0 repositories listed
-
Memory³: Language Modeling with Explicit Memory1 Jul 2024 0 repositories listed
-
Tokenize the World into Object-level Knowledge to Address Long-tail Events in Autonomous Driving1 Jul 2024 0 repositories listed
-
Characterizing Stereotypical Bias from Privacy-preserving Pre-Training30 Jun 2024 0 repositories listed
-
Explaining Chest X-ray Pathology Models using Textual Concepts30 Jun 2024 0 repositories listed
-
Scaling Technology Acceptance Analysis with Large Language Model (LLM) Annotation Systems30 Jun 2024 0 repositories listed
-
A Study on Effect of Reference Knowledge Choice in Generating Technical Content Relevant to SAPPhIRE Model Using Large Language Model29 Jun 2024 0 repositories listed
-
Answering real-world clinical questions using large language model based systems29 Jun 2024 0 repositories listed
-
Financial Knowledge Large Language Model29 Jun 2024 0 repositories listed
-
It's Morphing Time: Unleashing the Potential of Multiple LLMs via Multi-objective Optimization29 Jun 2024 0 repositories listed
-
Open-Source Conversational AI with SpeechBrain 1.029 Jun 2024 0 repositories listed
-
Potential Renovation of Information Search Process with the Power of Large Language Model for Healthcare29 Jun 2024 0 repositories listed
-
Self-Translate-Train: Enhancing Cross-Lingual Transfer of Large Language Models via Inherent Capability29 Jun 2024 0 repositories listed
-
BESTOW: Efficient and Streamable Speech Language Model with the Best of Two Worlds in GPT and T528 Jun 2024 0 repositories listed
-
Can GPT-4 Help Detect Quit Vaping Intentions? An Exploration of Automatic Data Annotation Approach28 Jun 2024 0 repositories listed
-
Designing and Evaluating Multi-Chatbot Interface for Human-AI Communication: Preliminary Findings from a Persuasion Task28 Jun 2024 0 repositories listed
-
Investigating the Timescales of Language Processing with EEG and Language Models28 Jun 2024 0 repositories listed
-
Simulating Financial Market via Large Language Model based Agents28 Jun 2024 0 repositories listed
-
The Qiyas Benchmark: Measuring ChatGPT Mathematical and Language Understanding in Arabic28 Jun 2024 0 repositories listed
-
Adaptive Draft-Verification for Efficient Large Language Model Decoding27 Jun 2024 0 repositories listed
-
IndoToxic2024: A Demographically-Enriched Dataset of Hate Speech and Toxicity Types for Indonesian Language27 Jun 2024 0 repositories listed
-
LICO: Large Language Models for In-Context Molecular Optimization27 Jun 2024 0 repositories listed
-
LongLaMP: A Benchmark for Personalized Long-form Text Generation27 Jun 2024 0 repositories listed
-
LoPT: Low-Rank Prompt Tuning for Parameter Efficient Language Models27 Jun 2024 0 repositories listed
-
Meta Large Language Model Compiler: Foundation Models of Compiler Optimization27 Jun 2024 0 repositories listed
-
MissionGNN: Hierarchical Multimodal GNN-based Weakly Supervised Video Anomaly Recognition with Mission-Specific Knowledge Graph Generation27 Jun 2024 0 repositories listed
-
PathAlign: A vision-language model for whole slide images in histopathology27 Jun 2024 0 repositories listed
-
xTower: A Multilingual LLM for Explaining and Correcting Translation Errors27 Jun 2024 0 repositories listed
-
Zero-shot Composed Image Retrieval Considering Query-target Relationship Leveraging Masked Image-text Pairs27 Jun 2024 0 repositories listed
-
Explicit Diversity Conditions for Effective Question Answer Generation with Large Language Models26 Jun 2024 0 repositories listed
-
Llamipa: An Incremental Discourse Parser26 Jun 2024 0 repositories listed
-
MammothModa: Multi-Modal Large Language Model26 Jun 2024 0 repositories listed
-
MUMU: Bootstrapping Multimodal Image Generation from Text-to-Image Data26 Jun 2024 0 repositories listed
-
Octo-planner: On-device Language Model for Planner-Action Agents26 Jun 2024 0 repositories listed
-
PharmaGPT: Domain-Specific Large Language Models for Bio-Pharmaceutical and Chemistry26 Jun 2024 0 repositories listed
-
Towards Large Language Model Aided Program Refinement26 Jun 2024 0 repositories listed
-
A Comprehensive Solution to Connect Speech Encoder and Large Language Model for ASR25 Jun 2024 0 repositories listed
-
Accelerating Clinical Evidence Synthesis with Large Language Models25 Jun 2024 0 repositories listed
-
AG-LSEC: Audio Grounded Lexical Speaker Error Correction25 Jun 2024 0 repositories listed
-
Beyond Demographics: Aligning Role-playing LLM-based Agents Using Human Belief Networks25 Jun 2024 0 repositories listed
-
Discrete Diffusion Language Model for Long Text Summarization25 Jun 2024 0 repositories listed
-
Find Parent then Label Children: A Two-stage Taxonomy Completion Method with Pre-trained Language Model25 Jun 2024 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.