Browse State-of-the-Art › Language Modeling › Papers, page 82
Language Modeling
Papers archive 2025-07-28
archive papers tagged: 14,182 · with a code link: 5,620 · where Syntology ran a sample: 1,894 (1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,894 of 14,182 tagged: 1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument)
Page 82 of 142: papers 8,101 to 8,200 of 14,182, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
SpeechPrompt: Prompting Speech Language Models for Speech Processing Tasks23 Aug 2024 0 repositories listed
-
Balancing Act: Prioritization Strategies for LLM-Designed Restless Bandit Rewards22 Aug 2024 0 repositories listed
-
Can You Trust Your Metric? Automatic Concatenation-Based Tests for Metric Validity22 Aug 2024 0 repositories listed
-
Enhancing Multi-hop Reasoning through Knowledge Erasure in Large Language Model Editing22 Aug 2024 0 repositories listed
-
Implicit Sentiment Analysis Based on Chain of Thought Prompting22 Aug 2024 0 repositories listed
-
22 Aug 2024 0 repositories listed
-
Multi-tool Integration Application for Math Reasoning Using Large Language Model22 Aug 2024 0 repositories listed
-
TRRG: Towards Truthful Radiology Report Generation With Cross-modal Disease Clue Enhanced Large Language Model22 Aug 2024 0 repositories listed
-
Vintern-1B: An Efficient Multimodal Large Language Model for Vietnamese22 Aug 2024 0 repositories listed
-
Automating Thought of Search: A Journey Towards Soundness and Completeness21 Aug 2024 0 repositories listed
-
EE-MLLM: A Data-Efficient and Compute-Efficient Multimodal Large Language Model21 Aug 2024 0 repositories listed
-
Estimating Contribution Quality in Online Deliberations Using a Large Language Model21 Aug 2024 0 repositories listed
-
GeoReasoner: Reasoning On Geospatially Grounded Context For Natural Language Understanding21 Aug 2024 0 repositories listed
-
Improving Speech Recognition Error Prediction for Modern and Off-the-shelf Speech Recognizers21 Aug 2024 0 repositories listed
-
LARR: Large Language Model Aided Real-time Scene Recommendation with Semantic Understanding21 Aug 2024 0 repositories listed
-
WeQA: A Benchmark for Retrieval Augmented Generation in Wind Energy Domain21 Aug 2024 0 repositories listed
-
21 Aug 2024 0 repositories listed
-
Swarm Intelligence in Geo-Localization: A Multi-Agent Large Vision-Language Model Collaborative Framework21 Aug 2024 0 repositories listed
-
Video Emotion Open-vocabulary Recognition Based on Multimodal Large Language Model21 Aug 2024 0 repositories listed
-
What are the limits of cross-lingual dense passage retrieval for low-resource languages?21 Aug 2024 0 repositories listed
-
Analysis of Plan-based Retrieval for Grounded Text Generation20 Aug 2024 0 repositories listed
-
Data Augmentation Integrating Dialogue Flow and Style to Adapt Spoken Dialogue Systems to Low-Resource User Groups20 Aug 2024 0 repositories listed
-
Fine-Tuning a Local LLaMA-3 Large Language Model for Automated Privacy-Preserving Physician Letter Generation in Radiation Oncology20 Aug 2024 0 repositories listed
-
Hide Your Malicious Goal Into Benign Narratives: Jailbreak Large Language Models through Carrier Articles20 Aug 2024 0 repositories listed
-
HMoE: Heterogeneous Mixture of Experts for Language Modeling20 Aug 2024 0 repositories listed
-
Large Language Model Driven Recommendation20 Aug 2024 0 repositories listed
-
Minor SFT loss for LLM fine-tune to increase performance and reduce model deviation20 Aug 2024 0 repositories listed
-
Predicting Rewards Alongside Tokens: Non-disruptive Parameter Insertion for Efficient Inference Intervention in Large Language Model20 Aug 2024 0 repositories listed
-
Unconditional Truthfulness: Learning Conditional Dependency for Uncertainty Quantification of Large Language Models20 Aug 2024 0 repositories listed
-
Beyond Relevant Documents: A Knowledge-Intensive Approach for Query-Focused Summarization using Large Language Models19 Aug 2024 0 repositories listed
-
Cross-composition Feature Disentanglement for Compositional Zero-shot Learning19 Aug 2024 0 repositories listed
-
Development of an AI Anti-Bullying System Using Large Language Model Key Topic Detection19 Aug 2024 0 repositories listed
-
Geometry Informed Tokenization of Molecules for Language Model Generation19 Aug 2024 0 repositories listed
-
MAPLE: Enhancing Review Generation with Multi-Aspect Prompt LEarning in Explainable Recommendation19 Aug 2024 0 repositories listed
-
MePT: Multi-Representation Guided Prompt Tuning for Vision-Language Model19 Aug 2024 0 repositories listed
-
Minor DPO reject penalty to increase training robustness19 Aug 2024 0 repositories listed
-
MoDeGPT: Modular Decomposition for Large Language Model Compression19 Aug 2024 0 repositories listed
-
MSDiagnosis: A Benchmark for Evaluating Large Language Models in Multi-Step Clinical Diagnosis19 Aug 2024 0 repositories listed
-
SSDTrain: An Activation Offloading Framework to SSDs for Faster Large Language Model Training19 Aug 2024 0 repositories listed
-
Crossing New Frontiers: Knowledge-Augmented Large Language Model Prompting for Zero-Shot Text-Based De Novo Molecule Design18 Aug 2024 0 repositories listed
-
FASST: Fast LLM-based Simultaneous Speech Translation18 Aug 2024 0 repositories listed
-
Grammatical Error Feedback: An Implicit Evaluation Approach18 Aug 2024 0 repositories listed
-
Moonshine: Distilling Game Content Generators into Steerable Generative Models18 Aug 2024 0 repositories listed
-
REFINE-LM: Mitigating Language Model Stereotypes via Reinforcement Learning18 Aug 2024 0 repositories listed
-
AI Managed Emergency Documentation with a Pretrained Model17 Aug 2024 0 repositories listed
-
Architectural Foundations for the Large Language Model Infrastructures17 Aug 2024 0 repositories listed
-
Chinese Metaphor Recognition Using a Multi-stage Prompting Large Language Model17 Aug 2024 0 repositories listed
-
Sentiment analysis of preservice teachers' reflections using a large language model17 Aug 2024 0 repositories listed
-
A Hassle-free Algorithm for Private Learning in Practice: Don't Use Tree Aggregation, Use BLTs16 Aug 2024 0 repositories listed
-
CIKMar: A Dual-Encoder Approach to Prompt-Based Reranking in Educational Dialogue Systems16 Aug 2024 0 repositories listed
-
Collaborative Cross-modal Fusion with Large Language Model for Recommendation16 Aug 2024 0 repositories listed
-
DePrompt: Desensitization and Evaluation of Personal Identifiable Information in Large Language Model Prompts16 Aug 2024 0 repositories listed
-
LLM-PCGC: Large Language Model-based Point Cloud Geometry Compression16 Aug 2024 0 repositories listed
-
Research on Personalized Compression Algorithm for Pre-trained Models Based on Homomorphic Entropy Increase16 Aug 2024 0 repositories listed
-
Risks and NLP Design: A Case Study on Procedural Document QA16 Aug 2024 0 repositories listed
-
Autonomous Behavior Planning For Humanoid Loco-manipulation Through Grounded Language Model15 Aug 2024 0 repositories listed
-
DaRec: A Disentangled Alignment Framework for Large Language Model and Recommender System15 Aug 2024 0 repositories listed
-
Enhancing Large Language Model-based Speech Recognition by Contextualization for Rare and Ambiguous Words15 Aug 2024 0 repositories listed
-
General-purpose Clothes Manipulation with Semantic Keypoints15 Aug 2024 0 repositories listed
-
LLM4DSR: Leveraing Large Language Model for Denoising Sequential Recommendation15 Aug 2024 0 repositories listed
-
P/D-Serve: Serving Disaggregated Large Language Model at Scale15 Aug 2024 0 repositories listed
-
Toward a Dialogue System Using a Large Language Model to Recognize User Emotions with a Camera15 Aug 2024 0 repositories listed
-
When Raw Data Prevails: Are Large Language Model Embeddings Effective in Numerical Data Representation for Medical Machine Learning Applications?15 Aug 2024 0 repositories listed
-
Cropper: Vision-Language Model for Image Cropping through In-Context Learning14 Aug 2024 0 repositories listed
-
Development of a Large Language Model-based Multi-Agent Clinical Decision Support System for Korean Triage and Acuity Scale (KTAS)-Based Triage and Treatment Planning in Emergency Departments14 Aug 2024 0 repositories listed
-
Do GPT Language Models Suffer From Split Personality Disorder? The Advent Of Substrate-Free Psychometrics14 Aug 2024 0 repositories listed
-
Bridging Information Asymmetry in Text-video Retrieval: A Data-centric Approach14 Aug 2024 0 repositories listed
-
Kraken: Inherently Parallel Transformers For Efficient Multi-Device Inference14 Aug 2024 0 repositories listed
-
Abstract Operations Research Modeling Using Natural Language Inputs14 Aug 2024 0 repositories listed
-
ONSEP: A Novel Online Neural-Symbolic Framework for Event Prediction Based on Large Language Model14 Aug 2024 0 repositories listed
-
Training Overhead Ratio: A Practical Reliability Metric for Large Language Model Training Systems14 Aug 2024 0 repositories listed
-
13 Aug 2024 0 repositories listed
-
Diversity Empowers Intelligence: Integrating Expertise of Software Engineering Agents13 Aug 2024 0 repositories listed
-
13 Aug 2024 0 repositories listed Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
MGH Radiology Llama: A Llama 3 70B Model for Radiology13 Aug 2024 0 repositories listed
-
Response Wide Shut: Surprising Observations in Basic Vision Language Model Capabilities13 Aug 2024 0 repositories listed
-
SceneGPT: A Language Model for 3D Scene Understanding13 Aug 2024 0 repositories listed
-
SparkRA: A Retrieval-Augmented Knowledge Service System Based on Spark Large Language Model13 Aug 2024 0 repositories listed
-
Style-Talker: Finetuning Audio Language Model and Style-Based Text-to-Speech Model for Fast Spoken Dialogue Generation13 Aug 2024 0 repositories listed
-
Vision Language Model for Interpretable and Fine-grained Detection of Safety Compliance in Diverse Workplaces13 Aug 2024 0 repositories listed
-
Space-LLaVA: a Vision-Language Model Adapted to Extraterrestrial Applications12 Aug 2024 0 repositories listed
-
AGE: Amharic, Ge’ez and English Parallel Dataset12 Aug 2024 0 repositories listed
-
Building Decision Making Models Through Language Model Regime12 Aug 2024 0 repositories listed
-
Creating Arabic LLM Prompts at Scale12 Aug 2024 0 repositories listed
-
Global-to-Local Support Spectrums for Language Model Explainability12 Aug 2024 0 repositories listed
-
LipidBERT: A Lipid Language Model Pre-trained on METiS de novo Lipid Library12 Aug 2024 0 repositories listed
-
Improving Whisper's Recognition Performance for Under-Represented Language Kazakh Leveraging Unpaired Speech and Text10 Aug 2024 0 repositories listed
-
Large Language Model-based Role-Playing for Personalized Medical Jargon Extraction10 Aug 2024 0 repositories listed
-
Path-LLM: A Shortest-Path-based LLM Learning for Unified Graph Representation10 Aug 2024 0 repositories listed
-
Speculative Diffusion Decoding: Accelerating Language Generation through Diffusion10 Aug 2024 0 repositories listed
-
ChatGPT Meets Iris Biometrics9 Aug 2024 0 repositories listed
-
ConfusedPilot: Confused Deputy Risks in RAG-based LLMs9 Aug 2024 0 repositories listed
-
Investigating a Benchmark for Training-set free Evaluation of Linguistic Capabilities in Machine Reading Comprehension9 Aug 2024 0 repositories listed
-
LLaMA based Punctuation Restoration With Forward Pass Only Decoding9 Aug 2024 0 repositories listed
-
MIDI-to-Tab: Guitar Tablature Inference via Masked Language Modeling9 Aug 2024 0 repositories listed
-
Report on the 1st Workshop on Large Language Model for Evaluation in Information Retrieval (LLM4Eval 2024) at SIGIR 20249 Aug 2024 0 repositories listed
-
Improving Mortality Prediction After Radiotherapy with Large Language Model Structuring of Large-Scale Unstructured Electronic Health Records9 Aug 2024 0 repositories listed
-
Analysis of Argument Structure Constructions in the Large Language Model BERT8 Aug 2024 0 repositories listed
-
Code-switching in text and speech reveals information-theoretic audience design8 Aug 2024 0 repositories listed
-
Dynamic Hypergraph-Enhanced Prediction of Sequential Medical Visits8 Aug 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.