Browse State-of-the-Art › Language Modelling › Papers, page 93
Language Modelling
Papers archive 2025-07-28
archive papers tagged: 17,610 · with a code link: 7,012 · where Syntology ran a sample: 2,428 (2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,428 of 17,610 tagged: 2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument)
Page 93 of 177: papers 9,201 to 9,300 of 17,610, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Disease Entity Recognition and Normalization is Improved with Large Language Model Derived Synthetic Normalized Mentions10 Oct 2024 0 repositories listed
-
Efficient Reinforcement Learning with Large Language Model Priors10 Oct 2024 0 repositories listed
-
Semantic Self-Consistency: Enhancing Language Model Reasoning via Semantic Weighting10 Oct 2024 0 repositories listed
-
Evolutionary Contrastive Distillation for Language Model Alignment10 Oct 2024 0 repositories listed
-
Language model developers should report train-test overlap10 Oct 2024 0 repositories listed
-
LecPrompt: A Prompt-based Approach for Logical Error Correction with CodeBERT10 Oct 2024 0 repositories listed
-
Mechanistic Permutability: Match Features Across Layers10 Oct 2024 0 repositories listed
-
PLaMo-100B: A Ground-Up Language Model Designed for Japanese Proficiency10 Oct 2024 0 repositories listed
-
Promptly Yours? A Human Subject Study on Prompt Inference in AI-Generated Art10 Oct 2024 0 repositories listed
-
Sample then Identify: A General Framework for Risk Control and Assessment in Multimodal Large Language Models10 Oct 2024 0 repositories listed
-
The Large Language Model GreekLegalRoBERTa10 Oct 2024 0 repositories listed
-
Uncovering Overfitting in Large Language Model Editing10 Oct 2024 0 repositories listed
-
β-calibration of Language Model Confidence Scores for Generative QA9 Oct 2024 0 repositories listed
-
Boosting Few-Shot Detection with Large Language Models and Layout-to-Image Synthesis9 Oct 2024 0 repositories listed
-
Exploring Efficient Foundational Multi-modal Models for Video Summarization9 Oct 2024 0 repositories listed
-
Exploring Prompt Engineering: A Systematic Review with SWOT Analysis9 Oct 2024 0 repositories listed
-
FltLM: An Intergrated Long-Context Large Language Model for Effective Context Filtering and Understanding9 Oct 2024 0 repositories listed
-
Generating long-horizon stock "buy" signals with a neural language model9 Oct 2024 0 repositories listed
-
Let's Ask GNN: Empowering Large Language Model for Graph In-Context Learning9 Oct 2024 0 repositories listed
-
Large Language Model Compression with Neural Architecture Search9 Oct 2024 0 repositories listed
-
Multi-Task Program Error Repair and Explanatory Diagnosis9 Oct 2024 0 repositories listed
-
Personal Intelligence System UniLM: Hybrid On-Device Small Language Model and Server-Based Large Language Model for Malay Nusantara9 Oct 2024 0 repositories listed
-
QuAILoRA: Quantization-Aware Initialization for LoRA9 Oct 2024 0 repositories listed
-
Recent advancements in LLM Red-Teaming: Techniques, Defenses, and Ethical Considerations9 Oct 2024 0 repositories listed
-
Reproducing and Extending Experiments in Behavioral Strategy with Large Language Models9 Oct 2024 0 repositories listed
-
Stuffed Mamba: State Collapse and State Capacity of RNN-Based Long-Context Modeling9 Oct 2024 0 repositories listed
-
TinyClick: Single-Turn Agent for Empowering GUI Automation9 Oct 2024 0 repositories listed
-
Towards Universality: Studying Mechanistic Similarity Across Language Model Architectures9 Oct 2024 0 repositories listed
-
Uncovering Factor Level Preferences to Improve Human-Model Alignment9 Oct 2024 0 repositories listed
-
Accelerated Preference Optimization for Large Language Model Alignment8 Oct 2024 0 repositories listed
-
Application of NotebookLM, a Large Language Model with Retrieval-Augmented Generation, for Lung Cancer Staging8 Oct 2024 0 repositories listed
-
Applying Refusal-Vector Ablation to Llama 3.1 70B Agents8 Oct 2024 0 repositories listed
-
ClaimBrush: A Novel Framework for Automated Patent Claim Refinement Based on Large Language Models8 Oct 2024 0 repositories listed
-
DecorateLM: Data Engineering through Corpus Rating, Tagging, and Editing with Language Models8 Oct 2024 0 repositories listed
-
FG-PRM: Fine-grained Hallucination Detection and Mitigation in Language Model Mathematical Reasoning8 Oct 2024 0 repositories listed
-
Jet Expansions of Residual Computation8 Oct 2024 0 repositories listed
-
Multi-Session Client-Centered Treatment Outcome Evaluation in Psychotherapy8 Oct 2024 0 repositories listed
-
ParallelSpec: Parallel Drafter for Efficient Speculative Decoding8 Oct 2024 0 repositories listed
-
Retrieving, Rethinking and Revising: The Chain-of-Verification Can Improve Retrieval Augmented Generation8 Oct 2024 0 repositories listed
-
TapType: Ten-finger text entry on everyday surfaces via Bayesian inference8 Oct 2024 0 repositories listed
-
TeaserGen: Generating Teasers for Long Documentaries8 Oct 2024 0 repositories listed
-
Activation Scaling for Steering and Interpreting Language Models7 Oct 2024 0 repositories listed
-
Chain and Causal Attention for Efficient Entity Tracking7 Oct 2024 0 repositories listed
-
Filtering Discomforting Recommendations with Large Language Models7 Oct 2024 0 repositories listed
-
DEPT: Decoupled Embeddings for Pre-training Language Models7 Oct 2024 0 repositories listed
-
Driving with Regulation: Interpretable Decision-Making for Autonomous Vehicles with Retrieval-Augmented Reasoning via LLM7 Oct 2024 0 repositories listed
-
Falcon Mamba: The First Competitive Attention-free 7B Language Model7 Oct 2024 0 repositories listed
-
In-the-loop Hyper-Parameter Optimization for LLM-Based Automated Design of Heuristics7 Oct 2024 0 repositories listed
-
Learning How Hard to Think: Input-Adaptive Allocation of LM Computation7 Oct 2024 0 repositories listed
-
Leverage Knowledge Graph and Large Language Model for Law Article Recommendation: A Case Study of Chinese Criminal Law7 Oct 2024 0 repositories listed
-
LPZero: Language Model Zero-cost Proxy Search from Zero7 Oct 2024 0 repositories listed
-
Reasoning Paths Optimization: Learning to Reason and Explore From Diverse Paths7 Oct 2024 0 repositories listed
-
RespLLM: Unifying Audio and Text with Multimodal LLMs for Generalized Respiratory Health Prediction7 Oct 2024 0 repositories listed
-
SFTMix: Elevating Language Model Instruction Tuning with Mixup Recipe7 Oct 2024 0 repositories listed
-
Towards the generation of hierarchical attack models from cybersecurity vulnerabilities using language models7 Oct 2024 0 repositories listed
-
Transformers learn variable-order Markov chains in-context7 Oct 2024 0 repositories listed
-
VLM2Vec: Training Vision-Language Models for Massive Multimodal Embedding Tasks7 Oct 2024 0 repositories listed
-
Wireless-Friendly Window Position Optimization for RIS-Aided Outdoor-to-Indoor Networks based on Multi-Modal Large Language Model7 Oct 2024 0 repositories listed
-
HALL-E: Hierarchical Neural Codec Language Model for Minute-Long Zero-Shot Text-to-Speech Synthesis6 Oct 2024 0 repositories listed
-
OD-Stega: LLM-Based Near-Imperceptible Steganography via Optimized Distributions6 Oct 2024 0 repositories listed
-
ReTok: Replacing Tokenizer to Enhance Representation Efficiency in Large Language Model6 Oct 2024 0 repositories listed
-
5 Oct 2024 0 repositories listed Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Assessing the Performance of Human-Capable LLMs -- Are LLMs Coming for Your Job?5 Oct 2024 0 repositories listed
-
Language Model-Driven Data Pruning Enables Efficient Active Learning5 Oct 2024 0 repositories listed
-
RoQLlama: A Lightweight Romanian Adapted Language Model5 Oct 2024 0 repositories listed
-
Toxic Subword Pruning for Dialogue Response Generation on Large Language Models5 Oct 2024 0 repositories listed
-
A Large Language Model-based Framework for Semi-Structured Tender Document Retrieval-Augmented Generation4 Oct 2024 0 repositories listed
-
An X-Ray Is Worth 15 Features: Sparse Autoencoders for Interpretable Radiology Report Generation4 Oct 2024 0 repositories listed
-
Audio-Agent: Leveraging LLMs For Audio Generation, Editing and Composition4 Oct 2024 0 repositories listed
-
Autoregressive Large Language Models are Computationally Universal4 Oct 2024 0 repositories listed
-
Function-Guided Conditional Generation Using Protein Language Models with Adapters4 Oct 2024 0 repositories listed
-
Cross-lingual Transfer for Automatic Question Generation by Learning Interrogative Structures in Target Languages4 Oct 2024 0 repositories listed
-
Enhancing Short-Text Topic Modeling with LLM-Driven Context Expansion and Prefix-Tuned VAEs4 Oct 2024 0 repositories listed
-
Image First or Text First? Optimising the Sequencing of Modalities in Large Language Model Prompting and Reasoning Tasks4 Oct 2024 0 repositories listed
-
Understanding Large Language Models in Your Pockets: Performance Study on COTS Mobile Devices4 Oct 2024 0 repositories listed
-
Large Language Models can be Strong Self-Detoxifiers4 Oct 2024 0 repositories listed
-
No Need to Talk: Asynchronous Mixture of Language Models4 Oct 2024 0 repositories listed
-
Parallel Corpus Augmentation using Masked Language Models4 Oct 2024 0 repositories listed
-
Permissive Information-Flow Analysis for Large Language Models4 Oct 2024 0 repositories listed
-
Scaling Parameter-Constrained Language Models with Quality Data4 Oct 2024 0 repositories listed
-
Surgical, Cheap, and Flexible: Mitigating False Refusal in Language Models via Single Vector Ablation4 Oct 2024 0 repositories listed
-
Textless Streaming Speech-to-Speech Translation using Semantic Speech Tokens4 Oct 2024 0 repositories listed
-
UNComp: Uncertainty-Aware Long-Context Compressor for Efficient Large Language Model Inference4 Oct 2024 0 repositories listed
-
Using Prompts to Guide Large Language Models in Imitating a Real Person's Language Style4 Oct 2024 0 repositories listed
-
Better Call SAUL: Fluent and Consistent Language Model Editing with Generation Regularization3 Oct 2024 0 repositories listed
-
BrainTransformers: SNN-LLM3 Oct 2024 0 repositories listed
-
CaLMFlow: Volterra Flow Matching using Causal Language Models3 Oct 2024 0 repositories listed
-
CodePMP: Scalable Preference Model Pretraining for Large Language Model Reasoning3 Oct 2024 0 repositories listed
-
Computational Modeling of Artistic Inspiration: A Framework for Predicting Aesthetic Preferences in Lyrical Lines Using Linguistic and Stylistic Features3 Oct 2024 0 repositories listed
-
Cut the Crap: An Economical Communication Pipeline for LLM-based Multi-Agent Systems3 Oct 2024 0 repositories listed
-
Determine-Then-Ensemble: Necessity of Top-k Union for Large Language Model Ensembling3 Oct 2024 0 repositories listed
-
Geometry is All You Need: A Unified Taxonomy of Matrix and Tensor Factorization for Compression of Generative Language Models3 Oct 2024 0 repositories listed
-
Grounding Large Language Models In Embodied Environment With Imperfect World Models3 Oct 2024 0 repositories listed
-
Large Language Model Aided Multi-objective Evolutionary Algorithm: a Low-cost Adaptive Approach3 Oct 2024 0 repositories listed
-
Large Language Model for Multi-Domain Translation: Benchmarking and Domain CoT Fine-tuning3 Oct 2024 0 repositories listed
-
Learning the Latent Rules of a Game from Data: A Chess Story3 Oct 2024 0 repositories listed
-
LLMCO2: Advancing Accurate Carbon Footprint Prediction for LLM Inferences3 Oct 2024 0 repositories listed
-
LoGra-Med: Long Context Multi-Graph Alignment for Medical Vision-Language Model3 Oct 2024 0 repositories listed
-
MedVisionLlama: Leveraging Pre-Trained Large Language Model Layers to Enhance Medical Image Segmentation3 Oct 2024 0 repositories listed
-
Morphological evaluation of subwords vocabulary used by BETO language model3 Oct 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.