Browse State-of-the-Art › Language Modeling › Papers, page 78
Language Modeling
Papers archive 2025-07-28
archive papers tagged: 14,182 · with a code link: 5,620 · where Syntology ran a sample: 1,894 (1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,894 of 14,182 tagged: 1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument)
Page 78 of 142: papers 7,701 to 7,800 of 14,182, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
To Err is AI : A Case Study Informing LLM Flaw Reporting Practices15 Oct 2024 0 repositories listed
-
Tokenization and Morphology in Multilingual Language Models: A Comparative Analysis of mT5 and ByT515 Oct 2024 0 repositories listed
-
Towards More Effective Table-to-Text Generation: Assessing In-Context Learning and Self-Evaluation with Open-Source Models15 Oct 2024 0 repositories listed
-
Y-Mol: A Multiscale Biomedical Knowledge-Guided Large Language Model for Drug Development15 Oct 2024 0 repositories listed
-
A Multi-Task Text Classification Pipeline with Natural Language Explanations: A User-Centric Evaluation in Sentiment Analysis and Offensive Language Identification in Greek Tweets14 Oct 2024 0 repositories listed
-
Not All Options Are Created Equal: Textual Option Weighting for Token-Efficient LLM-Based Knowledge Tracing14 Oct 2024 0 repositories listed
-
ForgeryGPT: Multimodal Large Language Model For Explainable Image Forgery Detection and Localization14 Oct 2024 0 repositories listed
-
Large Language Model-Enhanced Reinforcement Learning for Generic Bus Holding Control Strategies14 Oct 2024 0 repositories listed
-
LG-CAV: Train Any Concept Activation Vector with Language Guidance14 Oct 2024 0 repositories listed
-
LOBG:Less Overfitting for Better Generalization in Vision-Language Model14 Oct 2024 0 repositories listed
-
Model-based Large Language Model Customization as Service14 Oct 2024 0 repositories listed
-
Recipe for Zero-shot POS Tagging: Is It Useful in Realistic Scenarios?14 Oct 2024 0 repositories listed
-
Skill Learning Using Process Mining for Large Language Model Plan Generation14 Oct 2024 0 repositories listed
-
Unified Representation of Genomic and Biomedical Concepts through Multi-Task, Multi-Source Contrastive Learning14 Oct 2024 0 repositories listed
-
Adaptive Reasoning and Acting in Medical Language Agents13 Oct 2024 0 repositories listed
-
Collu-Bench: A Benchmark for Predicting Language Model Hallucinations in Code13 Oct 2024 0 repositories listed
-
EchoPrime: A Multi-Video View-Informed Vision-Language Model for Comprehensive Echocardiography Interpretation13 Oct 2024 0 repositories listed
-
Conversational Code Generation: a Case Study of Designing a Dialogue System for Generating Driving Scenarios for Testing Autonomous Vehicles13 Oct 2024 0 repositories listed
-
Learning to Rank for Multiple Retrieval-Augmented Models through Iterative Utility Maximization13 Oct 2024 0 repositories listed
-
LoRE: Logit-Ranked Retriever Ensemble for Enhancing Open-Domain Question Answering13 Oct 2024 0 repositories listed
-
MoIN: Mixture of Introvert Experts to Upcycle an LLM13 Oct 2024 0 repositories listed
-
Impeding LLM-assisted Cheating in Introductory Programming Assignments via Adversarial Perturbation12 Oct 2024 0 repositories listed
-
ACER: Automatic Language Model Context Extension via Retrieval11 Oct 2024 0 repositories listed
-
Aerial Vision-and-Language Navigation via Semantic-Topo-Metric Representation Guided LLM Reasoning11 Oct 2024 0 repositories listed
-
Calibrated Cache Model for Few-Shot Vision-Language Model Adaptation11 Oct 2024 0 repositories listed
-
Efficiently Scanning and Resampling Spatio-Temporal Tasks with Irregular Observations11 Oct 2024 0 repositories listed
-
∀uto∃∨∧L: Autonomous Evaluation of LLMs for Truth Maintenance and Reasoning Tasks11 Oct 2024 0 repositories listed
-
Hypothesis-only Biases in Large Language Model-Elicited Natural Language Inference11 Oct 2024 0 repositories listed
-
Language-Model-Assisted Bi-Level Programming for Reward Learning from Internet Videos11 Oct 2024 0 repositories listed
-
Lifelong Event Detection via Optimal Transport11 Oct 2024 0 repositories listed
-
LLMD: A Large Language Model for Interpreting Longitudinal Medical Records11 Oct 2024 0 repositories listed
-
nach0-pc: Multi-task Language Model with Molecular Point Cloud Encoder11 Oct 2024 0 repositories listed
-
Preferential Normalizing Flows11 Oct 2024 0 repositories listed
-
SimpleStrat: Diversifying Language Model Generation with Stratification11 Oct 2024 0 repositories listed
-
Simultaneous Reward Distillation and Preference Learning: Get You a Language Model Who Can Do Both11 Oct 2024 0 repositories listed
-
Emergent social conventions and collective bias in LLM populations11 Oct 2024 0 repositories listed
-
The Same But Different: Structural Similarities and Differences in Multilingual Language Modeling11 Oct 2024 0 repositories listed
-
Towards Trustworthy Knowledge Graph Reasoning: An Uncertainty Aware Perspective11 Oct 2024 0 repositories listed
-
ViT3D Alignment of LLaMA3: 3D Medical Image Report Generation11 Oct 2024 0 repositories listed
-
VLM See, Robot Do: Human Demo Video to Robot Action Plan via Vision Language Model11 Oct 2024 0 repositories listed
-
A Framework for Collaborating a Large Language Model Tool in Brainstorming for Triggering Creative Thoughts10 Oct 2024 0 repositories listed
-
CrossQuant: A Post-Training Quantization Method with Smaller Quantization Kernel for Precise Large Language Model Compression10 Oct 2024 0 repositories listed
-
DICE: Discrete Inversion Enabling Controllable Editing for Multinomial Diffusion and Masked Generative Models10 Oct 2024 0 repositories listed
-
Disease Entity Recognition and Normalization is Improved with Large Language Model Derived Synthetic Normalized Mentions10 Oct 2024 0 repositories listed
-
Efficient Reinforcement Learning with Large Language Model Priors10 Oct 2024 0 repositories listed
-
Semantic Self-Consistency: Enhancing Language Model Reasoning via Semantic Weighting10 Oct 2024 0 repositories listed
-
Evolutionary Contrastive Distillation for Language Model Alignment10 Oct 2024 0 repositories listed
-
Language model developers should report train-test overlap10 Oct 2024 0 repositories listed
-
LecPrompt: A Prompt-based Approach for Logical Error Correction with CodeBERT10 Oct 2024 0 repositories listed
-
Mechanistic Permutability: Match Features Across Layers10 Oct 2024 0 repositories listed
-
PLaMo-100B: A Ground-Up Language Model Designed for Japanese Proficiency10 Oct 2024 0 repositories listed
-
Sample then Identify: A General Framework for Risk Control and Assessment in Multimodal Large Language Models10 Oct 2024 0 repositories listed
-
The Large Language Model GreekLegalRoBERTa10 Oct 2024 0 repositories listed
-
Uncovering Overfitting in Large Language Model Editing10 Oct 2024 0 repositories listed
-
β-calibration of Language Model Confidence Scores for Generative QA9 Oct 2024 0 repositories listed
-
Exploring Efficient Foundational Multi-modal Models for Video Summarization9 Oct 2024 0 repositories listed
-
Exploring Prompt Engineering: A Systematic Review with SWOT Analysis9 Oct 2024 0 repositories listed
-
FltLM: An Intergrated Long-Context Large Language Model for Effective Context Filtering and Understanding9 Oct 2024 0 repositories listed
-
Generating long-horizon stock "buy" signals with a neural language model9 Oct 2024 0 repositories listed
-
Let's Ask GNN: Empowering Large Language Model for Graph In-Context Learning9 Oct 2024 0 repositories listed
-
Large Language Model Compression with Neural Architecture Search9 Oct 2024 0 repositories listed
-
Multi-Task Program Error Repair and Explanatory Diagnosis9 Oct 2024 0 repositories listed
-
Personal Intelligence System UniLM: Hybrid On-Device Small Language Model and Server-Based Large Language Model for Malay Nusantara9 Oct 2024 0 repositories listed
-
QuAILoRA: Quantization-Aware Initialization for LoRA9 Oct 2024 0 repositories listed
-
Recent advancements in LLM Red-Teaming: Techniques, Defenses, and Ethical Considerations9 Oct 2024 0 repositories listed
-
Reproducing and Extending Experiments in Behavioral Strategy with Large Language Models9 Oct 2024 0 repositories listed
-
Stuffed Mamba: State Collapse and State Capacity of RNN-Based Long-Context Modeling9 Oct 2024 0 repositories listed
-
TinyClick: Single-Turn Agent for Empowering GUI Automation9 Oct 2024 0 repositories listed
-
Towards Universality: Studying Mechanistic Similarity Across Language Model Architectures9 Oct 2024 0 repositories listed
-
Accelerated Preference Optimization for Large Language Model Alignment8 Oct 2024 0 repositories listed
-
Application of NotebookLM, a Large Language Model with Retrieval-Augmented Generation, for Lung Cancer Staging8 Oct 2024 0 repositories listed
-
Applying Refusal-Vector Ablation to Llama 3.1 70B Agents8 Oct 2024 0 repositories listed
-
ClaimBrush: A Novel Framework for Automated Patent Claim Refinement Based on Large Language Models8 Oct 2024 0 repositories listed
-
DecorateLM: Data Engineering through Corpus Rating, Tagging, and Editing with Language Models8 Oct 2024 0 repositories listed
-
FG-PRM: Fine-grained Hallucination Detection and Mitigation in Language Model Mathematical Reasoning8 Oct 2024 0 repositories listed
-
Jet Expansions of Residual Computation8 Oct 2024 0 repositories listed
-
Multi-Session Client-Centered Treatment Outcome Evaluation in Psychotherapy8 Oct 2024 0 repositories listed
-
ParallelSpec: Parallel Drafter for Efficient Speculative Decoding8 Oct 2024 0 repositories listed
-
Retrieving, Rethinking and Revising: The Chain-of-Verification Can Improve Retrieval Augmented Generation8 Oct 2024 0 repositories listed
-
Activation Scaling for Steering and Interpreting Language Models7 Oct 2024 0 repositories listed
-
Chain and Causal Attention for Efficient Entity Tracking7 Oct 2024 0 repositories listed
-
Filtering Discomforting Recommendations with Large Language Models7 Oct 2024 0 repositories listed
-
DEPT: Decoupled Embeddings for Pre-training Language Models7 Oct 2024 0 repositories listed
-
Driving with Regulation: Interpretable Decision-Making for Autonomous Vehicles with Retrieval-Augmented Reasoning via LLM7 Oct 2024 0 repositories listed
-
Falcon Mamba: The First Competitive Attention-free 7B Language Model7 Oct 2024 0 repositories listed
-
Leverage Knowledge Graph and Large Language Model for Law Article Recommendation: A Case Study of Chinese Criminal Law7 Oct 2024 0 repositories listed
-
LPZero: Language Model Zero-cost Proxy Search from Zero7 Oct 2024 0 repositories listed
-
Reasoning Paths Optimization: Learning to Reason and Explore From Diverse Paths7 Oct 2024 0 repositories listed
-
RespLLM: Unifying Audio and Text with Multimodal LLMs for Generalized Respiratory Health Prediction7 Oct 2024 0 repositories listed
-
SFTMix: Elevating Language Model Instruction Tuning with Mixup Recipe7 Oct 2024 0 repositories listed
-
Towards the generation of hierarchical attack models from cybersecurity vulnerabilities using language models7 Oct 2024 0 repositories listed
-
Transformers learn variable-order Markov chains in-context7 Oct 2024 0 repositories listed
-
VLM2Vec: Training Vision-Language Models for Massive Multimodal Embedding Tasks7 Oct 2024 0 repositories listed
-
Wireless-Friendly Window Position Optimization for RIS-Aided Outdoor-to-Indoor Networks based on Multi-Modal Large Language Model7 Oct 2024 0 repositories listed
-
HALL-E: Hierarchical Neural Codec Language Model for Minute-Long Zero-Shot Text-to-Speech Synthesis6 Oct 2024 0 repositories listed
-
OD-Stega: LLM-Based Near-Imperceptible Steganography via Optimized Distributions6 Oct 2024 0 repositories listed
-
ReTok: Replacing Tokenizer to Enhance Representation Efficiency in Large Language Model6 Oct 2024 0 repositories listed
-
5 Oct 2024 0 repositories listed Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Assessing the Performance of Human-Capable LLMs -- Are LLMs Coming for Your Job?5 Oct 2024 0 repositories listed
-
Language Model-Driven Data Pruning Enables Efficient Active Learning5 Oct 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.