Browse State-of-the-Art › Language Modeling › Papers, page 69
Language Modeling
Papers archive 2025-07-28
archive papers tagged: 14,182 · with a code link: 5,620 · where Syntology ran a sample: 1,894 (1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,894 of 14,182 tagged: 1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument)
Page 69 of 142: papers 6,801 to 6,900 of 14,182, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Concept Navigation and Classification via Open-Source Large Language Model Processing7 Feb 2025 0 repositories listed
-
DCFormer: Efficient 3D Vision-Language Modeling with Decomposed Convolutions7 Feb 2025 0 repositories listed
-
Learning the Language of NVMe Streams for Ransomware Detection7 Feb 2025 0 repositories listed
-
RAG-Verus: Repository-Level Program Verification with LLMs using Retrieval Augmented Generation7 Feb 2025 0 repositories listed
-
Refining Integration-by-Parts Reduction of Feynman Integrals with Machine Learning7 Feb 2025 0 repositories listed
-
Adaptive Semantic Prompt Caching with VectorQ6 Feb 2025 0 repositories listed
-
Contextual Gradient Flow Modeling for Large Language Model Generalization in Multi-Scale Feature Spaces6 Feb 2025 0 repositories listed
-
DiTAR: Diffusion Transformer Autoregressive Modeling for Speech Generation6 Feb 2025 0 repositories listed
-
FairT2I: Mitigating Social Bias in Text-to-Image Generation via Large Language Model-Assisted Detection and Attribute Rebalancing6 Feb 2025 0 repositories listed
-
RWKV-UI: UI Understanding with Enhanced Perception and Reasoning6 Feb 2025 0 repositories listed
-
Verifiable Format Control for Large Language Model Generations6 Feb 2025 0 repositories listed
-
Vision-Integrated LLMs for Autonomous Driving Assistance : Human Performance Comparison and Trust Evaluation6 Feb 2025 0 repositories listed
-
A Contemporary Survey of Large Language Model Assisted Program Analysis5 Feb 2025 0 repositories listed
-
Adapt-Pruner: Adaptive Structural Pruning for Efficient Small Language Model Training5 Feb 2025 0 repositories listed
-
Control Search Rankings, Control the World: What is a Good Search Engine?5 Feb 2025 0 repositories listed
-
Efficient Vision Language Model Fine-tuning for Text-based Person Anomaly Search5 Feb 2025 0 repositories listed
-
Entropy Adaptive Decoding: Dynamic Model Switching for Efficient Inference5 Feb 2025 0 repositories listed
-
Fine-grained Preference Optimization Improves Zero-shot Text-to-Speech5 Feb 2025 0 repositories listed
-
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling5 Feb 2025 0 repositories listed
-
Gold-medalist Performance in Solving Olympiad Geometry with AlphaGeometry25 Feb 2025 0 repositories listed
-
HACK: Homomorphic Acceleration via Compression of the Key-Value Cache for Disaggregated LLM Inference5 Feb 2025 0 repositories listed
-
Large Language Model Guided Self-Debugging Code Generation5 Feb 2025 0 repositories listed
-
Large Language Model as Universal Retriever in Industrial-Scale Recommender System5 Feb 2025 0 repositories listed
-
On Fairness of Unified Multimodal Large Language Model for Image Generation5 Feb 2025 0 repositories listed
-
Simplifying Formal Proof-Generating Models with ChatGPT and Basic Searching Techniques5 Feb 2025 0 repositories listed
-
Token Assorted: Mixing Latent and Text Tokens for Improved Language Model Reasoning5 Feb 2025 0 repositories listed
-
Analyzing Similarity Metrics for Data Selection for Language Model Pretraining4 Feb 2025 0 repositories listed
-
Automating Mathematical Proof Generation Using Large Language Model Agents and Knowledge Graphs4 Feb 2025 0 repositories listed
-
ComplexDec: A Domain-robust High-fidelity Neural Audio Codec with Complex Spectrum Modeling4 Feb 2025 0 repositories listed
-
EditIQ: Automated Cinematic Editing of Static Wide-Angle Videos via Dialogue Interpretation and Saliency Cues4 Feb 2025 0 repositories listed
-
FinBloom: Knowledge Grounding Large Language Model with Real-time Financial Data4 Feb 2025 0 repositories listed
-
Flatten Graphs as Sequences: Transformers are Scalable Graph Generators4 Feb 2025 0 repositories listed
-
JingFang: A Traditional Chinese Medicine Large Language Model of Expert-Level Medical Diagnosis and Syndrome Differentiation-Based Treatment4 Feb 2025 0 repositories listed
-
LLM-USO: Large Language Model-based Universal Sizing Optimizer4 Feb 2025 0 repositories listed
-
MPIC: Position-Independent Multimodal Context Caching System for Efficient MLLM Serving4 Feb 2025 0 repositories listed
-
Position: Stop Acting Like Language Model Agents Are Normal Agents4 Feb 2025 0 repositories listed
-
Prompt-based Depth Pruning of Large Language Models4 Feb 2025 0 repositories listed
-
Rethinking Homogeneity of Vision and Text Tokens in Large Vision-and-Language Models4 Feb 2025 0 repositories listed
-
SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model4 Feb 2025 0 repositories listed
-
Unlocking Efficient Large Inference Models: One-Bit Unrolling Tips the Scales4 Feb 2025 0 repositories listed
-
An Inquiry into Datacenter TCO for LLM Inference with FP83 Feb 2025 0 repositories listed
-
ConditionNET: Learning Preconditions and Effects for Execution Monitoring3 Feb 2025 0 repositories listed
-
Eliciting Language Model Behaviors with Investigator Agents3 Feb 2025 0 repositories listed
-
FALCON: Fine-grained Activation Manipulation by Contrastive Orthogonal Unalignment for Large Language Model3 Feb 2025 0 repositories listed
-
InfoBridge: Mutual Information estimation via Bridge Matching3 Feb 2025 0 repositories listed
-
Knowledge Synthesis of Photosynthesis Research Using a Large Language Model3 Feb 2025 0 repositories listed
-
Latent Lexical Projection in Large Language Models: A Novel Approach to Implicit Representation Refinement3 Feb 2025 0 repositories listed
-
Learning to Learn Weight Generation via Local Consistency Diffusion3 Feb 2025 0 repositories listed
-
Position: Towards a Responsible LLM-empowered Multi-Agent Systems3 Feb 2025 0 repositories listed
-
Scalable Language Models with Posterior Inference of Latent Thought Vectors3 Feb 2025 0 repositories listed
-
Scaling Embedding Layers in Language Models3 Feb 2025 0 repositories listed
-
Soup-of-Experts: Pretraining Specialist Models via Parameters Averaging3 Feb 2025 0 repositories listed
-
The Differences Between Direct Alignment Algorithms are a Blur3 Feb 2025 0 repositories listed
-
Agent-Based Uncertainty Awareness Improves Automated Radiology Report Labeling with an Open-Source Large Language Model2 Feb 2025 0 repositories listed
-
Decision-informed Neural Networks with Large Language Model Integration for Portfolio Optimization2 Feb 2025 0 repositories listed
-
Efficient Multi-Agent System Training with Data Influence-Oriented Tree Search2 Feb 2025 0 repositories listed
-
Language Models Use Trigonometry to Do Addition2 Feb 2025 0 repositories listed
-
LIBRA: Measuring Bias of Large Language Model from a Local Context2 Feb 2025 0 repositories listed
-
2 Feb 2025 0 repositories listed Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
A statistically consistent measure of Semantic Variability using Language Models1 Feb 2025 0 repositories listed
-
Doing More with Less -- Implementing Routing Strategies in Large Language Model-Based Systems: An Extended Survey1 Feb 2025 0 repositories listed
-
Enhancing Token Filtering Efficiency in Large Language Model Training with Collider1 Feb 2025 0 repositories listed
-
INSIGHT: Enhancing Autonomous Driving Safety through Vision-Language Models on Context-Aware Hazard Detection and Edge Case Evaluation1 Feb 2025 0 repositories listed
-
OrcaLoca: An LLM Agent Framework for Software Issue Localization1 Feb 2025 0 repositories listed
-
An Efficient Approach for Machine Translation on Low-resource Languages: A Case Study in Vietnamese-Chinese31 Jan 2025 0 repositories listed
-
Brain-inspired sparse training enables Transformers and LLMs to perform as fully connected31 Jan 2025 0 repositories listed
-
BRiTE: Bootstrapping Reinforced Thinking Process to Enhance Language Model Reasoning31 Jan 2025 0 repositories listed
-
Can AI Solve the Peer Review Crisis? A Large Scale Cross Model Experiment of LLMs' Performance and Biases in Evaluating over 1000 Economics Papers31 Jan 2025 0 repositories listed
-
Estimating the Probability of Sampling a Trained Neural Network at Random31 Jan 2025 0 repositories listed
-
Importing Phantoms: Measuring LLM Package Hallucination Vulnerabilities31 Jan 2025 0 repositories listed
-
Intrinsic Tensor Field Propagation in Large Language Models: A Novel Approach to Contextual Information Flow31 Jan 2025 0 repositories listed
-
Mobile Robot Navigation Using Hand-Drawn Maps: A Vision Language Model Approach31 Jan 2025 0 repositories listed
-
Offline Learning for Combinatorial Multi-armed Bandits31 Jan 2025 0 repositories listed
-
Resolving Editing-Unlearning Conflicts: A Knowledge Codebook Framework for Large Language Model Updating31 Jan 2025 0 repositories listed
-
Scaling Laws for Differentially Private Language Models31 Jan 2025 0 repositories listed
-
SELMA: A Speech-Enabled Language Model for Virtual Assistant Interactions31 Jan 2025 0 repositories listed
-
Structural Embedding Projection for Contextual Large Language Model Inference31 Jan 2025 0 repositories listed
-
Towards the Worst-case Robustness of Large Language Models31 Jan 2025 0 repositories listed
-
CALM: Unleashing the Cross-Lingual Self-Aligning Ability of Language Model Question Answering30 Jan 2025 0 repositories listed
-
CLEAR: Cue Learning using Evolution for Accurate Recognition Applied to Sustainability Data Extraction30 Jan 2025 0 repositories listed
-
Efficiency and Effectiveness of LLM-Based Summarization of Evidence in Crowdsourced Fact-Checking30 Jan 2025 0 repositories listed
-
Economic Rationality under Specialization: Evidence of Decision Bias in AI Agents30 Jan 2025 0 repositories listed
-
Enhancing Large Language Model Efficiencyvia Symbolic Compression: A Formal Approach Towards Interpretability30 Jan 2025 0 repositories listed
-
Exploring Audio Editing Features as User-Centric Privacy Defenses Against Large Language Model(LLM) Based Emotion Inference Attacks30 Jan 2025 0 repositories listed
-
Fine-tuning LLaMA 2 interference: a comparative study of language implementations for optimal efficiency30 Jan 2025 0 repositories listed
-
Investigating Tax Evasion Emergence Using Dual Large Language Model and Deep Reinforcement Learning Powered Agent-based Simulation30 Jan 2025 0 repositories listed
-
Loss Functions and Operators Generated by f-Divergences30 Jan 2025 0 repositories listed
-
Vision-Language Model Selection and Reuse for Downstream Adaptation30 Jan 2025 0 repositories listed
-
Token-Hungry, Yet Precise: DeepSeek R1 Highlights the Need for Multi-Step Reasoning Over Speed in MATH30 Jan 2025 0 repositories listed
-
BreezyVoice: Adapting TTS for Taiwanese Mandarin with Enhanced Polyphone Disambiguation -- Challenges and Insights29 Jan 2025 0 repositories listed
-
DINT Transformer29 Jan 2025 0 repositories listed
-
DReSS: Data-driven Regularized Structured Streamlining for Large Language Models29 Jan 2025 0 repositories listed
-
From tools to thieves: Measuring and understanding public perceptions of AI through crowdsourced metaphors29 Jan 2025 0 repositories listed
-
Large Language Models for Single-Step and Multi-Step Flight Trajectory Prediction29 Jan 2025 0 repositories listed
-
Learning Free Token Reduction for Multi-Modal Large Language Models29 Jan 2025 0 repositories listed
-
Planning with Vision-Language Models and a Use Case in Robot-Assisted Teaching29 Jan 2025 0 repositories listed
-
Prompt-oriented Output of Culture-Specific Items in Translated African Poetry by Large Language Model: An Initial Multi-layered Tabular Review29 Jan 2025 0 repositories listed
-
Query-Aware Learnable Graph Pooling Tokens as Prompt for Large Language Models29 Jan 2025 0 repositories listed
-
An LLM Benchmark for Addressee Recognition in Multi-modal Multi-party Dialogue28 Jan 2025 0 repositories listed
-
Implementation of a Generative AI Assistant in K-12 Education: The CyberScholar Initiative28 Jan 2025 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.