Browse State-of-the-Art › Language Modelling › Papers, page 128
Language Modelling
Papers archive 2025-07-28
archive papers tagged: 17,610 · with a code link: 7,012 · where Syntology ran a sample: 2,428 (2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,428 of 17,610 tagged: 2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument)
Page 128 of 177: papers 12,701 to 12,800 of 17,610, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Exposing Attention Glitches with Flip-Flop Language Modeling1 Jun 2023 0 repositories listed
-
Graph-Level Embedding for Time-Evolving Graphs1 Jun 2023 0 repositories listed
-
How Generative Spoken Language Modeling Encodes Noisy Speech: Investigation from Phonetics to Syntactics1 Jun 2023 0 repositories listed
-
Interpretable Math Word Problem Solution Generation Via Step-by-step Planning1 Jun 2023 0 repositories listed
-
Understanding Augmentation-based Self-Supervised Representation Learning via RKHS Approximation and Regression1 Jun 2023 0 repositories listed
-
Adverbs, Surprisingly31 May 2023 0 repositories listed
-
RealignDiff: Boosting Text-to-Image Diffusion Model with Coarse-to-fine Semantic Re-alignment31 May 2023 0 repositories listed
-
BotArtist: Generic approach for bot detection in Twitter via semi-automatic machine learning pipeline31 May 2023 0 repositories listed
-
Catalysis distillation neural network for the few shot open catalyst challenge31 May 2023 0 repositories listed
-
Human or Not? A Gamified Approach to the Turing Test31 May 2023 0 repositories listed
-
Domain Specialization as the Key to Make Large Language Models Disruptive: A Comprehensive Survey30 May 2023 0 repositories listed
-
GPT4GEO: How a Language Model Sees the World's Geography30 May 2023 0 repositories listed
-
KEYword based Sampling (KEYS) for Large Language Models30 May 2023 0 repositories listed
-
30 May 2023 0 repositories listed
-
PanoGen: Text-Conditioned Panoramic Environment Generation for Vision-and-Language Navigation30 May 2023 0 repositories listed
-
Universality and Limitations of Prompt Tuning30 May 2023 0 repositories listed
-
Coeditor: Leveraging Contextual Changes for Multi-round Code Auto-editing29 May 2023 0 repositories listed
-
GripRank: Bridging the Gap between Retrieval and Generation via the Generative Knowledge Improved Passage Ranking29 May 2023 0 repositories listed
-
LaFTer: Label-Free Tuning of Zero-shot Classifier using Language and Unlabeled Image Collections29 May 2023 0 repositories listed
-
Information Association for Language Model Updating by Mitigating LM-Logical Discrepancy29 May 2023 0 repositories listed
-
Short Answer Grading Using One-shot Prompting and Text Similarity Scoring Model29 May 2023 0 repositories listed
-
Image Captioning with Multi-Context Synthetic Data29 May 2023 0 repositories listed
-
Writing user personas with Large Language Models: Testing phase 6 of a Thematic Analysis of semi-structured interviews29 May 2023 0 repositories listed
-
A Quantitative Review on Language Model Efficiency Research28 May 2023 0 repositories listed
-
Feature-Learning Networks Are Consistent Across Widths At Realistic Scales28 May 2023 0 repositories listed
-
Semantic Segmentation with Bidirectional Language Models Improves Long-form ASR28 May 2023 0 repositories listed
-
Augmenting Large Language Model Translators via Translation Memories27 May 2023 0 repositories listed
-
CIF-PT: Bridging Speech and Text Representations for Spoken Language Understanding via Continuous Integrate-and-Fire Pre-Training27 May 2023 0 repositories listed
-
CONA: A novel CONtext-Aware instruction paradigm for communication using large language model26 May 2023 0 repositories listed
-
Distinguishing Human Generated Text From ChatGPT Generated Text Using Machine Learning26 May 2023 0 repositories listed
-
Emergent Agentic Transformer from Chain of Hindsight Experience26 May 2023 0 repositories listed
-
External Language Model Integration for Factorized Neural Transducers26 May 2023 0 repositories listed
-
From Dogwhistles to Bullhorns: Unveiling Coded Rhetoric with Language Models26 May 2023 0 repositories listed
-
Green Runner: A tool for efficient model selection from model repositories26 May 2023 0 repositories listed
-
Improving accuracy of GPT-3/4 results on biomedical data using a retrieval-augmented language model26 May 2023 0 repositories listed
-
Large language models improve Alzheimer's disease diagnosis using multi-modality data26 May 2023 0 repositories listed
-
Slide, Constrain, Parse, Repeat: Synchronous SlidingWindows for Document AMR Parsing26 May 2023 0 repositories listed
-
SQL-PaLM: Improved Large Language Model Adaptation for Text-to-SQL (extended)26 May 2023 0 repositories listed
-
BookGPT: A General Framework for Book Recommendation Empowered by Large Language Model25 May 2023 0 repositories listed
-
Improving Scheduled Sampling for Neural Transducer-based ASR25 May 2023 0 repositories listed
-
VioLA: Unified Codec Language Models for Speech Recognition, Synthesis, and Translation25 May 2023 0 repositories listed
-
24 May 2023 0 repositories listed
-
Estimating class separability of text embeddings with persistent homology24 May 2023 0 repositories listed
-
A Monte Carlo Language Model Pipeline for Zero-Shot Sociopolitical Event Extraction24 May 2023 0 repositories listed
-
Chain-of-Questions Training with Latent Answers for Robust Multistep Question Answering24 May 2023 0 repositories listed
-
Drafting Event Schemas using Language Models24 May 2023 0 repositories listed
-
Dynamic Masking Rate Schedules for MLM Pretraining24 May 2023 0 repositories listed
-
Eliciting the Translation Ability of Large Language Models via Multilingual Finetuning with Translation Instructions24 May 2023 0 repositories listed
-
EmbodiedGPT: Vision-Language Pre-Training via Embodied Chain of Thought24 May 2023 0 repositories listed
-
Emergent inabilities? Inverse scaling over the course of pretraining24 May 2023 0 repositories listed
-
Just CHOP: Embarrassingly Simple LLM Compression24 May 2023 0 repositories listed
-
Large Language Models are Few-Shot Health Learners24 May 2023 0 repositories listed
-
Lexinvariant Language Models24 May 2023 0 repositories listed
-
Spoken Question Answering and Speech Continuation Using Spectrogram-Powered LLM24 May 2023 0 repositories listed
-
Getting MoRE out of Mixture of Language Model Reasoning Experts24 May 2023 0 repositories listed
-
Neural Summarization of Electronic Health Records24 May 2023 0 repositories listed
-
PURR: Efficiently Editing Language Model Hallucinations by Denoising Language Model Corruptions24 May 2023 0 repositories listed
-
Structural Ambiguity and its Disambiguation in Language Model Based Parsers: the Case of Dutch Clause Relativization24 May 2023 0 repositories listed
-
Alt-Text with Context: Improving Accessibility for Images on Twitter24 May 2023 0 repositories listed
-
Towards Adaptive Prefix Tuning for Parameter-Efficient Language Model Fine-tuning24 May 2023 0 repositories listed
-
Acquiring Frame Element Knowledge with Deep Metric Learning for Semantic Frame Induction23 May 2023 0 repositories listed
-
Cascaded Beam Search: Plug-and-Play Terminology-Forcing For Neural Machine Translation23 May 2023 0 repositories listed
-
Do All Languages Cost the Same? Tokenization in the Era of Commercial Language Models23 May 2023 0 repositories listed
-
Dr.ICL: Demonstration-Retrieved In-context Learning23 May 2023 0 repositories listed
-
Enhancing Black-Box Few-Shot Text Classification with Prompt-Based Data Augmentation23 May 2023 0 repositories listed
-
From Characters to Words: Hierarchical Pre-trained Language Model for Open-vocabulary Language Understanding23 May 2023 0 repositories listed
-
GenSpectrum Chat: Data Exploration in Public Health Using Large Language Models23 May 2023 0 repositories listed
-
Graph Meets LLM: A Novel Approach to Collaborative Filtering for Robust Conversational Understanding23 May 2023 0 repositories listed
-
Language Model Self-improvement by Reinforcement Learning Contemplation23 May 2023 0 repositories listed
-
Latent Positional Information is in the Self-Attention Variance of Transformer Language Models Without Positional Embeddings23 May 2023 0 repositories listed
-
Memory-Efficient Fine-Tuning of Compressed Large Language Models via sub-4-bit Integer Quantization23 May 2023 0 repositories listed
-
Query Rewriting for Retrieval-Augmented Large Language Models23 May 2023 0 repositories listed
-
R2H: Building Multimodal Navigation Helpers that Respond to Help Requests23 May 2023 0 repositories listed
-
Regex-augmented Domain Transfer Topic Classification based on a Pre-trained Language Model: An application in Financial Domain23 May 2023 0 repositories listed
-
Robust Prompt Optimization for Large Language Models Against Distribution Shifts23 May 2023 0 repositories listed
-
Towards A Unified View of Sparse Feed-Forward Network in Pretraining Large Language Model23 May 2023 0 repositories listed
-
Can LLMs facilitate interpretation of pre-trained language models?22 May 2023 0 repositories listed
-
Enhance Reasoning Ability of Visual-Language Models via Large Language Models22 May 2023 0 repositories listed
-
Evaluating Pragmatic Abilities of Image Captioners on A3DS22 May 2023 0 repositories listed
-
Extrapolating Multilingual Understanding Models as Multilingual Generators22 May 2023 0 repositories listed
-
GPT-SW3: An Autoregressive Language Model for the Nordic Languages22 May 2023 0 repositories listed
-
Distilling Robustness into Natural Language Inference Models with Domain-Targeted Augmentation22 May 2023 0 repositories listed
-
Explaining Emergent In-Context Learning as Kernel Regression22 May 2023 0 repositories listed
-
Learning Easily Updated General Purpose Text Representations with Adaptable Task-Specific Prefixes22 May 2023 0 repositories listed
-
LMGQS: A Large-scale Dataset for Query-focused Summarization22 May 2023 0 repositories listed
-
Observations on LLMs for Telecom Domain: Capabilities and Limitations22 May 2023 0 repositories listed
-
Text-based Person Search without Parallel Image-Text Data22 May 2023 0 repositories listed
-
The Influence of ChatGPT on Artificial Intelligence Related Crypto Assets: Evidence from a Synthetic Control Analysis22 May 2023 0 repositories listed
-
A Pilot Study on Dialogue-Level Dependency Parsing for Chinese21 May 2023 0 repositories listed
-
Augmenting Autotelic Agents with Large Language Models21 May 2023 0 repositories listed
-
Direct Fact Retrieval from Knowledge Graphs without Entity Linking21 May 2023 0 repositories listed
-
DPIC: Decoupling Prompt and Intrinsic Characteristics for LLM Generated Text Detection21 May 2023 0 repositories listed
-
HIINT: Historical, Intra- and Inter- personal Dynamics Modeling with Cross-person Memory Transformer21 May 2023 0 repositories listed
-
Infor-Coef: Information Bottleneck-based Dynamic Token Downsampling for Compact and Efficient language model21 May 2023 0 repositories listed
-
21 May 2023 0 repositories listed
-
OntoType: Ontology-Guided and Pre-Trained Language Model Assisted Fine-Grained Entity Typing21 May 2023 0 repositories listed
-
Description-Based Text Similarity21 May 2023 0 repositories listed
-
SLaDe: A Portable Small Language Model Decompiler for Optimized Assembly21 May 2023 0 repositories listed
-
A Sequence-to-Sequence Approach for Arabic Pronoun Resolution19 May 2023 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.