Browse State-of-the-Art › Language Modelling › Papers, page 100
Language Modelling
Papers archive 2025-07-28
archive papers tagged: 17,610 · with a code link: 7,012 · where Syntology ran a sample: 2,428 (2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,428 of 17,610 tagged: 2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument)
Page 100 of 177: papers 9,901 to 10,000 of 17,610, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
F-HOI: Toward Fine-grained Semantic-Aligned 3D Human-Object Interactions17 Jul 2024 0 repositories listed
-
LLM Inference Serving: Survey of Recent Advances and Opportunities17 Jul 2024 0 repositories listed
-
Krutrim LLM: A Novel Tokenization Strategy for Multilingual Indic Languages with Petabyte-Scale Data Processing17 Jul 2024 0 repositories listed
-
R+X: Retrieval and Execution from Everyday Human Videos17 Jul 2024 0 repositories listed
-
Retrieval-Enhanced Machine Learning: Synthesis and Opportunities17 Jul 2024 0 repositories listed
-
SENTAUR: Security EnhaNced Trojan Assessment Using LLMs Against Undesirable Revisions17 Jul 2024 0 repositories listed
-
VisionTrap: Vision-Augmented Trajectory Prediction Guided by Textual Descriptions17 Jul 2024 0 repositories listed
-
A Language Modeling Approach to Diacritic-Free Hebrew TTS16 Jul 2024 0 repositories listed
-
A Pilot Study of GSLM-based Simulation of Foreign Accentuation Only Using Native Speech Corpora16 Jul 2024 0 repositories listed
-
BadRobot: Jailbreaking Embodied LLMs in the Physical World16 Jul 2024 0 repositories listed
-
CCoE: A Compact LLM with Collaboration of Experts16 Jul 2024 0 repositories listed
-
GPT Assisted Annotation of Rhetorical and Linguistic Features for Interpretable Propaganda Technique Detection in News Text16 Jul 2024 0 repositories listed
-
The Foundations of Tokenization: Statistical and Computational Concerns16 Jul 2024 0 repositories listed
-
BiasScanner: Automatic Detection and Classification of News Bias to Strengthen Democracy15 Jul 2024 0 repositories listed
-
Building Intelligence Identification System via Large Language Model Watermarking: A Survey and Beyond15 Jul 2024 0 repositories listed
-
Enhancing Medication Recommendation with LLM Text Representation15 Jul 2024 0 repositories listed
-
Fine-Tuning and Prompt Optimization: Two Great Steps that Work Better Together15 Jul 2024 0 repositories listed
-
GraphEval: A Knowledge-Graph Based LLM Hallucination Evaluation Framework15 Jul 2024 0 repositories listed
-
How and where does CLIP process negation?15 Jul 2024 0 repositories listed
-
Large Language Model-based FMRI Encoding of Language Functions for Subjects with Neurocognitive Disorder15 Jul 2024 0 repositories listed
-
MMM: Multilingual Mutual Reinforcement Effect Mix Datasets & Test with Open-domain Information Extraction Large Language Models15 Jul 2024 0 repositories listed
-
On Machine Learning Approaches for Protein-Ligand Binding Affinity Prediction15 Jul 2024 0 repositories listed
-
Enhancing Emotion Prediction in News Headlines: Insights from ChatGPT and Seq2Seq Models for Free-Text Generation14 Jul 2024 0 repositories listed
-
Key-Point-Driven Mathematical Reasoning Distillation of Large Language Model14 Jul 2024 0 repositories listed
-
LAB-Bench: Measuring Capabilities of Language Models for Biology Research14 Jul 2024 0 repositories listed
-
Lean-STaR: Learning to Interleave Thinking and Proving14 Jul 2024 0 repositories listed
-
14 Jul 2024 0 repositories listed Syntology 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Multi-Granularity Semantic Revision for Large Language Model Distillation14 Jul 2024 0 repositories listed
-
Rapid Biomedical Research Classification: The Pandemic PACT Advanced Categorisation Engine14 Jul 2024 0 repositories listed
-
Data Imputation using Large Language Model to Accelerate Recommendation System14 Jul 2024 0 repositories listed
-
Bilingual Adaptation of Monolingual Foundation Models13 Jul 2024 0 repositories listed
-
Explanation is All You Need in Distillation: Mitigating Bias and Shortcut Learning13 Jul 2024 0 repositories listed
-
ICCV23 Visual-Dialog Emotion Explanation Challenge: SEU_309 Team Technical Report13 Jul 2024 0 repositories listed
-
PFPs: Prompt-guided Flexible Pathological Segmentation for Diverse Potential Outcomes Using Large Vision and Language Models13 Jul 2024 0 repositories listed
-
A Neural Matrix Decomposition Recommender System Model based on the Multimodal Large Language Model12 Jul 2024 0 repositories listed
-
AI-Powered Immersive Assistance for Interactive Task Execution in Industrial Environments12 Jul 2024 0 repositories listed
-
Bridging Dictionary: AI-Generated Dictionary of Partisan Language Use12 Jul 2024 0 repositories listed
-
Deep Bag-of-Words Model: An Efficient and Interpretable Relevance Architecture for Chinese E-Commerce12 Jul 2024 0 repositories listed
-
Does Incomplete Syntax Influence Korean Language Model? Focusing on Word Order and Case Markers12 Jul 2024 0 repositories listed
-
Domain-Hierarchy Adaptation via Chain of Iterative Reasoning for Few-shot Hierarchical Text Classification12 Jul 2024 0 repositories listed
-
FairyLandAI: Personalized Fairy Tales utilizing ChatGPT and DALLE-312 Jul 2024 0 repositories listed
-
GPC: Generative and General Pathology Image Classifier12 Jul 2024 0 repositories listed
-
How Chinese are Chinese Language Models? The Puzzling Lack of Language Policy in China's LLMs12 Jul 2024 0 repositories listed
-
Optimized Multi-Token Joint Decoding with Auxiliary Model for LLM Inference12 Jul 2024 0 repositories listed
-
TensorTEE: Unifying Heterogeneous TEE Granularity for Efficient Secure Collaborative Tensor Computing12 Jul 2024 0 repositories listed
-
The Sociolinguistic Foundations of Language Modeling12 Jul 2024 0 repositories listed
-
Automata-based constraints for language model decoding11 Jul 2024 0 repositories listed
-
Autoregressive Speech Synthesis without Vector Quantization11 Jul 2024 0 repositories listed
-
Continually Learn to Map Visual Concepts to Large Language Models in Resource-constrained Environments11 Jul 2024 0 repositories listed
-
Evaluating Nuanced Bias in Large Language Model Free Response Answers11 Jul 2024 0 repositories listed
-
Fault Diagnosis in Power Grids with Large Language Model11 Jul 2024 0 repositories listed
-
GeNet: A Multimodal LLM-Based Co-Pilot for Network Topology and Configuration11 Jul 2024 0 repositories listed
-
MAGNET: Improving the Multilingual Fairness of Language Models with Adaptive Gradient-Based Tokenization11 Jul 2024 0 repositories listed
-
Rule-Based, Neural and LLM Back-Translation: Comparative Insights from a Variant of Ladin11 Jul 2024 0 repositories listed
-
Tamil Language Computing: the Present and the Future11 Jul 2024 0 repositories listed
-
Turn-Level Empathy Prediction Using Psychological Indicators11 Jul 2024 0 repositories listed
-
Unveiling Disparities in Maternity Care: A Topic Modelling Approach to Analysing Maternity Incident Investigation Reports11 Jul 2024 0 repositories listed
-
Automated Question Generation on Tabular Data for Conversational Data Exploration10 Jul 2024 0 repositories listed
-
Deconstructing What Makes a Good Optimizer for Language Models10 Jul 2024 0 repositories listed
-
Evaluating Voice Command Pipelines for Drone Control: From STT and LLM to Direct Classification and Siamese Networks10 Jul 2024 0 repositories listed
-
HiLight: Technical Report on the Motern AI Video Language Model10 Jul 2024 0 repositories listed
-
Knowledge Overshadowing Causes Amalgamated Hallucination in Large Language Models10 Jul 2024 0 repositories listed
-
Large Language Model-Augmented Auto-Delineation of Treatment Target Volume in Radiation Therapy10 Jul 2024 0 repositories listed
-
LokiLM: Technical Report10 Jul 2024 0 repositories listed
-
Malicious Path Manipulations via Exploitation of Representation Vulnerabilities of Vision-Language Navigation Systems10 Jul 2024 0 repositories listed
-
A Guide To Effectively Leveraging LLMs for Low-Resource Text Summarization: Data Augmentation and Semi-supervised Approaches10 Jul 2024 0 repositories listed
-
Token-Mol 1.0: Tokenized drug design with large language model10 Jul 2024 0 repositories listed
-
A Simple Architecture for Enterprise Large Language Model Applications based on Role based security and Clearance Levels using Retrieval-Augmented Generation or Mixture of Experts9 Jul 2024 0 repositories listed
-
AutoTask: Task Aware Multi-Faceted Single Model for Multi-Task Ads Relevance9 Jul 2024 0 repositories listed
-
Enhancing Low-Resource NMT with a Multilingual Encoder and Knowledge Distillation: A Case Study9 Jul 2024 0 repositories listed
-
LETS-C: Leveraging Text Embedding for Time Series Classification9 Jul 2024 0 repositories listed
-
Pseudo-perplexity in One Fell Swoop for Protein Fitness Estimation9 Jul 2024 0 repositories listed
-
SoftDedup: an Efficient Data Reweighting Method for Speeding Up Language Model Pre-training9 Jul 2024 0 repositories listed
-
Using Pretrained Large Language Model with Prompt Engineering to Answer Biomedical Questions9 Jul 2024 0 repositories listed
-
VQA-Diff: Exploiting VQA and Diffusion for Zero-Shot Image-to-3D Vehicle Asset Generation in Autonomous Driving9 Jul 2024 0 repositories listed
-
Artificial Intuition: Efficient Classification of Scientific Abstracts8 Jul 2024 0 repositories listed
-
B'MOJO: Hybrid State Space Realizations of Foundation Models with Eidetic and Fading Memory8 Jul 2024 0 repositories listed
-
CrowdMoGen: Zero-Shot Text-Driven Collective Motion Generation8 Jul 2024 0 repositories listed
-
Enhancing Language Model Rationality with Bi-Directional Deliberation Reasoning8 Jul 2024 0 repositories listed
-
GenArtist: Multimodal LLM as an Agent for Unified Image Generation and Editing8 Jul 2024 0 repositories listed
-
GenFollower: Enhancing Car-Following Prediction with Large Language Models8 Jul 2024 0 repositories listed
-
HyCIR: Boosting Zero-Shot Composed Image Retrieval with Synthetic Labels8 Jul 2024 0 repositories listed
-
Igea: a Decoder-Only Language Model for Biomedical Text Generation in Italian8 Jul 2024 0 repositories listed
-
Large Language Models for Judicial Entity Extraction: A Comparative Study8 Jul 2024 0 repositories listed
-
E²CFD: Towards Effective and Efficient Cost Function Design for Safe Reinforcement Learning via Large Language Model8 Jul 2024 0 repositories listed
-
On the Power of Convolution Augmented Transformer8 Jul 2024 0 repositories listed
-
OneDiff: A Generalist Model for Image Difference Captioning8 Jul 2024 0 repositories listed
-
Using Grammar Masking to Ensure Syntactic Validity in LLM-based Modeling Tasks8 Jul 2024 0 repositories listed
-
Variational Best-of-N Alignment8 Jul 2024 0 repositories listed
-
Biomedical Nested NER with Large Language Model and UMLS Heuristics7 Jul 2024 0 repositories listed
-
Large Language Model as an Assignment Evaluator: Insights, Feedback, and Challenges in a 1000+ Student Course7 Jul 2024 0 repositories listed
-
AI Safety in Generative AI Large Language Models: A Survey6 Jul 2024 0 repositories listed
-
Leveraging Task-Specific Knowledge from LLM for Semi-Supervised 3D Medical Image Segmentation6 Jul 2024 0 repositories listed
-
Aligning Model Evaluations with Human Preferences: Mitigating Token Count Bias in Language Model Assessments5 Jul 2024 0 repositories listed
-
Efficient Controlled Language Generation with Low-Rank Autoregressive Reward Models5 Jul 2024 0 repositories listed
-
Dude: Dual Distribution-Aware Context Prompt Learning For Large Vision-Language Model5 Jul 2024 0 repositories listed
-
EventChat: Implementation and user-centric evaluation of a large language model-driven conversational recommender system for exploring leisure events in an SME context5 Jul 2024 0 repositories listed
-
MobileFlow: A Multimodal LLM For Mobile GUI Agent5 Jul 2024 0 repositories listed
-
Romanization Encoding For Multilingual ASR5 Jul 2024 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.