Browse State-of-the-Art › Information Retrieval › Papers, page 13
Information Retrieval
Papers archive 2025-07-28
archive papers tagged: 4,740 · with a code link: 1,188 · where Syntology ran a sample: 191 (151 with a run with no instrument failure, 40 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (191 of 4,740 tagged: 151 with a run with no instrument failure, 40 where every run was a failure of Syntology's instrument)
Page 13 of 48: papers 1,201 to 1,300 of 4,740, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
InsertRank: LLMs can reason over BM25 scores to Improve Listwise Reranking17 Jun 2025 0 repositories listed
-
Hierarchical Multi-Positive Contrastive Learning for Patent Image Retrieval16 Jun 2025 0 repositories listed
-
xbench: Tracking Agents Productivity Scaling with Profession-Aligned Real-World Evaluations16 Jun 2025 0 repositories listed
-
eLog analysis for accelerators: status and future outlook15 Jun 2025 0 repositories listed
-
CMI-Bench: A Comprehensive Benchmark for Evaluating Music Instruction Following14 Jun 2025 0 repositories listed
-
Private Aggregation for Byzantine-Resilient Heterogeneous Federated Learning11 Jun 2025 0 repositories listed
-
ScholarSearch: Benchmarking Scholar Searching Ability of LLMs11 Jun 2025 0 repositories listed
-
Evaluation empirique de la sécurisation et de l'alignement de ChatGPT et Gemini: analyse comparative des vulnérabilités par expérimentations de jailbreaks10 Jun 2025 0 repositories listed
-
Multimodal Representation Alignment for Cross-modal Information Retrieval10 Jun 2025 0 repositories listed
-
Unlocking the Potential of Large Language Models in the Nuclear Industry with Synthetic Data10 Jun 2025 0 repositories listed
-
PolitiSky24: U.S. Political Bluesky Dataset with User Stance Labels9 Jun 2025 0 repositories listed
-
SAR2Struct: Extracting 3D Semantic Structural Representation of Aircraft Targets from Single-View SAR Image7 Jun 2025 0 repositories listed
-
BioMol-MQA: A Multi-Modal Question Answering Dataset For LLM Reasoning Over Bio-Molecular Interactions6 Jun 2025 0 repositories listed
-
A Survey on Vietnamese Document Analysis and Recognition: Challenges and Future Directions5 Jun 2025 0 repositories listed
-
APVR: Hour-Level Long Video Understanding with Adaptive Pivot Visual Information Retrieval5 Jun 2025 0 repositories listed
-
Preface to the Special Issue of the TAL Journal on Scholarly Document Processing4 Jun 2025 0 repositories listed
-
Evaluating the Unseen Capabilities: How Many Theorems Do LLMs Know?1 Jun 2025 0 repositories listed
-
CoQuIR: A Comprehensive Benchmark for Code Quality-Aware Information Retrieval31 May 2025 0 repositories listed
-
CLIP-AE: CLIP-assisted Cross-view Audio-Visual Enhancement for Unsupervised Temporal Action Localization29 May 2025 0 repositories listed
-
Map&Make: Schema Guided Text to Table Generation29 May 2025 0 repositories listed
-
Anveshana: A New Benchmark Dataset for Cross-Lingual Information Retrieval On English Queries and Sanskrit Documents26 May 2025 0 repositories listed
-
It's High Time: A Survey of Temporal Information Retrieval and Question Answering26 May 2025 0 repositories listed
-
DeepResearchGym: A Free, Transparent, and Reproducible Evaluation Sandbox for Deep Research25 May 2025 0 repositories listed
-
DocMMIR: A Framework for Document Multi-modal Information Retrieval25 May 2025 0 repositories listed
-
Likert or Not: LLM Absolute Relevance Judgments on Fine-Grained Ordinal Scales25 May 2025 0 repositories listed
-
InstructPart: Task-Oriented Part Segmentation with Instruction Reasoning23 May 2025 0 repositories listed
-
Intent Classification on Low-Resource Languages with Query Similarity Search23 May 2025 0 repositories listed
-
Reinforcement Speculative Decoding for Fast Ranking23 May 2025 0 repositories listed
-
Align-GRAG: Reasoning-Guided Dual Alignment for Graph Retrieval-Augmented Generation22 May 2025 0 repositories listed
-
Don't "Overthink" Passage Reranking: Is Reasoning Truly Necessary?22 May 2025 0 repositories listed
-
Fixing Data That Hurts Performance: Cascading LLMs to Relabel Hard Negatives for Robust Information Retrieval22 May 2025 0 repositories listed
-
Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models22 May 2025 0 repositories listed
-
Learning Normal Patterns in Musical Loops22 May 2025 0 repositories listed
-
MiLQ: Benchmarking IR Models for Bilingual Web Search with Mixed Language Queries22 May 2025 0 repositories listed
-
Search Wisely: Mitigating Sub-optimal Agentic Searches By Reducing Uncertainty22 May 2025 0 repositories listed
-
Tools in the Loop: Quantifying Uncertainty of LLM Question Answering Systems That Use Tools22 May 2025 0 repositories listed
-
Bridge the Gap between Past and Future: Siamese Model Optimization for Context-Aware Document Ranking20 May 2025 0 repositories listed
-
Interpretable Traces, Unexpected Outcomes: Investigating the Disconnect in Trace-Based Knowledge Distillation20 May 2025 0 repositories listed
-
NExT-Search: Rebuilding User Feedback Ecosystem for Generative AI Search20 May 2025 0 repositories listed
-
Unified Cross-modal Translation of Score Images, Symbolic Music, and Performance Audio19 May 2025 0 repositories listed
-
On Next-Token Prediction in LLMs: How End Goals Determine the Consistency of Decoding Algorithms16 May 2025 0 repositories listed
-
CRISP: Clustering Multi-Vector Representations for Denoising and Pruning16 May 2025 0 repositories listed
-
Comparing Lexical and Semantic Vector Search Methods When Classifying Medical Documents16 May 2025 0 repositories listed
-
Towards Robust Evaluation of STEM Education: Leveraging MLLMs in Project-Based Learning16 May 2025 0 repositories listed
-
ALOHA: Empowering Multilingual Agent for University Orientation with Hierarchical Retrieval13 May 2025 0 repositories listed
-
Evaluating LLM Metrics Through Real-World Capabilities13 May 2025 0 repositories listed
-
Hakim: Farsi Text Embedding Model13 May 2025 0 repositories listed
-
Lost in Transliteration: Bridging the Script Gap in Neural IR13 May 2025 0 repositories listed
-
TRAIL: Trace Reasoning and Agentic Issue Localization13 May 2025 0 repositories listed
-
MedEIR: A Specialized Medical Embedding Model for Enhanced Information Retrieval12 May 2025 0 repositories listed
-
QUPID: Quantified Understanding for Enhanced Performance, Insights, and Decisions in Korean Search Engines12 May 2025 0 repositories listed
-
Artifact Sharing for Information Retrieval Research8 May 2025 0 repositories listed
-
MARK: Memory Augmented Refinement of Knowledge8 May 2025 0 repositories listed
-
QBR: A Question-Bank-Based Approach to Fine-Grained Legal Knowledge Retrieval for the General Public8 May 2025 0 repositories listed
-
QBD-RankedDataGen: Generating Custom Ranked Datasets for Improving Query-By-Document Search Using LLM-Reranking with Reduced Human Effort7 May 2025 0 repositories listed
-
Towards Large-scale Generative Ranking7 May 2025 0 repositories listed
-
CB-cPIR: Code-Based Computational Private Information Retrieval6 May 2025 0 repositories listed
-
Analyzing Cognitive Differences Among Large Language Models through the Lens of Social Worldview4 May 2025 0 repositories listed
-
Bell's Inequalities and Entanglement in Corpora of Italian Language4 May 2025 0 repositories listed
-
Exploring new Approaches for Information Retrieval through Natural Language Processing4 May 2025 0 repositories listed
-
Scalable Unit Harmonization in Medical Informatics Using Bi-directional Transformers and Bayesian-Optimized BM25 and Sentence Embedding Retrieval1 May 2025 0 repositories listed
-
In a Few Words: Comparing Weak Supervision and LLMs for Short Query Intent Classification30 Apr 2025 0 repositories listed
-
TartuNLP at SemEval-2025 Task 5: Subject Tagging as Two-Stage Information Retrieval30 Apr 2025 0 repositories listed
-
Are Information Retrieval Approaches Good at Harmonising Longitudinal Survey Questions in Social Science?29 Apr 2025 0 repositories listed
-
Federated One-Shot Learning with Data Privacy and Objective-Hiding29 Apr 2025 0 repositories listed
-
OpenTCM: A GraphRAG-Empowered LLM-based System for Traditional Chinese Medicine Knowledge Retrieval and Diagnosis28 Apr 2025 0 repositories listed
-
LLM-Evaluation Tropes: Perspectives on the Validity of LLM-Evaluations27 Apr 2025 0 repositories listed
-
Speaker Retrieval in the Wild: Challenges, Effectiveness and Robustness26 Apr 2025 0 repositories listed
-
Pushing the boundary on Natural Language Inference25 Apr 2025 0 repositories listed
-
Replication and Exploration of Generative Retrieval over Dynamic Corpora24 Apr 2025 0 repositories listed
-
Unsupervised Corpus Poisoning Attacks in Continuous Space for Dense Retrieval24 Apr 2025 0 repositories listed
-
CiteFix: Enhancing RAG Accuracy Through Post-Processing Citation Correction22 Apr 2025 0 repositories listed
-
CLIRudit: Cross-Lingual Information Retrieval of Scientific Documents22 Apr 2025 0 repositories listed
-
Stitching Inner Product and Euclidean Metrics for Topology-aware Maximum Inner Product Search21 Apr 2025 0 repositories listed
-
The 1st EReL@MIR Workshop on Efficient Representation Learning for Multimodal Information Retrieval21 Apr 2025 0 repositories listed
-
LegalRAG: A Hybrid RAG System for Multilingual Legal Information Retrieval19 Apr 2025 0 repositories listed
-
Accommodate Knowledge Conflicts in Retrieval-augmented LLMs: Towards Reliable Response Generation in the Wild17 Apr 2025 0 repositories listed
-
FreshStack: Building Realistic Benchmarks for Evaluating Retrieval on Technical Documents17 Apr 2025 0 repositories listed
-
How Large Language Models Are Changing MOOC Essay Answers: A Comparison of Pre- and Post-LLM Responses17 Apr 2025 0 repositories listed
-
Validating LLM-Generated Relevance Labels for Educational Resource Search17 Apr 2025 0 repositories listed
-
Clarifying Ambiguities: on the Role of Ambiguity Types in Prompting Methods for Clarification Generation16 Apr 2025 0 repositories listed
-
Optimizing Compound Retrieval Systems16 Apr 2025 0 repositories listed
-
CSPLADE: Learned Sparse Retrieval with Causal Language Models15 Apr 2025 0 repositories listed
-
Progressive Rock Music Classification15 Apr 2025 0 repositories listed
-
Streamlining Biomedical Research with Specialized LLMs15 Apr 2025 0 repositories listed
-
Brain-Machine Interfaces & Information Retrieval Challenges and Opportunities14 Apr 2025 0 repositories listed
-
On Precomputation and Caching in Information Retrieval Experiments with Pipeline Architectures14 Apr 2025 0 repositories listed
-
Ordinary Least Squares as an Attention Mechanism13 Apr 2025 0 repositories listed
-
Span-level Emotion-Cause-Category Triplet Extraction with Instruction Tuning LLMs and Data Augmentation13 Apr 2025 0 repositories listed
-
A Reproducibility Study of Graph-Based Legal Case Retrieval11 Apr 2025 0 repositories listed
-
Code-Craft: Hierarchical Graph-Based Code Summarization for Enhanced Context Retrieval11 Apr 2025 0 repositories listed
-
Cross-Document Cross-Lingual NLI via RST-Enhanced Graph Fusion and Interpretability Prediction11 Apr 2025 0 repositories listed
-
Knowledge Graph-extended Retrieval Augmented Generation for Question Answering11 Apr 2025 0 repositories listed
-
Learning from Elders: Making an LLM-powered Chatbot for Retirement Communities more Accessible through User-centered Design11 Apr 2025 0 repositories listed
-
VLMT: Vision-Language Multimodal Transformer for Multimodal Multi-hop Question Answering11 Apr 2025 0 repositories listed
-
Evaluating Retrieval Augmented Generative Models for Document Queries in Transportation Safety9 Apr 2025 0 repositories listed
-
MicroNN: An On-device Disk-resident Updatable Vector Database8 Apr 2025 0 repositories listed
-
CCSK:Cognitive Convection of Self-Knowledge Based Retrieval Augmentation for Large Language Models7 Apr 2025 0 repositories listed
-
Unleashing the Power of LLMs in Dense Retrieval with Query Likelihood Modeling7 Apr 2025 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.