Browse State-of-the-Art › Question Answering › Papers, page 53
Question Answering
Papers archive 2025-07-28
archive papers tagged: 10,817 · with a code link: 4,171 · where Syntology ran a sample: 1,274 (1,073 with a run with no instrument failure, 201 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,274 of 10,817 tagged: 1,073 with a run with no instrument failure, 201 where every run was a failure of Syntology's instrument)
Page 53 of 109: papers 5,201 to 5,300 of 10,817, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Object-Centric Temporal Consistency via Conditional Autoregressive Inductive Biases21 Oct 2024 0 repositories listed
-
RAG4ITOps: A Supervised Fine-Tunable and Comprehensive RAG Framework for IT Operations and Maintenance21 Oct 2024 0 repositories listed
-
xGen-MM-Vid (BLIP-3-Video): You Only Need 32 Tokens to Represent a Video Even in VLMs21 Oct 2024 0 repositories listed
-
Evaluating Consistencies in LLM responses through a Semantic Clustering of Question Answering20 Oct 2024 0 repositories listed
-
ChitroJera: A Regionally Relevant Visual Question Answering Dataset for Bangla19 Oct 2024 0 repositories listed
-
Coarse-to-Fine Highlighting: Reducing Knowledge Hallucination in Large Language Models19 Oct 2024 0 repositories listed
-
LLaVA-Ultra: Large Chinese Language and Vision Assistant for Ultrasound19 Oct 2024 0 repositories listed
-
Addressing Blind Guessing: Calibration of Selection Bias in Multiple-Choice Question Answering by Video Language Models18 Oct 2024 0 repositories listed
-
Bridging the Training-Inference Gap in LLMs by Leveraging Self-Generated Tokens18 Oct 2024 0 repositories listed
-
DiscoGraMS: Enhancing Movie Screen-Play Summarization using Movie Character-Aware Discourse Graph18 Oct 2024 0 repositories listed
-
E3D-GPT: Enhanced 3D Visual Foundation for Medical Vision-Language Model18 Oct 2024 0 repositories listed
-
Electrocardiogram-Language Model for Few-Shot Question Answering with Meta Learning18 Oct 2024 0 repositories listed
-
MCSFF: Multi-modal Consistency and Specificity Fusion Framework for Entity Alignment18 Oct 2024 0 repositories listed
-
NaturalBench: Evaluating Vision-Language Models on Natural Adversarial Samples18 Oct 2024 0 repositories listed
-
Optimizing Retrieval-Augmented Generation with Elasticsearch for Enhanced Question-Answering Systems18 Oct 2024 0 repositories listed
-
RA-BLIP: Multimodal Adaptive Retrieval-Augmented Bootstrapping Language-Image Pre-training18 Oct 2024 0 repositories listed
-
SPFresh: Incremental In-Place Update for Billion-Scale Vector Search18 Oct 2024 0 repositories listed
-
SwaQuAD-24: QA Benchmark Dataset in Swahili18 Oct 2024 0 repositories listed
-
Zero-shot Action Localization via the Confidence of Large Vision-Language Models18 Oct 2024 0 repositories listed
-
Accounting for Sycophancy in Language Model Uncertainty Estimation17 Oct 2024 0 repositories listed
-
AdaSwitch: Adaptive Switching between Small and Large Agents for Effective Cloud-Local Collaborative Learning17 Oct 2024 0 repositories listed
-
Advancing Large Language Model Attribution through Self-Improving17 Oct 2024 0 repositories listed
-
BQA: Body Language Question Answering Dataset for Video Large Language Models17 Oct 2024 0 repositories listed
-
Evaluating Self-Generated Documents for Enhancing Retrieval-Augmented Generation with Large Language Models17 Oct 2024 0 repositories listed
-
FinQAPT: Empowering Financial Decisions with End-to-End LLM-driven Question Answering Pipeline17 Oct 2024 0 repositories listed
-
From Isolated Conversations to Hierarchical Schemas: Dynamic Tree Memory Representation for LLMs17 Oct 2024 0 repositories listed
-
RescueADI: Adaptive Disaster Interpretation in Remote Sensing Images with Autonomous Agents17 Oct 2024 0 repositories listed
-
An Automatic and Cost-Efficient Peer-Review Framework for Language Generation Evaluation16 Oct 2024 0 repositories listed
-
16 Oct 2024 0 repositories listed
-
Large Language Models as a Tool for Mining Object Knowledge16 Oct 2024 0 repositories listed
-
Open Domain Question Answering with Conflicting Contexts16 Oct 2024 0 repositories listed
-
Pyramid-Driven Alignment: Pyramid Principle Guided Integration of Large Language Models and Knowledge Graphs16 Oct 2024 0 repositories listed
-
REFINE on Scarce Data: Retrieval Enhancement through Fine-Tuning via Model Fusion of Embedding Models16 Oct 2024 0 repositories listed
-
AGENTiGraph: An Interactive Knowledge Graph Platform for LLM-based Chatbots Utilizing Private Data15 Oct 2024 0 repositories listed
-
Causal Reasoning in Large Language Models: A Knowledge Graph Approach15 Oct 2024 0 repositories listed
-
Empowering Users in Digital Privacy Management through Interactive LLM-Based Agents15 Oct 2024 0 repositories listed
-
LargePiG: Your Large Language Model is Secretly a Pointer Generator15 Oct 2024 0 repositories listed
-
OMCAT: Omni Context Aware Transformer15 Oct 2024 0 repositories listed
-
15 Oct 2024 0 repositories listed
-
Telco-DPR: A Hybrid Dataset for Evaluating Retrieval Models of 3GPP Technical Specifications15 Oct 2024 0 repositories listed
-
Unleashing the Power of LLMs as Multi-Modal Encoders for Text and Graph-Structured Data15 Oct 2024 0 repositories listed
-
BanglaQuAD: A Bengali Open-domain Question Answering Dataset14 Oct 2024 0 repositories listed
-
Eliminating the Language Bias for Visual Question Answering with fine-grained Causal Intervention14 Oct 2024 0 repositories listed
-
FLARE: Faithful Logic-Aided Reasoning and Exploration14 Oct 2024 0 repositories listed
-
A Step Towards Mixture of Grader: Statistical Analysis of Existing Automatic Evaluation Metrics13 Oct 2024 0 repositories listed
-
ChartKG: A Knowledge-Graph-Based Representation for Chart Images13 Oct 2024 0 repositories listed
-
LoRE: Logit-Ranked Retriever Ensemble for Enhancing Open-Domain Question Answering13 Oct 2024 0 repositories listed
-
13 Oct 2024 0 repositories listed
-
Surgical-LLaVA: Toward Surgical Scenario Understanding via Large Language and Vision Models13 Oct 2024 0 repositories listed
-
Enhanced Electronic Health Records Text Summarization Using Large Language Models12 Oct 2024 0 repositories listed
-
Prompting Video-Language Foundation Models with Domain-specific Fine-grained Heuristics for Video Question Answering12 Oct 2024 0 repositories listed
-
Measuring the Groundedness of Legal Question-Answering Systems11 Oct 2024 0 repositories listed
-
Optimized Biomedical Question-Answering Services with LLM and Multi-BERT Integration11 Oct 2024 0 repositories listed
-
Retrieving Contextual Information for Long-Form Question Answering using Weak Supervision11 Oct 2024 0 repositories listed
-
ViT3D Alignment of LLaMA3: 3D Medical Image Report Generation11 Oct 2024 0 repositories listed
-
Can Knowledge Graphs Make Large Language Models More Trustworthy? An Empirical Study over Open-ended Question Answering10 Oct 2024 0 repositories listed
-
Do You Know What You Are Talking About? Characterizing Query-Knowledge Relevance For Reliable Retrieval Augmented Generation10 Oct 2024 0 repositories listed
-
Emerging Pixel Grounding in Large Multimodal Models Without Grounding Supervision10 Oct 2024 0 repositories listed
-
Increasing the Difficulty of Automatically Generated Questions via Reinforcement Learning with Synthetic Preference10 Oct 2024 0 repositories listed
-
MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models10 Oct 2024 0 repositories listed
-
Optima: Optimizing Effectiveness and Efficiency for LLM-Based Multi-Agent System10 Oct 2024 0 repositories listed
-
Rewriting Conversational Utterances with Instructed Large Language Models10 Oct 2024 0 repositories listed
-
SAKA: An Intelligent Platform for Semi-automated Knowledge Graph Construction and Application10 Oct 2024 0 repositories listed
-
Sample then Identify: A General Framework for Risk Control and Assessment in Multimodal Large Language Models10 Oct 2024 0 repositories listed
-
10 Oct 2024 0 repositories listed
-
Do great minds think alike? Investigating Human-AI Complementarity in Question Answering with CAIMIRA9 Oct 2024 0 repositories listed
-
Enhancing Multimodal LLM for Detailed and Accurate Video Captioning using Multi-Round Preference Optimization9 Oct 2024 0 repositories listed
-
FltLM: An Intergrated Long-Context Large Language Model for Effective Context Filtering and Understanding9 Oct 2024 0 repositories listed
-
Personal Intelligence System UniLM: Hybrid On-Device Small Language Model and Server-Based Large Language Model for Malay Nusantara9 Oct 2024 0 repositories listed
-
PAR: Prompt-Aware Token Reduction Method for Efficient Large Multimodal Models9 Oct 2024 0 repositories listed
-
SEGMENT+: Long Text Processing with Short-Context Language Models9 Oct 2024 0 repositories listed
-
Uncovering Factor Level Preferences to Improve Human-Model Alignment9 Oct 2024 0 repositories listed
-
ActionAtlas: A VideoQA Benchmark for Domain-specialized Action Recognition8 Oct 2024 0 repositories listed
-
Beyond Captioning: Task-Specific Prompting for Improved VLM Performance in Mathematical Reasoning8 Oct 2024 0 repositories listed
-
Information Discovery in e-Commerce8 Oct 2024 0 repositories listed
-
PortLLM: Personalizing Evolving Large Language Models with Training-Free and Portable Model Patches8 Oct 2024 0 repositories listed
-
Temporal Reasoning Transfer from Text to Video8 Oct 2024 0 repositories listed
-
Document-level Causal Relation Extraction with Knowledge-guided Binary Question Answering7 Oct 2024 0 repositories listed
-
Mitigating the Risk of Health Inequity Exacerbated by Large Language Models7 Oct 2024 0 repositories listed
-
MM-R³: On (In-)Consistency of Multi-modal Large Language Models (MLLMs)7 Oct 2024 0 repositories listed
-
Precise Model Benchmarking with Only a Few Observations7 Oct 2024 0 repositories listed
-
VLM2Vec: Training Vision-Language Models for Massive Multimodal Embedding Tasks7 Oct 2024 0 repositories listed
-
5 Oct 2024 0 repositories listed Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Beyond Forecasting: Compositional Time Series Reasoning for End-to-End Task Execution5 Oct 2024 0 repositories listed
-
Overview of Factify5WQA: Fact Verification through 5W Question-Answering5 Oct 2024 0 repositories listed
-
ALR²: A Retrieve-then-Reason Framework for Long-context Question Answering4 Oct 2024 0 repositories listed
-
Cross-lingual Transfer for Automatic Question Generation by Learning Interrogative Structures in Target Languages4 Oct 2024 0 repositories listed
-
Frame-Voyager: Learning to Query Frames for Video Large Language Models4 Oct 2024 0 repositories listed
-
Learning Semantic Structure through First-Order-Logic Translation4 Oct 2024 0 repositories listed
-
Question-Answering System for Bangla: Fine-tuning BERT-Bangla for a Closed Domain4 Oct 2024 0 repositories listed
-
Structured List-Grounded Question Answering4 Oct 2024 0 repositories listed
-
A Comprehensive Survey of Retrieval-Augmented Generation (RAG): Evolution, Current Landscape and Future Directions3 Oct 2024 0 repositories listed
-
Coal Mining Question Answering with LLMs3 Oct 2024 0 repositories listed
-
Distilling an End-to-End Voice Assistant Without Instruction Training Data3 Oct 2024 0 repositories listed
-
Domain-Specific Retrieval-Augmented Generation Using Vector Stores, Knowledge Graphs, and Tensor Factorization3 Oct 2024 0 repositories listed
-
Grounding Large Language Models In Embodied Environment With Imperfect World Models3 Oct 2024 0 repositories listed
-
Listening to the Wise Few: Select-and-Copy Attention Heads for Multiple-Choice QA3 Oct 2024 0 repositories listed
-
3 Oct 2024 0 repositories listed
-
Unlocking Structured Thinking in Language Models with Cognitive Prompting3 Oct 2024 0 repositories listed
-
3 Oct 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.