Browse State-of-the-Art › Question Answering › Papers, page 61
Question Answering
Papers archive 2025-07-28
archive papers tagged: 10,817 · with a code link: 4,171 · where Syntology ran a sample: 1,274 (1,073 with a run with no instrument failure, 201 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,274 of 10,817 tagged: 1,073 with a run with no instrument failure, 201 where every run was a failure of Syntology's instrument)
Page 61 of 109: papers 6,001 to 6,100 of 10,817, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
TV-TREES: Multimodal Entailment Trees for Neuro-Symbolic Video Reasoning29 Feb 2024 0 repositories listed
-
A Cognitive Evaluation Benchmark of Image Reasoning and Description for Large Vision-Language Models28 Feb 2024 0 repositories listed
-
Can GPT Improve the State of Prior Authorization via Guideline Based Automated Question Answering?28 Feb 2024 0 repositories listed
-
ArcSin: Adaptive ranged cosine Similarity injected noise for Language-Driven Visual Tasks27 Feb 2024 0 repositories listed
-
Reasoning in Conversation: Solving Subjective Tasks through Dialogue Simulation for Large Language Models27 Feb 2024 0 repositories listed
-
27 Feb 2024 0 repositories listed
-
Self-Refinement of Language Models from External Proxy Metrics Feedback27 Feb 2024 0 repositories listed
-
Unsupervised multiple choices question answering via universal corpus27 Feb 2024 0 repositories listed
-
VCD: Knowledge Base Guided Visual Commonsense Discovery in Images27 Feb 2024 0 repositories listed
-
Read and Think: An Efficient Step-wise Multimodal Language Model for Document Understanding and Reasoning26 Feb 2024 0 repositories listed
-
GigaPevt: Multimodal Medical Assistant26 Feb 2024 0 repositories listed
-
PAQA: Toward ProActive Open-Retrieval Question Answering26 Feb 2024 0 repositories listed
-
PerLTQA: A Personal Long-Term Memory Dataset for Memory Classification, Retrieval, and Synthesis in Question Answering26 Feb 2024 0 repositories listed
-
Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts26 Feb 2024 0 repositories listed
-
26 Feb 2024 0 repositories listed
-
Deep Learning Approaches for Improving Question Answering Systems in Hepatocellular Carcinoma Research25 Feb 2024 0 repositories listed
-
Prompt Perturbation Consistency Learning for Robust Language Models24 Feb 2024 0 repositories listed
-
ArabianGPT: Native Arabic GPT-based Large Language Model23 Feb 2024 0 repositories listed
-
Cost-Adaptive Recourse Recommendation by Adaptive Preference Elicitation23 Feb 2024 0 repositories listed
-
DOSA: A Dataset of Social Artifacts from Different Indian Geographical Subcultures23 Feb 2024 0 repositories listed
-
Evaluating the Performance of ChatGPT for Spam Email Detection23 Feb 2024 0 repositories listed
-
23 Feb 2024 0 repositories listed
-
Multimodal Transformer With a Low-Computational-Cost Guarantee23 Feb 2024 0 repositories listed
-
VISREAS: Complex Visual Reasoning with Unanswerable Questions23 Feb 2024 0 repositories listed
-
Does the Generator Mind its Contexts? An Analysis of Generative Model Faithfulness under Context Transfer22 Feb 2024 0 repositories listed
-
Word-Sequence Entropy: Towards Uncertainty Estimation in Free-Form Medical Question Answering Applications and Beyond22 Feb 2024 0 repositories listed
-
LLMs Meet Long Video: Advancing Long Video Question Answering with An Interactive Visual Adapter in LLMs21 Feb 2024 0 repositories listed
-
Self-DC: When to Reason and When to Act? Self Divide-and-Conquer for Compositional Unknown Questions21 Feb 2024 0 repositories listed
-
Exploring the Frontier of Vision-Language Models: A Survey of Current Methodologies and Future Directions20 Feb 2024 0 repositories listed
-
Exploring the Impact of Table-to-Text Methods on Augmenting LLM-based Question Answering with Domain Hybrid Data20 Feb 2024 0 repositories listed
-
Modality-Aware Integration with Large Language Models for Knowledge-based Visual Question Answering20 Feb 2024 0 repositories listed
-
20 Feb 2024 0 repositories listed
-
Slot-VLM: SlowFast Slots for Video-Language Modeling20 Feb 2024 0 repositories listed
-
VideoPrism: A Foundational Visual Encoder for Video Understanding20 Feb 2024 0 repositories listed
-
BIDER: Bridging Knowledge Inconsistency for Efficient Retrieval-Augmented LLMs via Key Supporting Evidence19 Feb 2024 0 repositories listed
-
Graph-Based Retriever Captures the Long Tail of Biomedical Knowledge19 Feb 2024 0 repositories listed
-
Model Tailor: Mitigating Catastrophic Forgetting in Multi-modal Large Language Models19 Feb 2024 0 repositories listed
-
Cofca: A Step-Wise Counterfactual Multi-hop QA benchmark19 Feb 2024 0 repositories listed
-
RJUA-MedDQA: A Multimodal Benchmark for Medical Document Question Answering and Clinical Reasoning19 Feb 2024 0 repositories listed
-
Tables as Texts or Images: Evaluating the Table Reasoning Ability of LLMs and MLLMs19 Feb 2024 0 repositories listed
-
Training Table Question Answering via SQL Query Decomposition19 Feb 2024 0 repositories listed
-
Large Language Models Can Better Understand Knowledge Graphs Than We Thought18 Feb 2024 0 repositories listed
-
Question Answering Over Spatio-Temporal Knowledge Graph18 Feb 2024 0 repositories listed
-
A Question Answering Based Pipeline for Comprehensive Chinese EHR Information Extraction17 Feb 2024 0 repositories listed
-
Evaluating LLMs' Mathematical Reasoning in Financial Document Question Answering17 Feb 2024 0 repositories listed
-
CliqueParcel: An Approach For Batching LLM Prompts That Jointly Optimizes Efficiency And Faithfulness17 Feb 2024 0 repositories listed
-
GenDec: A robust generative Question-decomposition method for Multi-hop reasoning17 Feb 2024 0 repositories listed
-
BlendFilter: Advancing Retrieval-Augmented Large Language Models via Query Generation Blending and Knowledge Filtering16 Feb 2024 0 repositories listed
-
Construction of a Syntactic Analysis Map for Yi Shui School through Text Mining and Natural Language Processing Research16 Feb 2024 0 repositories listed
-
Exploring Hybrid Question Answering via Program-based Prompting16 Feb 2024 0 repositories listed
-
Inference to the Best Explanation in Large Language Models16 Feb 2024 0 repositories listed
-
PaLM2-VAdapter: Progressively Aligned Language Model Makes a Strong Vision-language Adapter16 Feb 2024 0 repositories listed
-
PAT-Questions: A Self-Updating Benchmark for Present-Anchored Temporal Question-Answering16 Feb 2024 0 repositories listed
-
16 Feb 2024 0 repositories listed Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Assessing biomedical knowledge robustness in large language models by query-efficient sampling attacks16 Feb 2024 0 repositories listed
-
A Dataset of Open-Domain Question Answering with Multiple-Span Answers15 Feb 2024 0 repositories listed
-
Enhancing Large Language Models with Pseudo- and Multisource- Knowledge Graphs for Open-ended Question Answering15 Feb 2024 0 repositories listed
-
15 Feb 2024 0 repositories listed
-
Prompt-based Personalized Federated Learning for Medical Visual Question Answering15 Feb 2024 0 repositories listed
-
Multi-Query Focused Disaster Summarization via Instruction-Based Prompting14 Feb 2024 0 repositories listed
-
Learning How To Ask: Cycle-Consistency Refines Prompts in Multimodal Foundation Models13 Feb 2024 0 repositories listed
-
Visual Question Answering Instruction: Unlocking Multimodal Large Language Model To Domain-Specific Visual Multitasks13 Feb 2024 0 repositories listed
-
BDIQA: A New Dataset for Video Question Answering to Explore Cognitive Reasoning through Theory of Mind12 Feb 2024 0 repositories listed
-
Lumos : Empowering Multimodal LLMs with Scene Text Recognition12 Feb 2024 0 repositories listed
-
PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs12 Feb 2024 0 repositories listed
-
Retrieval Augmented Thought Process for Private Data Handling in Healthcare12 Feb 2024 0 repositories listed
-
T-RAG: Lessons from the LLM Trenches12 Feb 2024 0 repositories listed
-
CPSDBench: A Large Language Model Evaluation Benchmark and Baseline for Chinese Public Security Domain11 Feb 2024 0 repositories listed
-
FaBERT: Pre-training BERT on Persian Blogs9 Feb 2024 0 repositories listed
-
The Generative AI Paradox on Evaluation: What It Can Solve, It May Not Evaluate9 Feb 2024 0 repositories listed
-
CIC: A Framework for Culturally-Aware Image Captioning8 Feb 2024 0 repositories listed
-
Efficient Models for the Detection of Hate, Abuse and Profanity8 Feb 2024 0 repositories listed
-
FAQ-Gen: An automated system to generate domain-specific FAQs to aid content comprehension8 Feb 2024 0 repositories listed
-
SubGen: Token Generation in Sublinear Time and Memory8 Feb 2024 0 repositories listed
-
Empowering Language Models with Active Inquiry for Deeper Understanding6 Feb 2024 0 repositories listed
-
Are Machines Better at Complex Reasoning? Unveiling Human-Machine Inference Gaps in Entailment Verification6 Feb 2024 0 repositories listed
-
SceMQA: A Scientific College Entrance Level Multimodal Question Answering Benchmark6 Feb 2024 0 repositories listed
-
A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications5 Feb 2024 0 repositories listed
-
LB-KBQA: Large-language-model and BERT based Knowledge-Based Question and Answering System5 Feb 2024 0 repositories listed
-
eXplainable Bayesian Multi-Perspective Generative Retrieval4 Feb 2024 0 repositories listed
-
Large Language Model for Table Processing: A Survey4 Feb 2024 0 repositories listed
-
PuzzleBench: Can LLMs Solve Challenging First-Order Combinatorial Reasoning Problems?4 Feb 2024 0 repositories listed
-
SemPool: Simple, robust, and interpretable KG pooling for enhancing language models3 Feb 2024 0 repositories listed
-
BAT: Learning to Reason about Spatial Sounds with Large Language Models2 Feb 2024 0 repositories listed
-
LLMs May Perform MCQA by Selecting the Least Incorrect Option2 Feb 2024 0 repositories listed
-
Efficient Prompt Caching via Embedding Similarity2 Feb 2024 0 repositories listed
-
A Chain-of-Thought Is as Strong as Its Weakest Link: A Benchmark for Verifiers of Reasoning Chains1 Feb 2024 0 repositories listed
-
An Exam-based Evaluation Approach Beyond Traditional Relevance Judgments1 Feb 2024 0 repositories listed
-
HiQA: A Hierarchical Contextual Augmentation RAG for Multi-Documents QA1 Feb 2024 0 repositories listed
-
Can Generative AI Support Patients' & Caregivers' Informational Needs? Towards Task-Centric Evaluation Of AI Systems31 Jan 2024 0 repositories listed
-
Binding Touch to Everything: Learning Unified Multimodal Tactile Representations31 Jan 2024 0 repositories listed
-
Desiderata for the Context Use of Question Answering Systems31 Jan 2024 0 repositories listed
-
Fine-tuning Transformer-based Encoder for Turkish Language Understanding Tasks30 Jan 2024 0 repositories listed
-
Security and Privacy Challenges of Large Language Models: A Survey30 Jan 2024 0 repositories listed
-
LCV2: An Efficient Pretraining-Free Framework for Grounded Visual Question Answering29 Jan 2024 0 repositories listed
-
Muffin or Chihuahua? Challenging Multimodal Large Language Models with Multipanel VQA29 Jan 2024 0 repositories listed
-
Towards Optimizing the Costs of LLM Usage29 Jan 2024 0 repositories listed
-
Improving Data Augmentation for Robust Visual Question Answering with Effective Curriculum Learning28 Jan 2024 0 repositories listed
-
A RAG-based Question Answering System Proposal for Understanding Islam: MufassirQAS LLM27 Jan 2024 0 repositories listed
-
DataFrame QA: A Universal LLM Framework on DataFrame Question Answering Without Data Exposure27 Jan 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.