Browse State-of-the-Art › Question Answering › Papers, page 54
Question Answering
Papers archive 2025-07-28
archive papers tagged: 10,817 · with a code link: 4,171 · where Syntology ran a sample: 1,274 (1,073 with a run with no instrument failure, 201 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,274 of 10,817 tagged: 1,073 with a run with no instrument failure, 201 where every run was a failure of Syntology's instrument)
Page 54 of 109: papers 5,301 to 5,400 of 10,817, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
AHP-Powered LLM Reasoning for Multi-Criteria Evaluation of Open-Ended Responses2 Oct 2024 0 repositories listed
-
Backdooring Vision-Language Models with Out-Of-Distribution Data2 Oct 2024 0 repositories listed
-
2 Oct 2024 0 repositories listed Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
CALF: Benchmarking Evaluation of LFQA Using Chinese Examinations2 Oct 2024 0 repositories listed
-
PCQPR: Proactive Conversational Question Planning with Reflection2 Oct 2024 0 repositories listed
-
Reasoning Elicitation in Language Models via Counterfactual Feedback2 Oct 2024 0 repositories listed
-
Why context matters in VQA and Reasoning: Semantic interventions for VLM input modalities2 Oct 2024 0 repositories listed
-
Addition is All You Need for Energy-efficient Language Models1 Oct 2024 0 repositories listed
-
Benchmarking Large Language Models for Conversational Question Answering in Multi-instructional Documents1 Oct 2024 0 repositories listed
-
FMBench: Benchmarking Fairness in Multimodal Large Language Models on Medical Tasks1 Oct 2024 0 repositories listed
-
Quantifying reliance on external information over parametric knowledge during Retrieval Augmented Generation (RAG) using mechanistic analysis1 Oct 2024 0 repositories listed
-
Self-Updatable Large Language Models with Parameter Integration1 Oct 2024 0 repositories listed
-
See then Tell: Enhancing Key Information Extraction with Vision Grounding29 Sep 2024 0 repositories listed
-
Video DataFlywheel: Resolving the Impossible Data Trinity in Video-Language Understanding29 Sep 2024 0 repositories listed
-
3D-CT-GPT: Generating 3D Radiology Reports through Integration of Large Vision-Language Models28 Sep 2024 0 repositories listed
-
HealthQ: Unveiling Questioning Capabilities of LLM Chains in Healthcare Conversations28 Sep 2024 0 repositories listed
-
TrojVLM: Backdoor Attack Against Vision Language Models28 Sep 2024 0 repositories listed
-
Zero-Shot Multi-Hop Question Answering via Monte-Carlo Tree Search with Large Language Models28 Sep 2024 0 repositories listed
-
AIPatient: Simulating Patients with EHRs and LLM Powered Agentic Workflow27 Sep 2024 0 repositories listed
-
Charting the Future: Using Chart Question-Answering for Scalable Evaluation of LLM-Driven Data Visualizations27 Sep 2024 0 repositories listed
-
Enhancing Explainability in Multimodal Large Language Models Using Ontological Context27 Sep 2024 0 repositories listed
-
Rehearsing Answers to Probable Questions with Perspective-Taking27 Sep 2024 0 repositories listed
-
Revisiting the Superficial Alignment Hypothesis27 Sep 2024 0 repositories listed
-
DARE: Diverse Visual Question Answering with Robustness Evaluation26 Sep 2024 0 repositories listed
-
Efficient In-Domain Question Answering for Resource-Constrained Environments26 Sep 2024 0 repositories listed
-
Episodic Memory Verbalization using Hierarchical Representations of Life-Long Robot Experience26 Sep 2024 0 repositories listed
-
Integrating Hierarchical Semantic into Iterative Generation Model for Entailment Tree Explanation26 Sep 2024 0 repositories listed
-
Robotic Environmental State Recognition with Pre-Trained Vision-Language Models and Black-Box Optimization26 Sep 2024 0 repositories listed
-
T3: A Novel Zero-shot Transfer Learning Framework Iteratively Training on an Assistant Task for a Target Task26 Sep 2024 0 repositories listed
-
ZALM3: Zero-Shot Enhancement of Vision-Language Alignment via In-Context Information in Multi-Turn Multimodal Medical Dialogue26 Sep 2024 0 repositories listed
-
Enhancing Post-Hoc Attributions in Long Document Comprehension via Coarse Grained Answer Decomposition25 Sep 2024 0 repositories listed
-
Enhancing Temporal Sensitivity and Reasoning for Time-Sensitive Question Answering25 Sep 2024 0 repositories listed
-
AsthmaBot: Multi-modal, Multi-Lingual Retrieval Augmented Generation For Asthma Patient Support24 Sep 2024 0 repositories listed
-
60 Data Points are Sufficient to Fine-Tune LLMs for Question-Answering24 Sep 2024 0 repositories listed
-
Expert-level vision-language foundation model for real-world radiology and comprehensive evaluation24 Sep 2024 0 repositories listed
-
From Pixels to Words: Leveraging Explainability in Face Recognition through Interactive Natural Language Processing24 Sep 2024 0 repositories listed
-
Lighter And Better: Towards Flexible Context Adaptation For Retrieval Augmented Generation24 Sep 2024 0 repositories listed
-
A Preliminary Study of o1 in Medicine: Are We Closer to an AI Doctor?23 Sep 2024 0 repositories listed
-
Can CLIP Count Stars? An Empirical Study on Quantity Bias in CLIP23 Sep 2024 0 repositories listed
-
Detect, Describe, Discriminate: Moving Beyond VQA for MLLM Evaluation23 Sep 2024 0 repositories listed
-
GEM-RAG: Graphical Eigen Memories For Retrieval Augmented Generation23 Sep 2024 0 repositories listed
-
Learning When to Retrieve, What to Rewrite, and How to Respond in Conversational QA23 Sep 2024 0 repositories listed
-
LINKAGE: Listwise Ranking among Varied-Quality References for Non-Factoid QA Evaluation via LLMs23 Sep 2024 0 repositories listed
-
Using Similarity to Evaluate Factual Consistency in Summaries23 Sep 2024 0 repositories listed
-
Evaluating the Performance and Robustness of LLMs in Materials Science Q&A and Property Predictions22 Sep 2024 0 repositories listed
-
SMART-RAG: Selection using Determinantal Matrices for Augmented Retrieval21 Sep 2024 0 repositories listed
-
@Bench: Benchmarking Vision-Language Models for Human-centered Assistive Technology21 Sep 2024 0 repositories listed
-
Drift to Remember21 Sep 2024 0 repositories listed
-
A Multimodal Dense Retrieval Approach for Speech-Based Open-Domain Question Answering20 Sep 2024 0 repositories listed
-
Enhancing Large Language Models with Domain-specific Retrieval Augment Generation: A Case Study on Long-form Consumer Health Question Answering in Ophthalmology20 Sep 2024 0 repositories listed
-
First Place Solution to the Multiple-choice Video QA Track of The Second Perception Test Challenge20 Sep 2024 0 repositories listed
-
Unlocking Memorization in Large Language Models with Dynamic Soft Prompting20 Sep 2024 0 repositories listed
-
CamelEval: Advancing Culturally Aligned Arabic Language Models and Benchmarks19 Sep 2024 0 repositories listed
-
Edu-Values: Towards Evaluating the Chinese Education Values of Large Language Models19 Sep 2024 0 repositories listed
-
TACO-RL: Task Aware Prompt Compression Optimization with Reinforcement Learning19 Sep 2024 0 repositories listed
-
Vision Language Models Can Parse Floor Plan Maps19 Sep 2024 0 repositories listed
-
Finetuning Language Models to Emit Linguistic Expressions of Uncertainty18 Sep 2024 0 repositories listed
-
MQA-KEAL: Multi-hop Question Answering under Knowledge Editing for Arabic Language18 Sep 2024 0 repositories listed
-
Contextual Breach: Assessing the Robustness of Transformer-based QA Models17 Sep 2024 0 repositories listed
-
OneEncoder: A Lightweight Framework for Progressive Alignment of Modalities17 Sep 2024 0 repositories listed
-
ProSLM : A Prolog Synergized Language Model for explainable Domain Specific Knowledge Based Question Answering17 Sep 2024 0 repositories listed
-
Sparks of Artificial General Intelligence(AGI) in Semiconductor Material Science: Early Explorations into the Next Frontier of Generative AI-Assisted Electron Micrograph Analysis17 Sep 2024 0 repositories listed
-
Uncertainty-Guided Self-Questioning and Answering for Video-Language Alignment17 Sep 2024 0 repositories listed
-
StruEdit: Structured Outputs Enable the Fast and Accurate Knowledge Editing for Large Language Models16 Sep 2024 0 repositories listed
-
A Benchmark Dataset with Larger Context for Non-Factoid Question Answering over Islamic Text15 Sep 2024 0 repositories listed
-
Explore the Hallucination on Low-level Perception for MLLMs15 Sep 2024 0 repositories listed
-
NEVLP: Noise-Robust Framework for Efficient Vision-Language Pre-training15 Sep 2024 0 repositories listed
-
QTG-VQA: Question-Type-Guided Architectural for VideoQA Systems14 Sep 2024 0 repositories listed
-
Contextual Evaluation of Large Language Models for Classifying Tropical and Infectious Diseases13 Sep 2024 0 repositories listed
-
Contri(e)ve: Context + Retrieve for Scholarly Question Answering13 Sep 2024 0 repositories listed
-
Electrocardiogram Report Generation and Question Answering via Retrieval-Augmented Self-Supervised Modeling13 Sep 2024 0 repositories listed
-
Expediting and Elevating Large Language Model Reasoning via Hidden Chain-of-Thought Decoding13 Sep 2024 0 repositories listed
-
KodeXv0.1: A Family of State-of-the-Art Financial Large Language Models13 Sep 2024 0 repositories listed
-
Enhancing Q&A Text Retrieval with Ranking Models: Benchmarking, fine-tuning and deploying Rerankers for RAG12 Sep 2024 0 repositories listed
-
Experimenting with Legal AI Solutions: The Case of Question-Answering for Access to Justice12 Sep 2024 0 repositories listed
-
Multi-object event graph representation learning for Video Question Answering12 Sep 2024 0 repositories listed
-
OmniQuery: Contextually Augmenting Captured Multimodal Memory to Enable Personal Question Answering12 Sep 2024 0 repositories listed
-
Source2Synth: Synthetic Data Generation and Curation Grounded in Real Data Sources12 Sep 2024 0 repositories listed
-
Top-down Activity Representation Learning for Video Question Answering12 Sep 2024 0 repositories listed
-
Integrating SPARQL and LLMs for Question Answering over Scholarly Data Sources11 Sep 2024 0 repositories listed
-
Learning to Compress Contexts for Efficient Knowledge-based Visual Question Answering11 Sep 2024 0 repositories listed
-
MEDIC: Towards a Comprehensive Framework for Evaluating LLMs in Clinical Applications11 Sep 2024 0 repositories listed
-
Securing Vision-Language Models with a Robust Encoder Against Jailbreak and Adversarial Attacks11 Sep 2024 0 repositories listed
-
Accelerating Large Language Model Pretraining via LFR Pedagogy: Learn, Focus, and Review10 Sep 2024 0 repositories listed
-
Enhancing Temporal Understanding in Audio Question Answering for Large Audio Language Models10 Sep 2024 0 repositories listed
-
Mitigating Hallucination in Visual-Language Models via Re-Balancing Contrastive Decoding10 Sep 2024 0 repositories listed
-
VisScience: An Extensive Benchmark for Evaluating K12 Educational Multi-modal Scientific Reasoning10 Sep 2024 0 repositories listed
-
Breaking Neural Network Scaling Laws with Modularity9 Sep 2024 0 repositories listed
-
MLLM-LLaVA-FL: Multimodal Large Language Model Assisted Federated Learning9 Sep 2024 0 repositories listed
-
Seek and Solve Reasoning for Table Question Answering9 Sep 2024 0 repositories listed
-
Towards Building a Robust Knowledge Intensive Question Answering Model with Large Language Models9 Sep 2024 0 repositories listed
-
Combining LLMs and Knowledge Graphs to Reduce Hallucinations in Question Answering6 Sep 2024 0 repositories listed
-
WebQuest: A Benchmark for Multimodal QA on Web Page Sequences6 Sep 2024 0 repositories listed
-
MARAGS: A Multi-Adapter System for Multi-Task Retrieval Augmented Generation Question Answering5 Sep 2024 0 repositories listed
-
OccLLaMA: An Occupancy-Language-Action Generative World Model for Autonomous Driving5 Sep 2024 0 repositories listed
-
RAG based Question-Answering for Contextual Response Prediction System5 Sep 2024 0 repositories listed
-
Vietnamese Legal Information Retrieval in Question-Answering System5 Sep 2024 0 repositories listed
-
Diversify-verify-adapt: Efficient and Robust Retrieval-Augmented Ambiguous Question Answering4 Sep 2024 0 repositories listed
-
GoT-CQA: Graph-of-Thought Guided Compositional Reasoning for Chart Question Answering4 Sep 2024 0 repositories listed
-
How Privacy-Savvy Are Large Language Models? A Case Study on Compliance and Privacy Technical Review4 Sep 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.