Browse State-of-the-Art › Question Answering › Papers, page 57
Question Answering
Papers archive 2025-07-28
archive papers tagged: 10,817 · with a code link: 4,171 · where Syntology ran a sample: 1,274 (1,073 with a run with no instrument failure, 201 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,274 of 10,817 tagged: 1,073 with a run with no instrument failure, 201 where every run was a failure of Syntology's instrument)
Page 57 of 109: papers 5,601 to 5,700 of 10,817, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Zero-Shot Long-Form Video Understanding through Screenplay25 Jun 2024 0 repositories listed
-
Is your benchmark truly adversarial? AdvScore: Evaluating Human-Grounded Adversarialness24 Jun 2024 0 repositories listed
-
24 Jun 2024 0 repositories listed
-
Context-augmented Retrieval: A Novel Framework for Fast Information Retrieval based Response Generation using Large Language Model24 Jun 2024 0 repositories listed
-
Directed Domain Fine-Tuning: Tailoring Separate Modalities for Specific Training Tasks24 Jun 2024 0 repositories listed
-
GPT-4V Explorations: Mining Autonomous Driving24 Jun 2024 0 repositories listed
-
MM-SpuBench: Towards Better Understanding of Spurious Biases in Multimodal LLMs24 Jun 2024 0 repositories listed
-
Modulating Language Model Experiences through Frictions24 Jun 2024 0 repositories listed
-
SEAM: A Stochastic Benchmark for Multi-Document Tasks23 Jun 2024 0 repositories listed
-
MR-MLLM: Mutual Reinforcement of Multimodal Comprehension and Vision Perception22 Jun 2024 0 repositories listed
-
70B-parameter large language models in Japanese medical question-answering21 Jun 2024 0 repositories listed
-
Generate-then-Ground in Retrieval-Augmented Generation for Multi-hop Question Answering21 Jun 2024 0 repositories listed
-
Prompting Whisper for QA-driven Zero-shot End-to-end Spoken Language Understanding21 Jun 2024 0 repositories listed
-
Sports Intelligence: Assessing the Sports Understanding Capabilities of Language Models through Question Answering from Text to Video21 Jun 2024 0 repositories listed
-
Towards Retrieval Augmented Generation over Large Video Libraries21 Jun 2024 0 repositories listed
-
Tri-VQA: Triangular Reasoning Medical Visual Question Answering for Multi-Attribute Analysis21 Jun 2024 0 repositories listed
-
A Learn-Then-Reason Model Towards Generalization in Knowledge Base Question Answering20 Jun 2024 0 repositories listed
-
Does Object Grounding Really Reduce Hallucination of Large Vision-Language Models?20 Jun 2024 0 repositories listed
-
Investigating Mysteries of CoT-Augmented Distillation20 Jun 2024 0 repositories listed
-
PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference20 Jun 2024 0 repositories listed
-
Ranking LLMs by compression20 Jun 2024 0 repositories listed
-
Robust Few-shot Transfer Learning for Knowledge Base Question Answering with Unanswerable Questions20 Jun 2024 0 repositories listed
-
SynDARin: Synthesising Datasets for Automated Reasoning in Low-Resource Languages20 Jun 2024 0 repositories listed
-
Temporal Knowledge Graph Question Answering: A Survey20 Jun 2024 0 repositories listed
-
The Fire Thief Is Also the Keeper: Balancing Usability and Privacy in Prompts20 Jun 2024 0 repositories listed
-
TTQA-RS- A break-down prompting approach for Multi-hop Table-Text Question Answering with Reasoning and Summarization20 Jun 2024 0 repositories listed
-
Understanding Finetuning for Factual Knowledge Extraction20 Jun 2024 0 repositories listed
-
Comparison of Open-Source and Proprietary LLMs for Machine Reading Comprehension: A Practical Analysis for Industrial Applications19 Jun 2024 0 repositories listed
-
FoRAG: Factuality-optimized Retrieval Augmented Generation for Web-enhanced Long-form Question Answering19 Jun 2024 0 repositories listed
-
QRMeM: Unleash the Length Limitation through Question then Reflection Memory Mechanism19 Jun 2024 0 repositories listed
-
Thread: A Logic-Based Data Organization Paradigm for How-To Question Answering with Retrieval Augmented Generation19 Jun 2024 0 repositories listed
-
Towards Robust Evaluation: A Comprehensive Taxonomy of Datasets and Metrics for Open Domain Question Answering in the Era of Large Language Models19 Jun 2024 0 repositories listed
-
Transferable speech-to-text large language model alignment module19 Jun 2024 0 repositories listed
-
Towards Understanding Domain Adapted Sentence Embeddings for Document Retrieval18 Jun 2024 0 repositories listed
-
From RAGs to rich parameters: Probing how language models utilize external knowledge over parametric information for factual queries18 Jun 2024 0 repositories listed
-
Intermediate Distillation: Data-Efficient Distillation from Black-Box LLMs for Information Retrieval18 Jun 2024 0 repositories listed
-
LightPAL: Lightweight Passage Retrieval for Open Domain Multi-Document Summarization18 Jun 2024 0 repositories listed
-
Exploring the Robustness of Language Models for Tabular Question Answering via Attention Analysis18 Jun 2024 0 repositories listed
-
Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs17 Jun 2024 0 repositories listed
-
SeRTS: Self-Rewarding Tree Search for Biomedical Retrieval-Augmented Generation17 Jun 2024 0 repositories listed
-
Hallucination Mitigation Prompts Long-term Video Understanding17 Jun 2024 0 repositories listed
-
InternalInspector I²: Robust Confidence Estimation in LLMs through Internal States17 Jun 2024 0 repositories listed
-
Iterative Utility Judgment Framework via LLMs Inspired by Relevance in Philosophy17 Jun 2024 0 repositories listed
-
LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning17 Jun 2024 0 repositories listed
-
Mitigating Large Language Model Hallucination with Faithful Finetuning17 Jun 2024 0 repositories listed
-
Context Graph17 Jun 2024 0 repositories listed
-
Program Synthesis Benchmark for Visual Programming in XLogoOnline Environment17 Jun 2024 0 repositories listed
-
Adaptive Query Rewriting: Aligning Rewriters through Marginal Probability of Conversational Answers16 Jun 2024 0 repositories listed
-
Multi-LLM QA with Embodied Exploration16 Jun 2024 0 repositories listed
-
HiddenTables & PyQTax: A Cooperative Game and Dataset For TableQA to Ensure Scale and Data Privacy Across a Myriad of Taxonomies16 Jun 2024 0 repositories listed
-
Towards Lifelong Dialogue Agents via Timeline-based Memory Management16 Jun 2024 0 repositories listed
-
On the Hardness of Faithful Chain-of-Thought Reasoning in Large Language Models15 Jun 2024 0 repositories listed
-
MMLU-SR: A Benchmark for Stress-Testing Reasoning Capability of Large Language Models15 Jun 2024 0 repositories listed
-
VCEval: Rethinking What is a Good Educational Video and How to Automatically Evaluate It15 Jun 2024 0 repositories listed
-
Datasets for Multilingual Answer Sentence Selection14 Jun 2024 0 repositories listed
-
Detecting and Evaluating Medical Hallucinations in Large Vision Language Models14 Jun 2024 0 repositories listed
-
14 Jun 2024 0 repositories listed
-
Enhancing Question Answering on Charts Through Effective Pre-training Tasks14 Jun 2024 0 repositories listed
-
EWEK-QA: Enhanced Web and Efficient Knowledge Graph Retrieval for Citation-based Question Answering Systems14 Jun 2024 0 repositories listed
-
GLiNER multi-task: Generalist Lightweight Model for Various Information Extraction Tasks14 Jun 2024 0 repositories listed
-
A Training-free Sub-quadratic Cost Transformer Model Serving Framework With Hierarchically Pruned Attention14 Jun 2024 0 repositories listed
-
Integrating Large Language Models with Graph-based Reasoning for Conversational Question Answering14 Jun 2024 0 repositories listed
-
Precision Empowers, Excess Distracts: Visual Question Answering With Dynamically Infused Knowledge In Language Models14 Jun 2024 0 repositories listed
-
SHMamba: Structured Hyperbolic State Space Model for Audio-Visual Question Answering14 Jun 2024 0 repositories listed
-
DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding13 Jun 2024 0 repositories listed
-
Multi-Modal Retrieval For Large Language Model Based Speech Recognition13 Jun 2024 0 repositories listed
-
Optimizing Visual Question Answering Models for Driving: Bridging the Gap Between Human and Machine Attention Patterns13 Jun 2024 0 repositories listed
-
DistilDoc: Knowledge Distillation for Visually-Rich Document Applications12 Jun 2024 0 repositories listed
-
Dynamic Stochastic Decoding Strategy for Open-Domain Dialogue Generation12 Jun 2024 0 repositories listed
-
Prediction of the Realisation of an Information Need: An EEG Study12 Jun 2024 0 repositories listed
-
Research Trends for the Interplay between Large Language Models and Knowledge Graphs12 Jun 2024 0 repositories listed
-
DR-RAG: Applying Dynamic Document Relevance to Retrieval-Augmented Generation for Question-Answering11 Jun 2024 0 repositories listed
-
Efficient Parallel Multi-Hop Reasoning: A Scalable Approach for Knowledge Graph Analysis11 Jun 2024 0 repositories listed
-
Paraphrasing in Affirmative Terms Improves Negation Understanding11 Jun 2024 0 repositories listed
-
Question-Answering (QA) Model for a Personalized Learning Assistant for Arabic Language11 Jun 2024 0 repositories listed
-
BrainChat: Decoding Semantic Information from fMRI using Vision-language Pretrained Models10 Jun 2024 0 repositories listed
-
CVQA: Culturally-diverse Multilingual Visual Question Answering Benchmark10 Jun 2024 0 repositories listed
-
Evaluating the Retrieval Component in LLM-Based Question Answering Systems10 Jun 2024 0 repositories listed
-
Harnessing AI for efficient analysis of complex policy documents: a case study of Executive Order 1411010 Jun 2024 0 repositories listed
-
HOLMES: Hyper-Relational Knowledge Graphs for Multi-hop Question Answering using LLMs10 Jun 2024 0 repositories listed
-
Solution for SMART-101 Challenge of CVPR Multi-modal Algorithmic Reasoning Task 202410 Jun 2024 0 repositories listed
-
Transforming Wearable Data into Health Insights using Large Language Model Agents10 Jun 2024 0 repositories listed
-
Do LLMs Exhibit Human-Like Reasoning? Evaluating Theory of Mind in LLMs for Open-Ended Responses9 Jun 2024 0 repositories listed
-
MedREQAL: Examining Medical Knowledge Recall of Large Language Models via Question Answering9 Jun 2024 0 repositories listed
-
MrRank: Improving Question Answering Retrieval System through Multi-Result Ranking Model9 Jun 2024 0 repositories listed
-
Zero-Shot End-To-End Spoken Question Answering In Medical Domain9 Jun 2024 0 repositories listed
-
CaLM: Contrasting Large and Small Language Models to Verify Grounded Generation8 Jun 2024 0 repositories listed
-
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts8 Jun 2024 0 repositories listed
-
Investigating and Addressing Hallucinations of LLMs in Tasks Involving Negation8 Jun 2024 0 repositories listed
-
Venn Diagram Prompting : Accelerating Comprehension with Scaffolding Effect8 Jun 2024 0 repositories listed
-
CRiskEval: A Chinese Multi-Level Risk Evaluation Benchmark Dataset for Large Language Models7 Jun 2024 0 repositories listed
-
MATTER: Memory-Augmented Transformer Using Heterogeneous Knowledge Sources7 Jun 2024 0 repositories listed
-
TCMD: A Traditional Chinese Medicine QA Dataset for Evaluating Large Language Models7 Jun 2024 0 repositories listed
-
Efficient Knowledge Infusion via KG-LLM Alignment6 Jun 2024 0 repositories listed
-
Synthesizing Conversations from Unlabeled Documents using Automatic Response Segmentation6 Jun 2024 0 repositories listed
-
Understanding Information Storage and Transfer in Multi-modal Large Language Models6 Jun 2024 0 repositories listed
-
Why Has Predicting Downstream Capabilities of Frontier AI Models with Scale Remained Elusive?6 Jun 2024 0 repositories listed
-
Balancing Performance and Efficiency in Zero-shot Robotic Navigation5 Jun 2024 0 repositories listed
-
Discovering Bias in Latent Space: An Unsupervised Debiasing Approach5 Jun 2024 0 repositories listed
-
IrokoBench: A New Benchmark for African Languages in the Age of Large Language Models5 Jun 2024 0 repositories listed