Browse State-of-the-Art › Question Answering › Papers, page 47
Question Answering
Papers archive 2025-07-28
archive papers tagged: 10,817 · with a code link: 4,171 · where Syntology ran a sample: 1,274 (1,073 with a run with no instrument failure, 201 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,274 of 10,817 tagged: 1,073 with a run with no instrument failure, 201 where every run was a failure of Syntology's instrument)
Page 47 of 109: papers 4,601 to 4,700 of 10,817, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
KoGNER: A Novel Framework for Knowledge Graph Distillation on Biomedical Named Entity Recognition19 Mar 2025 0 repositories listed
-
MAMM-Refine: A Recipe for Improving Faithfulness in Generation with Multi-Agent Collaboration19 Mar 2025 0 repositories listed
-
TruthLens:A Training-Free Paradigm for DeepFake Detection19 Mar 2025 0 repositories listed
-
UPME: An Unsupervised Peer Review Framework for Multimodal Large Language Model Evaluation19 Mar 2025 0 repositories listed
-
CARE: A QLoRA-Fine Tuned Multi-Domain Chatbot With Fast Learning On Minimal Hardware18 Mar 2025 0 repositories listed
-
EIAD: Explainable Industrial Anomaly Detection Via Multi-Modal Large Language Models18 Mar 2025 0 repositories listed
-
Identifying and Mitigating Position Bias of Multi-image Vision-Language Models18 Mar 2025 0 repositories listed
-
18 Mar 2025 0 repositories listed Syntology 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 3 harvested samples) · 3 pointer-only (licence)
-
Synthetic Clarification and Correction Dialogues about Data-Centric Tasks -- A Teacher-Student Approach18 Mar 2025 0 repositories listed
-
Synthetic Data Generation Using Large Language Models: Advances in Text and Code18 Mar 2025 0 repositories listed
-
Uncertainty Distillation: Teaching Language Models to Express Semantic Confidence18 Mar 2025 0 repositories listed
-
From Head to Tail: Towards Balanced Representation in Large Vision-Language Models through Adaptive Data Calibration17 Mar 2025 0 repositories listed
-
HIS-GPT: Towards 3D Human-In-Scene Multimodal Understanding17 Mar 2025 0 repositories listed
-
Knowledge-Aware Iterative Retrieval for Multi-Agent Systems17 Mar 2025 0 repositories listed
-
Logic-in-Frames: Dynamic Keyframe Search via Visual Semantic-Logical Verification for Long Video Understanding17 Mar 2025 0 repositories listed
-
Long-VMNet: Accelerating Long-Form Video Understanding via Fixed Memory17 Mar 2025 0 repositories listed
-
MAP: Evaluation and Multi-Agent Enhancement of Large Language Models for Inpatient Pathways17 Mar 2025 0 repositories listed
-
RAG-RL: Advancing Retrieval-Augmented Generation via RL and Curriculum Learning17 Mar 2025 0 repositories listed
-
Sightation Counts: Leveraging Sighted User Feedback in Building a BLV-aligned Dataset of Diagram Descriptions17 Mar 2025 0 repositories listed
-
Task-Oriented Feature Compression for Multimodal Understanding via Device-Edge Co-Inference17 Mar 2025 0 repositories listed
-
Unified Autoregressive Visual Generation and Understanding with Continuous Tokens17 Mar 2025 0 repositories listed
-
VITED: Video Temporal Evidence Distillation17 Mar 2025 0 repositories listed
-
General Table Question Answering via Answer-Formula Joint Generation16 Mar 2025 0 repositories listed
-
GeoRSMLLM: A Multimodal Large Language Model for Vision-Language Tasks in Geoscience and Remote Sensing16 Mar 2025 0 repositories listed
-
PEBench: A Fictitious Dataset to Benchmark Machine Unlearning for Multimodal Large Language Models16 Mar 2025 0 repositories listed
-
Beyond the Destination: A Novel Benchmark for Exploration-Aware Embodied Question Answering14 Mar 2025 0 repositories listed
-
DynRsl-VLM: Enhancing Autonomous Driving Perception with Dynamic Resolution Vision-Language Models14 Mar 2025 0 repositories listed
-
MUSS: Multilevel Subset Selection for Relevance and Diversity14 Mar 2025 0 repositories listed
-
UMB@PerAnsSumm 2025: Enhancing Perspective-Aware Summarization with Prompt Optimization and Supervised Fine-Tuning14 Mar 2025 0 repositories listed
-
KV-Distill: Nearly Lossless Learnable Context Compression for LLMs13 Mar 2025 0 repositories listed
-
Learning to Inference Adaptively for Multimodal Large Language Models13 Mar 2025 0 repositories listed
-
TIME: Temporal-sensitive Multi-dimensional Instruction Tuning and Benchmarking for Video-LLMs13 Mar 2025 0 repositories listed
-
Unlock the Power of Unlabeled Data in Language Driving Model13 Mar 2025 0 repositories listed
-
Everything Can Be Described in Words: A Simple Unified Multi-Modal Framework with Semantic and Temporal Alignment12 Mar 2025 0 repositories listed
-
FaVChat: Unlocking Fine-Grained Facail Video Understanding with Multimodal Large Language Models12 Mar 2025 0 repositories listed
-
On the Limitations of Vision-Language Models in Understanding Image Transforms12 Mar 2025 0 repositories listed
-
SurgicalVLM-Agent: Towards an Interactive AI Co-Pilot for Pituitary Surgery12 Mar 2025 0 repositories listed
-
A Survey on Knowledge-Oriented Retrieval-Augmented Generation11 Mar 2025 0 repositories listed
-
Bring Remote Sensing Object Detect Into Nature Language Model: Using SFT Method11 Mar 2025 0 repositories listed
-
DAFE: LLM-Based Evaluation Through Dynamic Arbitration for Free-Form Question-Answering11 Mar 2025 0 repositories listed
-
FASIONAD++ : Integrating High-Level Instruction and Information Bottleneck in FAt-Slow fusION Systems for Enhanced Safety in Autonomous Driving with Adaptive Feedback11 Mar 2025 0 repositories listed
-
Gradient-guided Attention Map Editing: Towards Efficient Contextual Hallucination Mitigation11 Mar 2025 0 repositories listed
-
Seeing and Reasoning with Confidence: Supercharging Multimodal LLMs with an Uncertainty-Aware Agentic Framework11 Mar 2025 0 repositories listed
-
UniF²ace: Fine-grained Face Understanding and Generation with Unified Multimodal Models11 Mar 2025 0 repositories listed
-
From Text to Visuals: Using LLMs to Generate Math Diagrams with Vector Graphics10 Mar 2025 0 repositories listed
-
MapQA: Open-domain Geospatial Question Answering on Map Data10 Mar 2025 0 repositories listed
-
ReAgent: Reversible Multi-Agent Reasoning for Knowledge-Enhanced Multi-Hop QA10 Mar 2025 0 repositories listed
-
Robusto-1 Dataset: Comparing Humans and VLMs on real out-of-distribution Autonomous Driving VQA from Peru10 Mar 2025 0 repositories listed
-
Taking Notes Brings Focus? Towards Multi-Turn Multimodal Dialogue Learning10 Mar 2025 0 repositories listed
-
Talking to GDELT Through Knowledge Graphs10 Mar 2025 0 repositories listed
-
Towards Fine-Grained Video Question Answering10 Mar 2025 0 repositories listed
-
Delusions of Large Language Models9 Mar 2025 0 repositories listed
-
Green Prompting9 Mar 2025 0 repositories listed
-
Human Cognition Inspired RAG with Knowledge Graph for Complex Problem Solving9 Mar 2025 0 repositories listed
-
TI-JEPA: An Innovative Energy-based Joint Embedding Strategy for Text-Image Multimodal Systems9 Mar 2025 0 repositories listed
-
Vector Quantized Feature Fields for Fast 3D Semantic Lifting9 Mar 2025 0 repositories listed
-
VisualSimpleQA: A Benchmark for Decoupled Evaluation of Large Vision-Language Models in Fact-Seeking Question Answering9 Mar 2025 0 repositories listed
-
Integrating Frequency-Domain Representations with Low-Rank Adaptation in Vision-Language Models8 Mar 2025 0 repositories listed
-
MoEMoE: Question Guided Dense and Scalable Sparse Mixture-of-Expert for Multi-source Multi-modal Answering8 Mar 2025 0 repositories listed
-
SplatTalk: 3D VQA with Gaussian Splatting8 Mar 2025 0 repositories listed
-
Correctness Coverage Evaluation for Medical Multiple-Choice Question Answering Based on the Enhanced Conformal Prediction Framework7 Mar 2025 0 repositories listed
-
Architecture for a Trustworthy Quantum Chatbot6 Mar 2025 0 repositories listed
-
Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities6 Mar 2025 0 repositories listed
-
Chart-HQA: A Benchmark for Hypothetical Question Answering in Charts6 Mar 2025 0 repositories listed
-
Dynamic-KGQA: A Scalable Framework for Generating Adaptive Question Answering Datasets6 Mar 2025 0 repositories listed
-
Enhancing SAM with Efficient Prompting and Preference Optimization for Semi-supervised Medical Image Segmentation6 Mar 2025 0 repositories listed
-
Evaluating Answer Reranking Strategies in Time-sensitive Question Answering6 Mar 2025 0 repositories listed
-
LVLM-Compress-Bench: Benchmarking the Broader Impact of Large Vision-Language Model Compression6 Mar 2025 0 repositories listed
-
FANS -- Formal Answer Selection for Natural Language Math Reasoning Using Lean45 Mar 2025 0 repositories listed
-
Structured Outputs Enable General-Purpose LLMs to be Medical Experts5 Mar 2025 0 repositories listed
-
Towards Understanding Multi-Round Large Language Model Reasoning: Approximability, Learnability and Generalizability5 Mar 2025 0 repositories listed
-
Vision-Language Models Struggle to Align Entities across Modalities5 Mar 2025 0 repositories listed
-
EchoQA: A Large Collection of Instruction Tuning Data for Echocardiogram Reports4 Mar 2025 0 repositories listed
-
Optimizing open-domain question answering with graph-based retrieval augmented generation4 Mar 2025 0 repositories listed
-
OWLViz: An Open-World Benchmark for Visual Question Answering4 Mar 2025 0 repositories listed
-
Beyond Prompting: An Efficient Embedding Framework for Open-Domain Question Answering3 Mar 2025 0 repositories listed
-
SAGE: A Framework of Precise Retrieval for RAG3 Mar 2025 0 repositories listed
-
Causal Tree Extraction from Medical Case Reports: A Novel Task for Experts-like Text Comprehension3 Mar 2025 0 repositories listed
-
Generate, Discriminate, Evolve: Enhancing Context Faithfulness via Fine-Grained Sentence-Level Self-Evolution3 Mar 2025 0 repositories listed
-
Parameter-free Video Segmentation for Vision and Language Understanding3 Mar 2025 0 repositories listed
-
SRAG: Structured Retrieval-Augmented Generation for Multi-Entity Question Answering over Wikipedia Graph3 Mar 2025 0 repositories listed
-
Towards Efficient Educational Chatbots: Benchmarking RAG Frameworks2 Mar 2025 0 repositories listed
-
ER-RAG: Enhance RAG with ER-Based Unified Modeling of Heterogeneous Data Sources2 Mar 2025 0 repositories listed
-
FunBench: Benchmarking Fundus Reading Skills of MLLMs2 Mar 2025 0 repositories listed
-
Optimizing Multi-Hop Document Retrieval Through Intermediate Representations2 Mar 2025 0 repositories listed
-
CL-MoE: Enhancing Multimodal Large Language Model with Dual Momentum Mixture-of-Experts for Continual Visual Question Answering1 Mar 2025 0 repositories listed
-
TempRetriever: Fusion-based Temporal Dense Passage Retrieval for Time-Sensitive Questions28 Feb 2025 0 repositories listed
-
PreMind: Multi-Agent Video Understanding for Advanced Indexing of Presentation-style Videos28 Feb 2025 0 repositories listed
-
Fine-Grained Retrieval-Augmented Generation for Visual Question Answering28 Feb 2025 0 repositories listed
-
WebFAQ: A Multilingual Collection of Natural Q&A Datasets for Dense Retrieval28 Feb 2025 0 repositories listed
-
Bisecting K-Means in RAG for Enhancing Question-Answering Tasks Performance in Telecommunications27 Feb 2025 0 repositories listed
-
Can Large Language Models Unveil the Mysteries? An Exploration of Their Ability to Unlock Information in Complex Scenarios27 Feb 2025 0 repositories listed
-
From Retrieval to Generation: Comparing Different Approaches27 Feb 2025 0 repositories listed
-
M-LLM Based Video Frame Selection for Efficient Video Understanding27 Feb 2025 0 repositories listed
-
Med-RLVR: Emerging Medical Reasoning from a 3B base model via reinforcement Learning27 Feb 2025 0 repositories listed
-
END: Early Noise Dropping for Efficient and Effective Context Denoising26 Feb 2025 0 repositories listed
-
MEBench: Benchmarking Large Language Models for Cross-Document Multi-Entity Question Answering26 Feb 2025 0 repositories listed
-
MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning26 Feb 2025 0 repositories listed
-
Nexus: An Omni-Perceptive And -Interactive Model for Language, Audio, And Vision26 Feb 2025 0 repositories listed
-
Talking to the brain: Using Large Language Models as Proxies to Model Brain Semantic Representation26 Feb 2025 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.