Methods › General › Attention Mechanisms › Attention › Papers, page 130
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 130 of 316: papers 12,901 to 13,000 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Text-aware Speech Separation for Multi-talker Keyword Spotting 18 Jun 2024 · 1 repository · arXiv:2406.12447
-
The Limits of Pure Exploration in POMDPs: When the Observation Entropy is Enough 18 Jun 2024 · 0 repositories · arXiv:2406.12795
-
Towards a Client-Centered Assessment of LLM Therapists by Client Simulation 18 Jun 2024 · 1 repository · arXiv:2406.12266Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Traffic Prediction considering Multiple Levels of Spatial-temporal Information: A Multi-scale Graph Wavelet-based Approach 18 Jun 2024 · 0 repositories · arXiv:2406.13038
-
Translation Equivariant Transformer Neural Processes 18 Jun 2024 · 1 repository · arXiv:2406.12409
-
UBENCH: Benchmarking Uncertainty in Large Language Models with Multiple Choice Questions 18 Jun 2024 · 1 repository · arXiv:2406.12784
-
Unified Active Retrieval for Retrieval Augmented Generation 18 Jun 2024 · 1 repository · arXiv:2406.12534
-
UrbanLLM: Autonomous Urban Activity Planning and Management with Large Language Models 18 Jun 2024 · 0 repositories · arXiv:2406.12360
-
Vernacular? I Barely Know Her: Challenges with Style Control and Stereotyping 18 Jun 2024 · 0 repositories · arXiv:2406.12679
-
VIA: Unified Spatiotemporal Video Adaptation Framework for Global and Local Video Editing 18 Jun 2024 · 0 repositories · arXiv:2406.12831
-
VoCo-LLaMA: Towards Vision Compression with Large Language Models 18 Jun 2024 · 1 repository · arXiv:2406.12275Syntology official (archive's flag): 10 ran · 12 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 6 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 5 pointer-only (licence)
-
Weakly Supervised Learning of Cortical Surface Reconstruction from Segmentations 18 Jun 2024 · 1 repository · arXiv:2406.12650Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
What Makes Two Language Models Think Alike? 18 Jun 2024 · 0 repositories · arXiv:2406.12620
-
"You Gotta be a Doctor, Lin": An Investigation of Name-Based Bias of Large Language Models in Employment Recommendations 18 Jun 2024 · 0 repositories · arXiv:2406.12232
-
A Critical Study of What Code-LLMs (Do Not) Learn 17 Jun 2024 · 1 repository · arXiv:2406.11930
-
A Personalised Learning Tool for Physics Undergraduate Students Built On a Large Language Model for Symbolic Regression 17 Jun 2024 · 0 repositories · arXiv:2407.00065
-
A Simple and Effective L₂ Norm-Based Strategy for KV Cache Compression 17 Jun 2024 · 2 repositories · arXiv:2406.11430Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A Two-dimensional Zero-shot Dialogue State Tracking Evaluation Method using GPT-4 17 Jun 2024 · 1 repository · arXiv:2406.11651
-
An Exploration of Length Generalization in Transformer-Based Speech Enhancement 17 Jun 2024 · 0 repositories · arXiv:2406.11401
-
An Online Approach and Evaluation Method for Tracking People Across Cameras in Extremely Long Video Sequence 17 Jun 2024 · 0 repositories
-
Are Large Language Models True Healthcare Jacks-of-All-Trades? Benchmarking Across Health Professions Beyond Physician Exams 17 Jun 2024 · 1 repository · arXiv:2406.11328
-
Attention-Based Deep Reinforcement Learning for Qubit Allocation in Modular Quantum Architectures 17 Jun 2024 · 0 repositories · arXiv:2406.11452
-
AV-CrossNet: an Audiovisual Complex Spectral Mapping Network for Speech Separation By Leveraging Narrow- and Cross-Band Modeling 17 Jun 2024 · 1 repository · arXiv:2406.11619
-
Beyond Boundaries: Learning a Universal Entity Taxonomy across Datasets and Languages for Open Named Entity Recognition 17 Jun 2024 · 1 repository · arXiv:2406.11192
-
Brain-inspired Computational Modeling of Action Recognition with Recurrent Spiking Neural Networks Equipped with Reinforcement Delay Learning 17 Jun 2024 · 0 repositories · arXiv:2406.11778
-
Breaking Boundaries: Investigating the Effects of Model Editing on Cross-linguistic Performance 17 Jun 2024 · 0 repositories · arXiv:2406.11139
-
Bridging Design Gaps: A Parametric Data Completion Approach With Graph Guided Diffusion Models 17 Jun 2024 · 0 repositories · arXiv:2406.11934
-
Building another Spanish dictionary, this time with GPT-4 17 Jun 2024 · 1 repository · arXiv:2406.11218
-
CItruS: Chunked Instruction-aware State Eviction for Long Sequence Modeling 17 Jun 2024 · 1 repository · arXiv:2406.12018Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Composing Object Relations and Attributes for Image-Text Matching 17 Jun 2024 · 1 repository · arXiv:2406.11820Syntology official (archive's flag): 4 ran · 4 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Computing in the Life Sciences: From Early Algorithms to Modern AI 17 Jun 2024 · 1 repository · arXiv:2406.12108
-
CrAM: Credibility-Aware Attention Modification in LLMs for Combating Misinformation in RAG 17 Jun 2024 · 1 repository · arXiv:2406.11497
-
Crossfusor: A Cross-Attention Transformer Enhanced Conditional Diffusion Model for Car-Following Trajectory Prediction 17 Jun 2024 · 0 repositories · arXiv:2406.11941
-
Cultural Conditioning or Placebo? On the Effectiveness of Socio-Demographic Prompting 17 Jun 2024 · 0 repositories · arXiv:2406.11661
-
Decoding the Narratives: Analyzing Personal Drug Experiences Shared on Reddit 17 Jun 2024 · 0 repositories · arXiv:2406.12117
-
Diffusion-Based Adaptation for Classification of Unknown Degraded Images 17 Jun 2024 · 1 repository
-
Diffusion Models in Low-Level Vision: A Survey 17 Jun 2024 · 1 repository · arXiv:2406.11138
-
Discriminative Hamiltonian Variational Autoencoder for Accurate Tumor Segmentation in Data-Scarce Regimes 17 Jun 2024 · 0 repositories · arXiv:2406.11659
-
Distributed Stochastic Gradient Descent with Staleness: A Stochastic Delay Differential Equation Based Framework 17 Jun 2024 · 0 repositories · arXiv:2406.11159
-
DiTTo-TTS: Diffusion Transformers for Scalable Text-to-Speech without Domain-Specific Factors 17 Jun 2024 · 1 repository · arXiv:2406.11427
-
Duoduo CLIP: Efficient 3D Understanding with Multi-View Images 17 Jun 2024 · 1 repository · arXiv:2406.11579Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Embodied Instruction Following in Unknown Environments 17 Jun 2024 · 0 repositories · arXiv:2406.11818
-
Enabling robots to follow abstract instructions and complete complex dynamic tasks 17 Jun 2024 · 0 repositories · arXiv:2406.11231
-
SeRTS: Self-Rewarding Tree Search for Biomedical Retrieval-Augmented Generation 17 Jun 2024 · 0 repositories · arXiv:2406.11258
-
Enhancing Text Classification through LLM-Driven Active Learning and Human Annotation 17 Jun 2024 · 1 repository · arXiv:2406.12114
-
Estimating the Increase in Emissions caused by AI-augmented Search 17 Jun 2024 · 0 repositories · arXiv:2407.16894
-
Are Small Language Models Ready to Compete with Large Language Models for Practical Applications? 17 Jun 2024 · 1 repository · arXiv:2406.11402
-
Evaluating the Efficacy of Open-Source LLMs in Enterprise-Specific RAG Systems: A Comparative Study of Performance and Scalability 17 Jun 2024 · 1 repository · arXiv:2406.11424
-
ExCP: Extreme LLM Checkpoint Compression via Weight-Momentum Joint Shrinking 17 Jun 2024 · 1 repository · arXiv:2406.11257Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples)
-
Exploring Safety-Utility Trade-Offs in Personalized Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.11107
-
Exploring the Role of Large Language Models in Prompt Encoding for Diffusion Models 17 Jun 2024 · 0 repositories · arXiv:2406.11831
-
Fine-Tuning or Fine-Failing? Debunking Performance Myths in Large Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.11201
-
FinTruthQA: A Benchmark Dataset for Evaluating the Quality of Financial Information Disclosure 17 Jun 2024 · 1 repository · arXiv:2406.12009Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
GameVibe: A Multimodal Affective Game Corpus 17 Jun 2024 · 0 repositories · arXiv:2407.12787
-
GeoGPT4V: Towards Geometric Multi-modal Large Language Models with Geometric Image Generation 17 Jun 2024 · 1 repository · arXiv:2406.11503Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Problematic Tokens: Tokenizer Bias in Large Language Models 17 Jun 2024 · 1 repository · arXiv:2406.11214
-
GPT-Powered Elicitation Interview Script Generator for Requirements Engineering Training 17 Jun 2024 · 0 repositories · arXiv:2406.11439
-
HyperSIGMA: Hyperspectral Intelligence Comprehension Foundation Model 17 Jun 2024 · 1 repository · arXiv:2406.11519Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Improving Multi-Agent Debate with Sparse Communication Topology 17 Jun 2024 · 0 repositories · arXiv:2406.11776
-
Inpainting the Gaps: A Novel Framework for Evaluating Explanation Methods in Vision Transformers 17 Jun 2024 · 0 repositories · arXiv:2406.11534
-
InternalInspector I²: Robust Confidence Estimation in LLMs through Internal States 17 Jun 2024 · 0 repositories · arXiv:2406.12053
-
Intrinsic Evaluation of Unlearning Using Parametric Knowledge Traces 17 Jun 2024 · 1 repository · arXiv:2406.11614Syntology official (archive's flag): 18 ran · 18 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 20 harvested samples) · 20 pointer-only (licence)
-
Investigating Annotator Bias in Large Language Models for Hate Speech Detection 17 Jun 2024 · 3 repositories · arXiv:2406.11109
-
Iterative Length-Regularized Direct Preference Optimization: A Case Study on Improving 7B Language Models to GPT-4 Level 17 Jun 2024 · 0 repositories · arXiv:2406.11817
-
Iterative Utility Judgment Framework via LLMs Inspired by Relevance in Philosophy 17 Jun 2024 · 0 repositories · arXiv:2406.11290
-
JobFair: A Framework for Benchmarking Gender Hiring Bias in Large Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.15484
-
Large Language Models and Knowledge Graphs for Astronomical Entity Disambiguation 17 Jun 2024 · 0 repositories · arXiv:2406.11400
-
Large Scale Transfer Learning for Tabular Data via Language Modeling 17 Jun 2024 · 2 repositories · arXiv:2406.12031Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples)
-
MetaGPT: Merging Large Language Models Using Model Exclusive Task Arithmetic 17 Jun 2024 · 0 repositories · arXiv:2406.11385
-
MFC-Bench: Benchmarking Multimodal Fact-Checking with Large Vision-Language Models 17 Jun 2024 · 1 repository · arXiv:2406.11288
-
Multiple Descents in Unsupervised Learning: The Role of Noise, Domain Shift and Anomalies 17 Jun 2024 · 0 repositories · arXiv:2406.11703
-
SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention 17 Jun 2024 · 0 repositories · arXiv:2406.15486
-
PhyBench: A Physical Commonsense Benchmark for Evaluating Text-to-Image Models 17 Jun 2024 · 0 repositories · arXiv:2406.11802
-
Prior Normality Prompt Transformer for Multi-class Industrial Image Anomaly Detection 17 Jun 2024 · 0 repositories · arXiv:2406.11507
-
Promises, Outlooks and Challenges of Diffusion Language Modeling 17 Jun 2024 · 0 repositories · arXiv:2406.11473
-
R-Eval: A Unified Toolkit for Evaluating Domain Knowledge of Retrieval Augmented Large Language Models 17 Jun 2024 · 1 repository · arXiv:2406.11681
-
Rethinking Spatio-Temporal Transformer for Traffic Prediction:Multi-level Multi-view Augmented Learning Framework 17 Jun 2024 · 0 repositories · arXiv:2406.11921
-
Satyrn: A Platform for Analytics Augmented Generation 17 Jun 2024 · 1 repository · arXiv:2406.12069
-
Scaling the Codebook Size of VQGAN to 100,000 with a Utilization Rate of 99% 17 Jun 2024 · 1 repository · arXiv:2406.11837Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
SEFraud: Graph-based Self-Explainable Fraud Detection via Interpretative Mask Learning 17 Jun 2024 · 0 repositories · arXiv:2406.11389
-
Self and Cross-Model Distillation for LLMs: Effective Methods for Refusal Pattern Alignment 17 Jun 2024 · 0 repositories · arXiv:2406.11285
-
Should AI Optimize Your Code? A Comparative Study of Classical Optimizing Compilers Versus Current Large Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.12146
-
Skip-Layer Attention: Bridging Abstract and Detailed Dependencies in Transformers 17 Jun 2024 · 0 repositories · arXiv:2406.11274
-
Small Agent Can Also Rock! Empowering Small Language Models as Hallucination Detector 17 Jun 2024 · 1 repository · arXiv:2406.11277
-
STNAGNN: Data-driven Spatio-temporal Brain Connectivity beyond FC 17 Jun 2024 · 0 repositories · arXiv:2406.12065
-
Super(ficial)-alignment: Strong Models May Deceive Weak Models in Weak-to-Strong Generalization 17 Jun 2024 · 1 repository · arXiv:2406.11431Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
SWCF-Net: Similarity-weighted Convolution and Local-global Fusion for Efficient Large-scale Point Cloud Semantic Segmentation 17 Jun 2024 · 1 repository · arXiv:2406.11441
-
Refiner: Restructure Retrieval Content Efficiently to Advance Question-Answering Capabilities 17 Jun 2024 · 1 repository · arXiv:2406.11357Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
TRACE the Evidence: Constructing Knowledge-Grounded Reasoning Chains for Retrieval-Augmented Generation 17 Jun 2024 · 2 repositories · arXiv:2406.11460Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
UDBRNet: A novel uncertainty driven boundary refined network for organ at risk segmentation 17 Jun 2024 · 1 repository
-
Understanding Multi-Granularity for Open-Vocabulary Part Segmentation 17 Jun 2024 · 2 repositories · arXiv:2406.11384Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples) · 3 pointer-only (licence)
-
Unveiling and Mitigating Bias in Mental Health Analysis with Large Language Models 17 Jun 2024 · 1 repository · arXiv:2406.12033
-
WellDunn: On the Robustness and Explainability of Language Models and Large Language Models in Identifying Wellness Dimensions 17 Jun 2024 · 1 repository · arXiv:2406.12058Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
YOLO9tr: A Lightweight Model for Pavement Damage Detection Utilizing a Generalized Efficient Layer Aggregation Network and Attention Mechanism 17 Jun 2024 · 1 repository · arXiv:2406.11254
-
Can LLMs Understand the Implication of Emphasized Sentences in Dialogue? 16 Jun 2024 · 1 repository · arXiv:2406.11065
-
Distilling Opinions at Scale: Incremental Opinion Summarization using XL-OPSUMM 16 Jun 2024 · 0 repositories · arXiv:2406.10886
-
Enhancing Supermarket Robot Interaction: A Multi-Level LLM Conversational Interface for Handling Diverse Customer Intents 16 Jun 2024 · 0 repositories · arXiv:2406.11047
-
Exposing the Achilles' Heel: Evaluating LLMs Ability to Handle Mistakes in Mathematical Reasoning 16 Jun 2024 · 0 repositories · arXiv:2406.10834
-
Generating Tables from the Parametric Knowledge of Language Models 16 Jun 2024 · 1 repository · arXiv:2406.10922
-
Geometric Distortion Guided Transformer for Omnidirectional Image Super-Resolution 16 Jun 2024 · 0 repositories · arXiv:2406.10869