Methods › General › Attention Mechanisms › Attention › Papers, page 39
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 39 of 316: papers 3,801 to 3,900 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Hallucination Detection in LLMs Using Spectral Features of Attention Maps 24 Feb 2025 · 1 repository · arXiv:2502.17598
-
Improving the Transferability of Adversarial Examples by Inverse Knowledge Distillation 24 Feb 2025 · 0 repositories · arXiv:2502.17003
-
LettuceDetect: A Hallucination Detection Framework for RAG Applications 24 Feb 2025 · 2 repositories · arXiv:2502.17125
-
LLM Inference Acceleration via Efficient Operation Fusion 24 Feb 2025 · 0 repositories · arXiv:2502.17728
-
Logic Haystacks: Probing LLMs Long-Context Logical Reasoning (Without Easily Identifiable Unrelated Padding) 24 Feb 2025 · 0 repositories · arXiv:2502.17169
-
LongSafety: Evaluating Long-Context Safety of Large Language Models 24 Feb 2025 · 1 repository · arXiv:2502.16971Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
LongSpec: Long-Context Speculative Decoding with Efficient Drafting and Verification 24 Feb 2025 · 1 repository · arXiv:2502.17421Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
MambaFlow: A Novel and Flow-guided State Space Model for Scene Flow Estimation 24 Feb 2025 · 1 repository · arXiv:2502.16907Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
MaxGlaViT: A novel lightweight vision transformer-based approach for early diagnosis of glaucoma stages from fundus images 24 Feb 2025 · 1 repository · arXiv:2502.17154
-
MDN: Mamba-Driven Dualstream Network For Medical Hyperspectral Image Segmentation 24 Feb 2025 · 0 repositories · arXiv:2502.17255
-
MEDA: Dynamic KV Cache Allocation for Efficient Multimodal Long-Context Inference 24 Feb 2025 · 1 repository · arXiv:2502.17599Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
MEMERAG: A Multilingual End-to-End Meta-Evaluation Benchmark for Retrieval Augmented Generation 24 Feb 2025 · 1 repository · arXiv:2502.17163
-
Mitigating Bias in RAG: Controlling the Embedder 24 Feb 2025 · 1 repository · arXiv:2502.17390
-
Mitigating Hallucinations in Diffusion Models through Adaptive Attention Modulation 24 Feb 2025 · 0 repositories · arXiv:2502.16872
-
MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs 24 Feb 2025 · 1 repository · arXiv:2502.17422Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Mutual Reinforcement of LLM Dialogue Synthesis and Summarization Capabilities for Few-Shot Dialogue Summarization 24 Feb 2025 · 0 repositories · arXiv:2502.17328
-
Neural Attention: A Novel Mechanism for Enhanced Expressive Power in Transformer Models 24 Feb 2025 · 0 repositories · arXiv:2502.17206
-
NUTSHELL: A Dataset for Abstract Generation from Scientific Talks 24 Feb 2025 · 0 repositories · arXiv:2502.16942
-
Optimal Recovery Meets Minimax Estimation 24 Feb 2025 · 0 repositories · arXiv:2502.17671
-
Order Matters: Investigate the Position Bias in Multi-constraint Instruction Following 24 Feb 2025 · 1 repository · arXiv:2502.17204
-
Quantifying Logical Consistency in Transformers via Query-Key Alignment 24 Feb 2025 · 0 repositories · arXiv:2502.17017
-
Shakti-VLMs: Scalable Vision-Language Models for Enterprise AI 24 Feb 2025 · 0 repositories · arXiv:2502.17092
-
TabulaTime: A Novel Multimodal Deep Learning Framework for Advancing Acute Coronary Syndrome Prediction through Environmental and Clinical Data Integration 24 Feb 2025 · 0 repositories · arXiv:2502.17049
-
The Lottery LLM Hypothesis, Rethinking What Abilities Should LLM Compression Preserve? 24 Feb 2025 · 0 repositories · arXiv:2502.17535
-
The Role of Sparsity for Length Generalization in Transformers 24 Feb 2025 · 0 repositories · arXiv:2502.16792
-
Towards Typologically Aware Rescoring to Mitigate Unfaithfulness in Lower-Resource Languages 24 Feb 2025 · 0 repositories · arXiv:2502.17664
-
Unraveling the geometry of visual relational reasoning 24 Feb 2025 · 1 repository · arXiv:2502.17382
-
VGFL-SA: Vertical Graph Federated Learning Structure Attack Based on Contrastive Learning 24 Feb 2025 · 0 repositories · arXiv:2502.16793
-
VideoGrain: Modulating Space-Time Attention for Multi-grained Video Editing 24 Feb 2025 · 0 repositories · arXiv:2502.17258
-
VR-Pipe: Streamlining Hardware Graphics Pipeline for Volume Rendering 24 Feb 2025 · 0 repositories · arXiv:2502.17078
-
X-Dancer: Expressive Music to Human Dance Video Generation 24 Feb 2025 · 0 repositories · arXiv:2502.17414
-
A Fine-Tuning Approach for T5 Using Knowledge Graphs to Address Complex Tasks 23 Feb 2025 · 0 repositories · arXiv:2502.16484
-
A Reverse Mamba Attention Network for Pathological Liver Segmentation 23 Feb 2025 · 1 repository · arXiv:2502.18232
-
A Split-Window Transformer for Multi-Model Sequence Spammer Detection using Multi-Model Variational Autoencoder 23 Feb 2025 · 0 repositories · arXiv:2502.16483
-
A Survey of Graph Transformers: Architectures, Theories and Applications 23 Feb 2025 · 0 repositories · arXiv:2502.16533
-
D2S-FLOW: Automated Parameter Extraction from Datasheets for SPICE Model Generation Using Large Language Models 23 Feb 2025 · 0 repositories · arXiv:2502.16540
-
AeroReformer: Aerial Referring Transformer for UAV-based Referring Image Segmentation 23 Feb 2025 · 1 repository · arXiv:2502.16680
-
Attention-based UAV Trajectory Optimization for Wireless Power Transfer-assisted IoT Systems 23 Feb 2025 · 0 repositories · arXiv:2502.17517
-
Benchmarking Online Object Trackers for Underwater Robot Position Locking Applications 23 Feb 2025 · 0 repositories · arXiv:2502.16569
-
Class-Conditional Neural Polarizer: A Lightweight and Effective Backdoor Defense by Purifying Poisoned Features 23 Feb 2025 · 0 repositories · arXiv:2502.18520
-
Co-MTP: A Cooperative Trajectory Prediction Framework with Multi-Temporal Fusion for Autonomous Driving 23 Feb 2025 · 1 repository · arXiv:2502.16589
-
Code Summarization Beyond Function Level 23 Feb 2025 · 1 repository · arXiv:2502.16704
-
CodeCriticBench: A Holistic Code Critique Benchmark for Large Language Models 23 Feb 2025 · 1 repository · arXiv:2502.16614
-
Dynamic LLM Routing and Selection based on User Preferences: Balancing Performance, Cost, and Ethics 23 Feb 2025 · 0 repositories · arXiv:2502.16696
-
Geometry-Aware 3D Salient Object Detection Network 23 Feb 2025 · 0 repositories · arXiv:2502.16488
-
GS-TransUNet: Integrated 2D Gaussian Splatting and Transformer UNet for Accurate Skin Lesion Analysis 23 Feb 2025 · 1 repository · arXiv:2502.16748
-
Intrinsic Model Weaknesses: How Priming Attacks Unveil Vulnerabilities in Large Language Models 23 Feb 2025 · 0 repositories · arXiv:2502.16491
-
Layer-Wise Evolution of Representations in Fine-Tuned Transformers: Insights from Sparse AutoEncoders 23 Feb 2025 · 0 repositories · arXiv:2502.16722
-
Liver Cirrhosis Stage Estimation from MRI with Deep Learning 23 Feb 2025 · 1 repository · arXiv:2502.18225
-
MemeIntel: Explainable Detection of Propagandistic and Hateful Memes 23 Feb 2025 · 0 repositories · arXiv:2502.16612
-
Optimizing Retrieval-Augmented Generation of Medical Content for Spaced Repetition Learning 23 Feb 2025 · 0 repositories · arXiv:2503.01859
-
PMAT: Optimizing Action Generation Order in Multi-Agent Reinforcement Learning 23 Feb 2025 · 1 repository · arXiv:2502.16496
-
Reasoning about Affordances: Causal and Compositional Reasoning in LLMs 23 Feb 2025 · 0 repositories · arXiv:2502.16606
-
Reasoning About Persuasion: Can LLMs Enable Explainable Propaganda Detection? 23 Feb 2025 · 0 repositories · arXiv:2502.16550
-
Recent Advances in Large Langauge Model Benchmarks against Data Contamination: From Static to Dynamic Evaluation 23 Feb 2025 · 1 repository · arXiv:2502.17521
-
Retrieval-Augmented Visual Question Answering via Built-in Autoregressive Search Engines 23 Feb 2025 · 0 repositories · arXiv:2502.16641
-
Sensing-Assisted Channel Estimation for OFDM ISAC Systems: Framework, Algorithm, and Analysis 23 Feb 2025 · 0 repositories · arXiv:2502.16436
-
UniDyG: A Unified and Effective Representation Learning Approach for Large Dynamic Graphs 23 Feb 2025 · 0 repositories · arXiv:2502.16431
-
Visual-RAG: Benchmarking Text-to-Image Retrieval Augmented Generation for Visual Knowledge Intensive Queries 23 Feb 2025 · 1 repository · arXiv:2502.16636
-
VPNeXt -- Rethinking Dense Decoding for Plain Vision Transformer 23 Feb 2025 · 0 repositories · arXiv:2502.16654
-
An End-to-End Homomorphically Encrypted Neural Network 22 Feb 2025 · 0 repositories · arXiv:2502.16176
-
Be a Multitude to Itself: A Prompt Evolution Framework for Red Teaming 22 Feb 2025 · 0 repositories · arXiv:2502.16109
-
Concept Corrector: Erase concepts on the fly for text-to-image diffusion models 22 Feb 2025 · 0 repositories · arXiv:2502.16368
-
Demand Forecasting for Electric Vehicle Charging Stations using Multivariate Time-Series Analysis 22 Feb 2025 · 0 repositories · arXiv:2502.16365
-
Destroy and Repair Using Hyper Graphs for Routing 22 Feb 2025 · 1 repository · arXiv:2502.16170Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Enhancing LLMs for Identifying and Prioritizing Important Medical Jargons from Electronic Health Record Notes Utilizing Data Augmentation 22 Feb 2025 · 0 repositories · arXiv:2502.16022
-
FHGE: A Fast Heterogeneous Graph Embedding with Ad-hoc Meta-paths 22 Feb 2025 · 0 repositories · arXiv:2502.16281
-
Iterative Auto-Annotation for Scientific Named Entity Recognition Using BERT-Based Models 22 Feb 2025 · 0 repositories · arXiv:2502.16312
-
Joint Similarity Item Exploration and Overlapped User Guidance for Multi-Modal Cross-Domain Recommendation 22 Feb 2025 · 0 repositories · arXiv:2502.16068
-
Linear Attention for Efficient Bidirectional Sequence Modeling 22 Feb 2025 · 1 repository · arXiv:2502.16249Syntology official (archive's flag): 14 ran · 14 ran (of which 1 constructed an object rather than computing a result; 4 with no instrument failure: 3 honoured, 0 violated, 1 with no contract checked; 10 where Syntology's instrument failed) · 0 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
RAG-Enhanced Collaborative LLM Agents for Drug Discovery 22 Feb 2025 · 0 repositories · arXiv:2502.17506
-
Robustness and Cybersecurity in the EU Artificial Intelligence Act 22 Feb 2025 · 0 repositories · arXiv:2502.16184
-
SAE-V: Interpreting Multimodal Models for Enhanced Alignment 22 Feb 2025 · 0 repositories · arXiv:2502.17514
-
SalM2: An Extremely Lightweight Saliency Mamba Model for Real-Time Cognitive Awareness of Driver Attention 22 Feb 2025 · 1 repository · arXiv:2502.16214
-
Single Domain Generalization with Model-aware Parametric Batch-wise Mixup 22 Feb 2025 · 0 repositories · arXiv:2502.16064
-
Uncertainty-Aware Fusion: An Ensemble Framework for Mitigating Hallucinations in Large Language Models 22 Feb 2025 · 0 repositories · arXiv:2503.05757
-
Vision Transformer Accelerator ASIC for Real-Time, Low-Power Sleep Staging 22 Feb 2025 · 0 repositories · arXiv:2502.16334
-
Worse than Zero-shot? A Fact-Checking Dataset for Evaluating the Robustness of RAG Against Misleading Retrievals 22 Feb 2025 · 0 repositories · arXiv:2502.16101
-
A Close Look at Decomposition-based XAI-Methods for Transformer Language Models 21 Feb 2025 · 2 repositories · arXiv:2502.15886
-
Attention Eclipse: Manipulating Attention to Bypass LLM Safety-Alignment 21 Feb 2025 · 0 repositories · arXiv:2502.15334
-
AttentionEngine: A Versatile Framework for Efficient Attention Mechanisms on Diverse Hardware Platforms 21 Feb 2025 · 1 repository · arXiv:2502.15349
-
Auto-Bench: An Automated Benchmark for Scientific Discovery in LLMs 21 Feb 2025 · 0 repositories · arXiv:2502.15224
-
AutoMedPrompt: A New Framework for Optimizing LLM Medical Prompts Using Textual Gradients 21 Feb 2025 · 0 repositories · arXiv:2502.15944
-
BP-GPT: Auditory Neural Decoding Using fMRI-prompted LLM 21 Feb 2025 · 1 repository · arXiv:2502.15172Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Chain-of-Rank: Enhancing Large Language Models for Domain-Specific RAG in Edge Device 21 Feb 2025 · 0 repositories · arXiv:2502.15134
-
CoKV: Optimizing KV Cache Allocation via Cooperative Game 21 Feb 2025 · 1 repository · arXiv:2502.17501
-
Comparative Analysis of Large Language Models for Context-Aware Code Completion using SAFIM Framework 21 Feb 2025 · 0 repositories · arXiv:2502.15243
-
Compression Barriers for Autoregressive Transformers 21 Feb 2025 · 0 repositories · arXiv:2502.15955
-
Connecting the geometry and dynamics of many-body complex systems with message passing neural operators 21 Feb 2025 · 0 repositories · arXiv:2502.15913
-
Corrections Meet Explanations: A Unified Framework for Explainable Grammatical Error Correction 21 Feb 2025 · 0 repositories · arXiv:2502.15261
-
CoT-ICL Lab: A Petri Dish for Studying Chain-of-Thought Learning from In-Context Demonstrations 21 Feb 2025 · 1 repository · arXiv:2502.15132
-
Cross-Format Retrieval-Augmented Generation in XR with LLMs for Context-Aware Maintenance Assistance 21 Feb 2025 · 0 repositories · arXiv:2502.15604
-
Depth-aware Fusion Method based on Image and 4D Radar Spectrum for 3D Object Detection 21 Feb 2025 · 0 repositories · arXiv:2502.15516
-
DOEI: Dual Optimization of Embedding Information for Attention-Enhanced Class Activation Maps 21 Feb 2025 · 1 repository · arXiv:2502.15885
-
Empowering LLMs with Logical Reasoning: A Comprehensive Survey 21 Feb 2025 · 0 repositories · arXiv:2502.15652
-
Enhancing Domain-Specific Retrieval-Augmented Generation: Synthetic Data Generation and Evaluation using Reasoning Models 21 Feb 2025 · 1 repository · arXiv:2502.15854
-
Enhancing RWKV-based Language Models for Long-Sequence Text Generation 21 Feb 2025 · 1 repository · arXiv:2502.15485
-
Enhancing Vehicle Make and Model Recognition with 3D Attention Modules 21 Feb 2025 · 0 repositories · arXiv:2502.15398
-
Exploring Embodied Multimodal Large Models: Development, Datasets, and Future Directions 21 Feb 2025 · 0 repositories · arXiv:2502.15336
-
Extraction multi-étiquettes de relations en utilisant des couches de Transformer 21 Feb 2025 · 0 repositories · arXiv:2502.15619