Methods › General › Attention Mechanisms › Attention › Papers, page 77
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 77 of 316: papers 7,601 to 7,700 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
ESC-MISR: Enhancing Spatial Correlations for Multi-Image Super-Resolution in Remote Sensing 7 Nov 2024 · 0 repositories · arXiv:2411.04706
-
Financial Fraud Detection using Jump-Attentive Graph Neural Networks 7 Nov 2024 · 1 repository · arXiv:2411.05857
-
FineTuneBench: How well do commercial fine-tuning APIs infuse knowledge into LLMs? 7 Nov 2024 · 1 repository · arXiv:2411.05059
-
Generating Highly Designable Proteins with Geometric Algebra Flow Matching 7 Nov 2024 · 1 repository · arXiv:2411.05238Syntology official (archive's flag): 6 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
GPT-Guided Monte Carlo Tree Search for Symbolic Regression in Financial Fraud Detection 7 Nov 2024 · 0 repositories · arXiv:2411.04459
-
HourVideo: 1-Hour Video-Language Understanding 7 Nov 2024 · 1 repository · arXiv:2411.04998Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 16 harvested samples) · 1 pointer-only (licence)
-
Hypercube Policy Regularization Framework for Offline Reinforcement Learning 7 Nov 2024 · 1 repository · arXiv:2411.04534
-
LLM2CLIP: Powerful Language Model Unlocks Richer Visual Representation 7 Nov 2024 · 1 repository · arXiv:2411.04997
-
LoFi: Neural Local Fields for Scalable Image Reconstruction 7 Nov 2024 · 1 repository · arXiv:2411.04995
-
M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding 7 Nov 2024 · 0 repositories · arXiv:2411.04952
-
Measure-to-measure interpolation using Transformers 7 Nov 2024 · 0 repositories · arXiv:2411.04551
-
Measuring short-form factuality in large language models 7 Nov 2024 · 1 repository · arXiv:2411.04368
-
Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models 7 Nov 2024 · 1 repository · arXiv:2411.04996Syntology 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples) · 2 pointer-only (licence)
-
Multi-temporal crack segmentation in concrete structure using deep learning approaches 7 Nov 2024 · 0 repositories · arXiv:2411.04620
-
Peri-midFormer: Periodic Pyramid Transformer for Time Series Analysis 7 Nov 2024 · 1 repository · arXiv:2411.04554Syntology official (archive's flag): 10 ran · 10 ran (of which 9 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Pose2Trajectory: Using Transformers on Body Pose to Predict Tennis Player's Trajectory 7 Nov 2024 · 0 repositories · arXiv:2411.04501
-
Private Algorithms for Stochastic Saddle Points and Variational Inequalities: Beyond Euclidean Geometry 7 Nov 2024 · 0 repositories · arXiv:2411.05198
-
ProverbEval: Exploring LLM Evaluation Challenges for Low-resource Language Understanding 7 Nov 2024 · 0 repositories · arXiv:2411.05049
-
Pruning Literals for Highly Efficient Explainability at Word Level 7 Nov 2024 · 0 repositories · arXiv:2411.04557
-
RetrieveGPT: Merging Prompts and Mathematical Models for Enhanced Code-Mixed Information Retrieval 7 Nov 2024 · 0 repositories · arXiv:2411.04752
-
SaSR-Net: Source-Aware Semantic Representation Network for Enhancing Audio-Visual Question Answering 7 Nov 2024 · 0 repositories · arXiv:2411.04933
-
Selecting Between BERT and GPT for Text Classification in Political Science Research 7 Nov 2024 · 0 repositories · arXiv:2411.05050
-
STAND-Guard: A Small Task-Adaptive Content Moderation Model 7 Nov 2024 · 0 repositories · arXiv:2411.05214
-
Synergy-Guided Regional Supervision of Pseudo Labels for Semi-Supervised Medical Image Segmentation 7 Nov 2024 · 0 repositories · arXiv:2411.04493
-
Toward Cultural Interpretability: A Linguistic Anthropological Framework for Describing and Evaluating Large Language Models (LLMs) 7 Nov 2024 · 0 repositories · arXiv:2411.05200
-
Towards Competitive Search Relevance For Inference-Free Learned Sparse Retrievers 7 Nov 2024 · 1 repository · arXiv:2411.04403
-
Unlearning in- vs. out-of-distribution data in LLMs under gradient-based method 7 Nov 2024 · 0 repositories · arXiv:2411.04388
-
Unveiling Placental Development in Circadian Rhythm-Disrupted Mice: A Photo-acoustic Imaging Study on Unstained Tissue 7 Nov 2024 · 0 repositories · arXiv:2411.04694
-
Words that Move Markets- Quantifying the Impact of RBI's Monetary Policy Communications on Indian Financial Market 7 Nov 2024 · 0 repositories · arXiv:2411.04808
-
A Comparative Study of Recent Large Language Models on Generating Hospital Discharge Summaries for Lung Cancer Patients 6 Nov 2024 · 0 repositories · arXiv:2411.03805
-
A Contrastive Self-Supervised Learning scheme for beat tracking amenable to few-shot learning 6 Nov 2024 · 0 repositories · arXiv:2411.04152
-
A Multilingual Sentiment Lexicon for Low-Resource Language Translation using Large Languages Models and Explainable AI 6 Nov 2024 · 0 repositories · arXiv:2411.04316
-
Advanced RAG Models with Graph Structures: Optimizing Complex Knowledge Reasoning and Text Generation 6 Nov 2024 · 0 repositories · arXiv:2411.03572
-
Analyzing Multimodal Features of Spontaneous Voice Assistant Commands for Mild Cognitive Impairment Detection 6 Nov 2024 · 0 repositories · arXiv:2411.04158
-
Bio-xLSTM: Generative modeling, representation and in-context learning of biological and chemical sequences 6 Nov 2024 · 3 repositories · arXiv:2411.04165Syntology official (archive's flag): 27 ran · 27 ran (of which 0 constructed an object rather than computing a result; 25 with no instrument failure: 0 honoured, 1 violated, 24 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 29 harvested samples) · 1 pointer-only (licence)
-
Can Custom Models Learn In-Context? An Exploration of Hybrid Architecture Performance on In-Context Learning Tasks 6 Nov 2024 · 1 repository · arXiv:2411.03945
-
Can Graph Neural Networks Expose Training Data Properties? An Efficient Risk Assessment Approach 6 Nov 2024 · 1 repository · arXiv:2411.03663Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Customized Multiple Clustering via Multi-Modal Subspace Proxy Learning 6 Nov 2024 · 1 repository · arXiv:2411.03978Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Diversity Helps Jailbreak Large Language Models 6 Nov 2024 · 0 repositories · arXiv:2411.04223
-
Fine-Grained Guidance for Retrievers: Leveraging LLMs' Feedback in Retrieval-Augmented Generation 6 Nov 2024 · 0 repositories · arXiv:2411.03957
-
From Medprompt to o1: Exploration of Run-Time Strategies for Medical Challenge Problems and Beyond 6 Nov 2024 · 0 repositories · arXiv:2411.03590
-
From Word Vectors to Multimodal Embeddings: Techniques, Applications, and Future Directions For Large Language Models 6 Nov 2024 · 0 repositories · arXiv:2411.05036
-
Generalized Trusted Multi-view Classification Framework with Hierarchical Opinion Aggregation 6 Nov 2024 · 1 repository · arXiv:2411.03713
-
GS2Pose: Two-stage 6D Object Pose Estimation Guided by Gaussian Splatting 6 Nov 2024 · 0 repositories · arXiv:2411.03807
-
How Transformers Solve Propositional Logic Problems: A Mechanistic Analysis 6 Nov 2024 · 0 repositories · arXiv:2411.04105
-
Hybrid Attention for Robust RGB-T Pedestrian Detection in Real-World Conditions 6 Nov 2024 · 0 repositories · arXiv:2411.03576
-
kNN Attention Demystified: A Theoretical Exploration for Scalable Transformers 6 Nov 2024 · 1 repository · arXiv:2411.04013
-
Learning Generalizable Policy for Obstacle-Aware Autonomous Drone Racing 6 Nov 2024 · 1 repository · arXiv:2411.04246
-
MambaPEFT: Exploring Parameter-Efficient Fine-Tuning for Mamba 6 Nov 2024 · 1 repository · arXiv:2411.03855Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples)
-
Number Cookbook: Number Understanding of Language Models and How to Improve It 6 Nov 2024 · 1 repository · arXiv:2411.03766Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
OccLoff: Learning Optimized Feature Fusion for 3D Occupancy Prediction 6 Nov 2024 · 0 repositories · arXiv:2411.03696
-
On-Device Emoji Classifier Trained with GPT-based Data Augmentation for a Mobile Keyboard 6 Nov 2024 · 0 repositories · arXiv:2411.05031
-
PhDGPT: Introducing a psychometric and linguistic dataset about how large language models perceive graduate students and professors in psychology 6 Nov 2024 · 0 repositories · arXiv:2411.10473
-
Prion-ViT: Prions-Inspired Vision Transformers for Temperature prediction with Specklegrams 6 Nov 2024 · 0 repositories · arXiv:2411.05836
-
Prompt Engineering Using GPT for Word-Level Code-Mixed Language Identification in Low-Resource Dravidian Languages 6 Nov 2024 · 0 repositories · arXiv:2411.04025
-
RAGulator: Lightweight Out-of-Context Detectors for Grounded Text Generation 6 Nov 2024 · 0 repositories · arXiv:2411.03920
-
Reducing catastrophic forgetting of incremental learning in the absence of rehearsal memory with task-specific token 6 Nov 2024 · 0 repositories · arXiv:2411.05846
-
♠ SPADE ♠ Split Peak Attention DEcomposition 6 Nov 2024 · 0 repositories · arXiv:2411.05852
-
These Maps Are Made by Propagation: Adapting Deep Stereo Networks to Road Scenarios with Decisive Disparity Diffusion 6 Nov 2024 · 0 repositories · arXiv:2411.03717
-
Towards 3D Semantic Scene Completion for Autonomous Driving: A Meta-Learning Framework Empowered by Deformable Large-Kernel Attention and Mamba Model 6 Nov 2024 · 0 repositories · arXiv:2411.03672
-
Towards Interpreting Language Models: A Case Study in Multi-Hop Reasoning 6 Nov 2024 · 1 repository · arXiv:2411.05037Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Understanding the Effects of Human-written Paraphrases in LLM-generated Text Detection 6 Nov 2024 · 1 repository · arXiv:2411.03806
-
YouTube Comments Decoded: Leveraging LLMs for Low Resource Language Classification 6 Nov 2024 · 0 repositories · arXiv:2411.05039
-
A Convex Relaxation Approach to Generalization Analysis for Parallel Positively Homogeneous Networks 5 Nov 2024 · 0 repositories · arXiv:2411.02767
-
A Deep-Based Approach for Multi-Descriptor Feature Extraction: Applications on SAR Image Registration 5 Nov 2024 · 1 repository
-
A Mamba Foundation Model for Time Series Forecasting 5 Nov 2024 · 0 repositories · arXiv:2411.02941
-
AtlasSeg: Atlas Prior Guided Dual-U-Net for Cortical Segmentation in Fetal Brain MRI 5 Nov 2024 · 0 repositories · arXiv:2411.02867
-
Automatic Generation of Question Hints for Mathematics Problems using Large Language Models in Educational Technology 5 Nov 2024 · 0 repositories · arXiv:2411.03495
-
Curiosity-Driven Science: The in Situ Jungle Biomechanics Lab in the Amazon Rainforest 5 Nov 2024 · 0 repositories · arXiv:2411.03498
-
Decoupling Fine Detail and Global Geometry for Compressed Depth Map Super-Resolution 5 Nov 2024 · 1 repository · arXiv:2411.03239
-
DiT4Edit: Diffusion Transformer for Image Editing 5 Nov 2024 · 0 repositories · arXiv:2411.03286
-
Efficient Feature Aggregation and Scale-Aware Regression for Monocular 3D Object Detection 5 Nov 2024 · 1 repository · arXiv:2411.02747
-
Enhanced Real-Time Threat Detection in 5G Networks: A Self-Attention RNN Autoencoder Approach for Spectral Intrusion Analysis 5 Nov 2024 · 0 repositories · arXiv:2411.03365
-
Enhancing Transformer Training Efficiency with Dynamic Dropout 5 Nov 2024 · 0 repositories · arXiv:2411.03236
-
Exploring the Benefits of Domain-Pretraining of Generative Large Language Models for Chemistry 5 Nov 2024 · 0 repositories · arXiv:2411.03542
-
Exploring the Potentials and Challenges of Using Large Language Models for the Analysis of Transcriptional Regulation of Long Non-coding RNAs 5 Nov 2024 · 0 repositories · arXiv:2411.03522
-
Foundation AI Model for Medical Image Segmentation 5 Nov 2024 · 0 repositories · arXiv:2411.02745
-
From Pixels to Prose: Advancing Multi-Modal Language Models for Remote Sensing 5 Nov 2024 · 0 repositories · arXiv:2411.05826
-
HtmlRAG: HTML is Better Than Plain Text for Modeling Retrieved Knowledge in RAG Systems 5 Nov 2024 · 1 repository · arXiv:2411.02959
-
Kernel Approximation using Analog In-Memory Computing 5 Nov 2024 · 1 repository · arXiv:2411.03375
-
LASER: Attention with Exponential Transformation 5 Nov 2024 · 0 repositories · arXiv:2411.03493
-
Leveraging Transfer Learning and Multiple Instance Learning for HER2 Automatic Scoring of H&E Whole Slide Images 5 Nov 2024 · 1 repository · arXiv:2411.05028
-
LiVOS: Light Video Object Segmentation with Gated Linear Matching 5 Nov 2024 · 1 repository · arXiv:2411.02818
-
LLMs for Domain Generation Algorithm Detection 5 Nov 2024 · 0 repositories · arXiv:2411.03307
-
Long Context RAG Performance of Large Language Models 5 Nov 2024 · 0 repositories · arXiv:2411.03538
-
Mixtures of In-Context Learners 5 Nov 2024 · 0 repositories · arXiv:2411.02830
-
Neurons for Neutrons: A Transformer Model for Computation Load Estimation on Domain-Decomposed Neutron Transport Problems 5 Nov 2024 · 0 repositories · arXiv:2411.03389
-
On the Loss of Context-awareness in General Instruction Fine-tuning 5 Nov 2024 · 1 repository · arXiv:2411.02688Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
P-MOSS: Learned Scheduling For Indexes Over NUMA Servers Using Low-Level Hardware Statistics 5 Nov 2024 · 0 repositories · arXiv:2411.02933
-
PersianRAG: A Retrieval-Augmented Generation System for Persian Language 5 Nov 2024 · 0 repositories · arXiv:2411.02832
-
Predictor-Corrector Enhanced Transformers with Exponential Moving Average Coefficient Learning 5 Nov 2024 · 0 repositories · arXiv:2411.03042
-
Quantifying Aleatoric Uncertainty of the Treatment Effect: A Novel Orthogonal Learner 5 Nov 2024 · 1 repository · arXiv:2411.03387Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Query-Efficient Adversarial Attack Against Vertical Federated Graph Learning 5 Nov 2024 · 1 repository · arXiv:2411.02809
-
Rethinking Decoders for Transformer-based Semantic Segmentation: A Compression Perspective 5 Nov 2024 · 1 repository · arXiv:2411.03033
-
Sparse Orthogonal Parameters Tuning for Continual Learning 5 Nov 2024 · 0 repositories · arXiv:2411.02813
-
TDDBench: A Benchmark for Training data detection 5 Nov 2024 · 0 repositories · arXiv:2411.03363Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
The Evolution of RWKV: Advancements in Efficient Language Modeling 5 Nov 2024 · 0 repositories · arXiv:2411.02795
-
TIP-I2V: A Million-Scale Real Text and Image Prompt Dataset for Image-to-Video Generation 5 Nov 2024 · 0 repositories · arXiv:2411.04709
-
TokenSelect: Efficient Long-Context Inference and Length Extrapolation for LLMs via Dynamic Token-Level KV Cache Selection 5 Nov 2024 · 0 repositories · arXiv:2411.02886
-
TopoTxR: A topology-guided deep convolutional network for breast parenchyma learning on DCE-MRIs 5 Nov 2024 · 1 repository · arXiv:2411.03464