Methods › General › Attention Mechanisms › Attention › Papers, page 43
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 43 of 316: papers 4,201 to 4,300 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
MEADOW: Memory-efficient Dataflow and Data Packing for Low Power Edge LLMs 14 Feb 2025 · 0 repositories · arXiv:2503.11663
-
Modern Hopfield Networks with Continuous-Time Memories 14 Feb 2025 · 0 repositories · arXiv:2502.10122
-
Named entity recognition for Serbian legal documents: Design, methodology and dataset development 14 Feb 2025 · 0 repositories · arXiv:2502.10582
-
Post-training an LLM for RAG? Train on Self-Generated Demonstrations 14 Feb 2025 · 0 repositories · arXiv:2502.10596
-
PromptArtisan: Multi-instruction Image Editing in Single Pass with Complete Attention Control 14 Feb 2025 · 0 repositories · arXiv:2502.10258
-
QMaxViT-Unet+: A Query-Based MaxViT-Unet with Edge Enhancement for Scribble-Supervised Segmentation of Medical Images 14 Feb 2025 · 1 repository · arXiv:2502.10294
-
Simplifying DINO via Coding Rate Regularization 14 Feb 2025 · 0 repositories · arXiv:2502.10385
-
STAR: Spectral Truncation and Rescale for Model Merging 14 Feb 2025 · 1 repository · arXiv:2502.10339Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model 14 Feb 2025 · 3 repositories · arXiv:2502.10248Syntology official (archive's flag): 1 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
A Contextual-Aware Position Encoding for Sequential Recommendation 13 Feb 2025 · 1 repository · arXiv:2502.09027
-
A Hybrid Transformer Model for Fake News Detection: Leveraging Bayesian Optimization and Bidirectional Recurrent Unit 13 Feb 2025 · 0 repositories · arXiv:2502.09097
-
A Physics-Informed Deep Learning Model for MRI Brain Motion Correction 13 Feb 2025 · 1 repository · arXiv:2502.09296
-
AnomalyGFM: Graph Foundation Model for Zero/Few-shot Anomaly Detection 13 Feb 2025 · 1 repository · arXiv:2502.09254
-
Application of Tabular Transformer Architectures for Operating System Fingerprinting 13 Feb 2025 · 1 repository · arXiv:2502.09084
-
AttentionSmithy: A Modular Framework for Rapid Transformer Development and Customization 13 Feb 2025 · 0 repositories · arXiv:2502.09503
-
Biologically Plausible Brain Graph Transformer 13 Feb 2025 · 1 repository · arXiv:2502.08958Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Can Uniform Meaning Representation Help GPT-4 Translate from Indigenous Languages? 13 Feb 2025 · 0 repositories · arXiv:2502.08900
-
Channel Dependence, Limited Lookback Windows, and the Simplicity of Datasets: How Biased is Time Series Forecasting? 13 Feb 2025 · 0 repositories · arXiv:2502.09683
-
CLEAR: Cluster-based Prompt Learning on Heterogeneous Graphs 13 Feb 2025 · 0 repositories · arXiv:2502.08918
-
Diverse Transformer Decoding for Offline Reinforcement Learning Using Financial Algorithmic Approaches 13 Feb 2025 · 0 repositories · arXiv:2502.10473
-
DynSegNet:Dynamic Architecture Adjustment for Adversarial Learning in Segmenting Hemorrhagic Lesions from Fundus Images 13 Feb 2025 · 0 repositories · arXiv:2502.09256
-
E-MD3C: Taming Masked Diffusion Transformers for Efficient Zero-Shot Object Customization 13 Feb 2025 · 0 repositories · arXiv:2502.09164
-
Enhancing RAG with Active Learning on Conversation Records: Reject Incapables and Answer Capables 13 Feb 2025 · 0 repositories · arXiv:2502.09073
-
FARM: Frequency-Aware Model for Cross-Domain Live-Streaming Recommendation 13 Feb 2025 · 0 repositories · arXiv:2502.09375
-
Feature-based Graph Attention Networks Improve Online Continual Learning 13 Feb 2025 · 0 repositories · arXiv:2502.09143
-
Hierarchical Vision Transformer with Prototypes for Interpretable Medical Image Classification 13 Feb 2025 · 0 repositories · arXiv:2502.08997
-
Human-LLM Coevolution: Evidence from Academic Writing 13 Feb 2025 · 0 repositories · arXiv:2502.09606
-
Improving TCM Question Answering through Tree-Organized Self-Reflective Retrieval with LLMs 13 Feb 2025 · 0 repositories · arXiv:2502.09156
-
InfiniteHiP: Extending Language Model Context Up to 3 Million Tokens on a Single GPU 13 Feb 2025 · 0 repositories · arXiv:2502.08910
-
Joint Attention Mechanism Learning to Facilitate Opto-physiological Monitoring during Physical Activity 13 Feb 2025 · 0 repositories · arXiv:2502.09291
-
KIMAs: A Configurable Knowledge Integrated Multi-Agent System 13 Feb 2025 · 0 repositories · arXiv:2502.09596
-
Linear-Time User-Level DP-SCO via Robust Statistics 13 Feb 2025 · 0 repositories · arXiv:2502.08889
-
LLM-Enhanced Multiple Instance Learning for Joint Rumor and Stance Detection with Social Context Information 13 Feb 2025 · 0 repositories · arXiv:2502.08888
-
Long-Term TalkingFace Generation via Motion-Prior Conditional Diffusion Model 13 Feb 2025 · 0 repositories · arXiv:2502.09533
-
Machine learning for modelling unstructured grid data in computational physics: a review 13 Feb 2025 · 0 repositories · arXiv:2502.09346
-
MC2SleepNet: Multi-modal Cross-masking with Contrastive Learning for Sleep Stage Classification 13 Feb 2025 · 1 repository · arXiv:2502.17470
-
Mechanistic Unveiling of Transformer Circuits: Self-Influence as a Key to Model Reasoning 13 Feb 2025 · 0 repositories · arXiv:2502.09022
-
Predicting Cognitive Decline: A Multimodal AI Approach to Dementia Screening from Speech 13 Feb 2025 · 0 repositories · arXiv:2502.08862
-
RefineCoder: Iterative Improving of Large Language Models via Adaptive Critique Refinement for Code Generation 13 Feb 2025 · 0 repositories · arXiv:2502.09183
-
Residual Transformer Fusion Network for Salt and Pepper Image Denoising 13 Feb 2025 · 0 repositories · arXiv:2502.09000
-
Rethinking Evaluation Metrics for Grammatical Error Correction: Why Use a Different Evaluation Process than Human? 13 Feb 2025 · 1 repository · arXiv:2502.09416Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
RLSA-PFL: Robust Lightweight Secure Aggregation with Model Inconsistency Detection in Privacy-Preserving Federated Learning 13 Feb 2025 · 0 repositories · arXiv:2502.08989
-
Structured Convergence in Large Language Model Representations via Hierarchical Latent Space Folding 13 Feb 2025 · 0 repositories · arXiv:2502.08947
-
Can Vision-Language Models Infer Speaker's Ignorance? The Role of Visual and Linguistic Cues 13 Feb 2025 · 0 repositories · arXiv:2502.09120
-
The Joint Entity-Relation Extraction Model Based on Span and Interactive Fusion Representation for Chinese Medical Texts with Complex Semantics 13 Feb 2025 · 0 repositories · arXiv:2502.09247
-
The Stochastic Parrot on LLM's Shoulder: A Summative Assessment of Physical Concept Understanding 13 Feb 2025 · 1 repository · arXiv:2502.08946
-
Transformer-Enhanced Variational Autoencoder for Crystal Structure Prediction 13 Feb 2025 · 0 repositories · arXiv:2502.09423
-
Unleashing the Power of Large Language Model for Denoising Recommendation 13 Feb 2025 · 0 repositories · arXiv:2502.09058
-
Unlocking the Potential of Classic GNNs for Graph-level Tasks: Simple Architectures Meet Excellence 13 Feb 2025 · 1 repository · arXiv:2502.09263Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples) · 3 pointer-only (licence)
-
Utilizing Pre-trained and Large Language Models for 10-K Items Segmentation 13 Feb 2025 · 0 repositories · arXiv:2502.08875
-
Vertical Federated Continual Learning via Evolving Prototype Knowledge 13 Feb 2025 · 0 repositories · arXiv:2502.09152
-
What exactly has TabPFN learned to do? 13 Feb 2025 · 1 repository · arXiv:2502.08978
-
A comparative study of different TSO-DSO coordination in the reserve market 12 Feb 2025 · 0 repositories · arXiv:2502.08782
-
A Survey on Image Quality Assessment: Insights, Analysis, and Future Outlook 12 Feb 2025 · 0 repositories · arXiv:2502.08540
-
A Systematic Review on the Evaluation of Large Language Models in Theory of Mind Tasks 12 Feb 2025 · 1 repository · arXiv:2502.08796
-
Ask in Any Modality: A Comprehensive Survey on Multimodal Retrieval-Augmented Generation 12 Feb 2025 · 1 repository · arXiv:2502.08826
-
ChorusCVR: Chorus Supervision for Entire Space Post-Click Conversion Rate Modeling 12 Feb 2025 · 0 repositories · arXiv:2502.08277
-
TANTE: Time-Adaptive Operator Learning via Neural Taylor Expansion 12 Feb 2025 · 0 repositories · arXiv:2502.08574
-
Contextual Compression Encoding for Large Language Models: A Novel Framework for Multi-Layered Parameter Space Pruning 12 Feb 2025 · 0 repositories · arXiv:2502.08323
-
COutfitGAN: Learning to Synthesize Compatible Outfits Supervised by Silhouette Masks and Fashion Styles 12 Feb 2025 · 0 repositories · arXiv:2502.08674
-
Enhanced Load Forecasting with GAT-LSTM: Leveraging Grid and Temporal Features 12 Feb 2025 · 1 repository · arXiv:2502.08376
-
Enhancing Auto-regressive Chain-of-Thought through Loop-Aligned Reasoning 12 Feb 2025 · 0 repositories · arXiv:2502.08482
-
Fino1: On the Transferability of Reasoning Enhanced LLMs to Finance 12 Feb 2025 · 1 repository · arXiv:2502.08127Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Generalized Class Discovery in Instance Segmentation 12 Feb 2025 · 0 repositories · arXiv:2502.08149
-
HDT: Hierarchical Discrete Transformer for Multivariate Time Series Forecasting 12 Feb 2025 · 1 repository · arXiv:2502.08302
-
Hi-End-MAE: Hierarchical encoder-driven masked autoencoders are stronger vision learners for medical image segmentation 12 Feb 2025 · 1 repository · arXiv:2502.08347
-
InTAR: Inter-Task Auto-Reconfigurable Accelerator Design for High Data Volume Variation in DNNs 12 Feb 2025 · 1 repository · arXiv:2502.08807
-
Joint Transmit and Pinching Beamforming for Pinching Antenna Systems (PASS): Optimization-Based or Learning-Based? 12 Feb 2025 · 0 repositories · arXiv:2502.08637
-
Light-A-Video: Training-free Video Relighting via Progressive Light Fusion 12 Feb 2025 · 1 repository · arXiv:2502.08590Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Mapping the Landscape of Generative AI in Network Monitoring and Management 12 Feb 2025 · 0 repositories · arXiv:2502.08576
-
mmE5: Improving Multimodal Multilingual Embeddings via High-quality Synthetic Data 12 Feb 2025 · 1 repository · arXiv:2502.08468
-
On Mechanistic Circuits for Extractive Question-Answering 12 Feb 2025 · 0 repositories · arXiv:2502.08059
-
ParetoRAG: Leveraging Sentence-Context Attention for Robust and Efficient Retrieval-Augmented Generation 12 Feb 2025 · 0 repositories · arXiv:2502.08178
-
Representation Learning to Advance Multi-institutional Studies with Electronic Health Record Data 12 Feb 2025 · 0 repositories · arXiv:2502.08547
-
Rethinking Tokenized Graph Transformers for Node Classification 12 Feb 2025 · 0 repositories · arXiv:2502.08101
-
Scalable Thermodynamic Second-order Optimization 12 Feb 2025 · 0 repositories · arXiv:2502.08603
-
Self-Evaluation for Job-Shop Scheduling 12 Feb 2025 · 0 repositories · arXiv:2502.08684
-
SelfElicit: Your Language Model Secretly Knows Where is the Relevant Evidence 12 Feb 2025 · 1 repository · arXiv:2502.08767
-
Systematic Knowledge Injection into Large Language Models via Diverse Augmentation for Domain-Specific RAG 12 Feb 2025 · 1 repository · arXiv:2502.08356Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
TLOB: A Novel Transformer Model with Dual Attention for Price Trend Prediction with Limit Order Book Data 12 Feb 2025 · 1 repository · arXiv:2502.15757Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Top-Theta Attention: Sparsifying Transformers by Compensated Thresholding 12 Feb 2025 · 1 repository · arXiv:2502.08363
-
YNote: A Novel Music Notation for Fine-Tuning LLMs in Music Generation 12 Feb 2025 · 0 repositories · arXiv:2502.10467
-
5D Neural Surrogates for Nonlinear Gyrokinetic Simulations of Plasma Turbulence 11 Feb 2025 · 0 repositories · arXiv:2502.07469
-
A Large-Scale Benchmark for Vietnamese Sentence Paraphrases 11 Feb 2025 · 1 repository · arXiv:2502.07188
-
A Survey on Mamba Architecture for Vision Applications 11 Feb 2025 · 0 repositories · arXiv:2502.07161
-
An Advanced NLP Framework for Automated Medical Diagnosis with DeBERTa and Dynamic Contextual Positional Gating 11 Feb 2025 · 0 repositories · arXiv:2502.07755
-
Attention Learning is Needed to Efficiently Learn Parity Function 11 Feb 2025 · 0 repositories · arXiv:2502.07553
-
Auditing Prompt Caching in Language Model APIs 11 Feb 2025 · 1 repository · arXiv:2502.07776Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Automated Capability Discovery via Model Self-Exploration 11 Feb 2025 · 2 repositories · arXiv:2502.07577
-
Streaming Attention Approximation via Discrepancy Theory 11 Feb 2025 · 0 repositories · arXiv:2502.07861
-
Breaking Down Bias: On The Limits of Generalizable Pruning Strategies 11 Feb 2025 · 0 repositories · arXiv:2502.07771
-
CausalGeD: Blending Causality and Diffusion for Spatial Gene Expression Generation 11 Feb 2025 · 0 repositories · arXiv:2502.07751
-
CodePhys: Robust Video-based Remote Physiological Measurement through Latent Codebook Querying 11 Feb 2025 · 0 repositories · arXiv:2502.07526
-
Dataset Ownership Verification in Contrastive Pre-trained Models 11 Feb 2025 · 1 repository · arXiv:2502.07276
-
Deep Semantic Graph Learning via LLM based Node Enhancement 11 Feb 2025 · 0 repositories · arXiv:2502.07982
-
Dense Object Detection Based on De-homogenized Queries 11 Feb 2025 · 0 repositories · arXiv:2502.07194
-
DSV: Exploiting Dynamic Sparsity to Accelerate Large-Scale Video DiT Training 11 Feb 2025 · 0 repositories · arXiv:2502.07590
-
Enhance-A-Video: Better Generated Video for Free 11 Feb 2025 · 1 repository · arXiv:2502.07508
-
Enviro-IoT: Calibrating Low-Cost Environmental Sensors in Urban Settings 11 Feb 2025 · 0 repositories · arXiv:2502.07596
-
Fast-COS: A Fast One-Stage Object Detector Based on Reparameterized Attention Vision Transformer for Autonomous Driving 11 Feb 2025 · 0 repositories · arXiv:2502.07417