Methods › General › Attention Mechanisms › Attention › Papers, page 121
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 121 of 316: papers 12,001 to 12,100 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Evaluating Large Language Models with Grid-Based Game Competitions: An Extensible LLM Benchmark and Leaderboard 10 Jul 2024 · 1 repository · arXiv:2407.07796
-
FACTS About Building Retrieval Augmented Generation-based Chatbots 10 Jul 2024 · 0 repositories · arXiv:2407.07858
-
Real World Federated Learning with a Knowledge Distilled Transformer for Cardiac CT Imaging 10 Jul 2024 · 2 repositories · arXiv:2407.07557
-
FsPONER: Few-shot Prompt Optimization for Named Entity Recognition in Domain-specific Scenarios 10 Jul 2024 · 1 repository · arXiv:2407.08035
-
Fusion of Short-term and Long-term Attention for Video Mirror Detection 10 Jul 2024 · 1 repository · arXiv:2407.07999
-
Greit-HRNet: Grouped Lightweight High-Resolution Network for Human Pose Estimation 10 Jul 2024 · 0 repositories · arXiv:2407.07389
-
H-FCBFormer Hierarchical Fully Convolutional Branch Transformer for Occlusal Contact Segmentation with Articulating Paper 10 Jul 2024 · 1 repository · arXiv:2407.07604
-
HAFormer: Unleashing the Power of Hierarchy-Aware Features for Lightweight Semantic Segmentation 10 Jul 2024 · 0 repositories · arXiv:2407.07441
-
HDKD: Hybrid Data-Efficient Knowledge Distillation Network for Medical Image Classification 10 Jul 2024 · 1 repository · arXiv:2407.07516
-
iiANET: Inception Inspired Attention Hybrid Network for efficient Long-Range Dependency 10 Jul 2024 · 1 repository · arXiv:2407.07603
-
Industrial-Grade Time-Dependent Counterfactual Root Cause Analysis through the Unanticipated Point of Incipient Failure: a Proof of Concept 10 Jul 2024 · 1 repository · arXiv:2407.11056
-
Inference Performance Optimization for Large Language Models on CPUs 10 Jul 2024 · 1 repository · arXiv:2407.07304Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
KpopMT: Translation Dataset with Terminology for Kpop Fandom 10 Jul 2024 · 1 repository · arXiv:2407.07413
-
Large Language Model-Augmented Auto-Delineation of Treatment Target Volume in Radiation Therapy 10 Jul 2024 · 0 repositories · arXiv:2407.07296
-
Let Occ Flow: Self-Supervised 3D Occupancy Flow Prediction 10 Jul 2024 · 0 repositories · arXiv:2407.07587
-
LitSearch: A Retrieval Benchmark for Scientific Literature Search 10 Jul 2024 · 1 repository · arXiv:2407.18940Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Enabling Massive Index Modulation Systems via Combinatorics-Free Detection 10 Jul 2024 · 0 repositories · arXiv:2407.08087
-
Micro-Expression Recognition by Motion Feature Extraction based on Pre-training 10 Jul 2024 · 0 repositories · arXiv:2407.07345
-
A Guide To Effectively Leveraging LLMs for Low-Resource Text Summarization: Data Augmentation and Semi-supervised Approaches 10 Jul 2024 · 0 repositories · arXiv:2407.07341
-
Multilingual Blending: LLM Safety Alignment Evaluation with Language Mixture 10 Jul 2024 · 0 repositories · arXiv:2407.07342
-
PosFormer: Recognizing Complex Handwritten Mathematical Expression with Position Forest Transformer 10 Jul 2024 · 1 repository · arXiv:2407.07764Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Probability of Differentiation Reveals Brittleness of Homogeneity Bias in GPT-4 10 Jul 2024 · 0 repositories · arXiv:2407.07329
-
Examining Long-Context Large Language Models for Environmental Review Document Comprehension 10 Jul 2024 · 0 repositories · arXiv:2407.07321
-
Rectifier: Code Translation with Corrector via LLMs 10 Jul 2024 · 1 repository · arXiv:2407.07472Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Resource Allocation for Twin Maintenance and Computing Task Processing in Digital Twin Vehicular Edge Computing Network 10 Jul 2024 · 1 repository · arXiv:2407.07575
-
ROSA: Random Subspace Adaptation for Efficient Fine-Tuning 10 Jul 2024 · 1 repository · arXiv:2407.07802
-
RT-LA-VocE: Real-Time Low-SNR Audio-Visual Speech Enhancement 10 Jul 2024 · 0 repositories · arXiv:2407.07825
-
Search, Examine and Early-Termination: Fake News Detection with Annotation-Free Evidences 10 Jul 2024 · 0 repositories · arXiv:2407.07931
-
Source Tracing of Audio Deepfake Systems 10 Jul 2024 · 0 repositories · arXiv:2407.08016
-
Spatial-Temporal Attention Model for Traffic State Estimation with Sparse Internet of Vehicles 10 Jul 2024 · 0 repositories · arXiv:2407.08047
-
SPIN: SE(3)-Invariant Physics Informed Network for Binding Affinity Prediction 10 Jul 2024 · 0 repositories · arXiv:2407.11057
-
Study on Aspect Ratio Variability toward Robustness of Vision Transformer-based Vehicle Re-identification 10 Jul 2024 · 0 repositories · arXiv:2407.07842
-
Swin SMT: Global Sequential Modeling in 3D Medical Image Segmentation 10 Jul 2024 · 1 repository · arXiv:2407.07514
-
Swiss DINO: Efficient and Versatile Vision Framework for On-device Personal Object Search 10 Jul 2024 · 1 repository · arXiv:2407.07541
-
Teaching Transformers Causal Reasoning through Axiomatic Training 10 Jul 2024 · 0 repositories · arXiv:2407.07612
-
Disturbance-based Discretization, Differentiable IDS Channel, and an IDS-Correcting Code for DNA Storage 10 Jul 2024 · 0 repositories · arXiv:2407.18929
-
Toto: Time Series Optimized Transformer for Observability 10 Jul 2024 · 0 repositories · arXiv:2407.07874
-
Unified Embedding Alignment for Open-Vocabulary Video Instance Segmentation 10 Jul 2024 · 1 repository · arXiv:2407.07427
-
Unity in Diversity: Multi-expert Knowledge Confrontation and Collaboration for Generalizable Vehicle Re-identification 10 Jul 2024 · 0 repositories · arXiv:2407.07351
-
Video In-context Learning 10 Jul 2024 · 0 repositories · arXiv:2407.07356
-
When to Accept Automated Predictions and When to Defer to Human Judgment? 10 Jul 2024 · 0 repositories · arXiv:2407.07821
-
WorldAPIs: The World Is Worth How Many APIs? A Thought Experiment 10 Jul 2024 · 0 repositories · arXiv:2407.07778
-
A Predictive Model Based on Transformer with Statistical Feature Embedding in Manufacturing Sensor Dataset 9 Jul 2024 · 0 repositories · arXiv:2407.06682
-
A Simple Architecture for Enterprise Large Language Model Applications based on Role based security and Clearance Levels using Retrieval-Augmented Generation or Mixture of Experts 9 Jul 2024 · 0 repositories · arXiv:2407.06718
-
AI AI Bias: Large Language Models Favor Their Own Generated Content 9 Jul 2024 · 1 repository · arXiv:2407.12856
-
Aligning Cyber Space with Physical World: A Comprehensive Survey on Embodied AI 9 Jul 2024 · 1 repository · arXiv:2407.06886
-
AnatoMask: Enhancing Medical Image Segmentation with Reconstruction-guided Self-masking 9 Jul 2024 · 1 repository · arXiv:2407.06468Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Automated Peer Reviewing in Paper SEA: Standardization, Evaluation, and Analysis 9 Jul 2024 · 1 repository · arXiv:2407.12857Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 14 harvested samples)
-
AutoTask: Task Aware Multi-Faceted Single Model for Multi-Task Ads Relevance 9 Jul 2024 · 0 repositories · arXiv:2407.06549
-
CamFreeDiff: Camera-free Image to Panorama Generation with Diffusion Model 9 Jul 2024 · 0 repositories · arXiv:2407.07174
-
CAPformer: Compression-Aware Pre-trained Transformer for Low-Light Image Enhancement 9 Jul 2024 · 0 repositories · arXiv:2407.07056
-
ChatGPT Doesn't Trust Chargers Fans: Guardrail Sensitivity in Context 9 Jul 2024 · 1 repository · arXiv:2407.06866Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
ConvNLP: Image-based AI Text Detection 9 Jul 2024 · 0 repositories · arXiv:2407.07225
-
CorMulT: A Semi-supervised Modality Correlation-aware Multimodal Transformer for Sentiment Analysis 9 Jul 2024 · 0 repositories · arXiv:2407.07046
-
CTRL-F: Pairing Convolution with Transformer for Image Classification via Multi-Level Feature Cross-Attention and Representation Learning Fusion 9 Jul 2024 · 1 repository · arXiv:2407.06673
-
Decoding Climate Disagreement: A Graph Neural Network-Based Approach to Understanding Social Media Dynamics 9 Jul 2024 · 0 repositories · arXiv:2407.07038
-
Deep-Motion-Net: GNN-based volumetric organ shape reconstruction from single-view 2D projections 9 Jul 2024 · 0 repositories · arXiv:2407.06692
-
Empirical analysis of Binding Precedent efficiency in the Brazilian Supreme Court via Similar Case Retrieval 9 Jul 2024 · 0 repositories · arXiv:2407.07004
-
ERQ: Error Reduction for Post-Training Quantization of Vision Transformers 9 Jul 2024 · 0 repositories · arXiv:2407.06794
-
Event Trojan: Asynchronous Event-based Backdoor Attacks 9 Jul 2024 · 1 repository · arXiv:2407.06838Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Exploring Camera Encoder Designs for Autonomous Driving Perception 9 Jul 2024 · 0 repositories · arXiv:2407.07276
-
Fine-Tuning Attention Modules Only: Enhancing Weight Disentanglement in Task Arithmetic 9 Jul 2024 · 2 repositories · arXiv:2407.07089Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Graph-Based Captioning: Enhancing Visual Descriptions by Interconnecting Region Captions 9 Jul 2024 · 0 repositories · arXiv:2407.06723
-
Graph Neural Networks and Deep Reinforcement Learning Based Resource Allocation for V2X Communications 9 Jul 2024 · 1 repository · arXiv:2407.06518
-
Hypergraph based Understanding for Document Semantic Entity Recognition 9 Jul 2024 · 1 repository · arXiv:2407.06904Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Identification of emotions on Twitter during the 2022 electoral process in Colombia 9 Jul 2024 · 0 repositories · arXiv:2407.07258
-
Induction Heads as an Essential Mechanism for Pattern Matching in In-context Learning 9 Jul 2024 · 0 repositories · arXiv:2407.07011
-
Less is More: Efficient Brain-Inspired Learning for Autonomous Driving Trajectory Prediction 9 Jul 2024 · 0 repositories · arXiv:2407.07020
-
Lookback Lens: Detecting and Mitigating Contextual Hallucinations in Large Language Models Using Only Attention Maps 9 Jul 2024 · 1 repository · arXiv:2407.07071Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 9 unverified (of 20 harvested samples) · 20 pointer-only (licence)
-
LVLM-empowered Multi-modal Representation Learning for Visual Place Recognition 9 Jul 2024 · 0 repositories · arXiv:2407.06730
-
Measuring Sustainability Intention of ESG Fund Disclosure using Few-Shot Learning 9 Jul 2024 · 0 repositories · arXiv:2407.06893
-
Microsoft Cloud-based Digitization Workflow with Rich Metadata Acquisition for Cultural Heritage Objects 9 Jul 2024 · 0 repositories · arXiv:2407.06972
-
Mixture-of-Modules: Reinventing Transformers as Dynamic Assemblies of Modules 9 Jul 2024 · 0 repositories · arXiv:2407.06677
-
MRI Volume-Based Robust Brain Age Estimation Using Weight-Shared Spatial Attention in 3D CNNs 9 Jul 2024 · 0 repositories · arXiv:2407.06686
-
Multiple Instance Verification 9 Jul 2024 · 0 repositories · arXiv:2407.06544
-
Parameter-Efficient and Memory-Efficient Tuning for Vision Transformer: A Disentangled Approach 9 Jul 2024 · 1 repository · arXiv:2407.06964
-
PEER: Expertizing Domain-Specific Tasks with a Multi-Agent Framework and Tuning Methods 9 Jul 2024 · 1 repository · arXiv:2407.06985
-
Prompting Techniques for Secure Code Generation: A Systematic Investigation 9 Jul 2024 · 0 repositories · arXiv:2407.07064
-
Raply: A profanity-mitigated rap generator 9 Jul 2024 · 0 repositories · arXiv:2407.06941
-
Rethinking Image-to-Video Adaptation: An Object-centric Perspective 9 Jul 2024 · 0 repositories · arXiv:2407.06871
-
Segment-Based Interactive Machine Translation for Pre-trained Models 9 Jul 2024 · 0 repositories · arXiv:2407.06990
-
Solving General Natural-Language-Description Optimization Problems with Large Language Models 9 Jul 2024 · 0 repositories · arXiv:2407.07924
-
Source Code Summarization in the Era of Large Language Models 9 Jul 2024 · 1 repository · arXiv:2407.07959
-
Spanish TrOCR: Leveraging Transfer Learning for Language Adaptation 9 Jul 2024 · 1 repository · arXiv:2407.06950
-
TeVAE: A Variational Autoencoder Approach for Discrete Online Anomaly Detection in Variable-state Multivariate Time-series Data 9 Jul 2024 · 1 repository · arXiv:2407.06849
-
Toward Motion Robustness: A masked attention regularization framework in remote photoplethysmography 9 Jul 2024 · 0 repositories · arXiv:2407.06653
-
TrackFormers: In Search of Transformer-Based Particle Tracking for the High-Luminosity LHC Era 9 Jul 2024 · 1 repository · arXiv:2407.07179
-
UnmixingSR: Material-aware Network with Unsupervised Unmixing as Auxiliary Task for Hyperspectral Image Super-resolution 9 Jul 2024 · 0 repositories · arXiv:2407.06525
-
Using Large Language Models for Generating Smart Contracts for Health Insurance from Textual Policies 9 Jul 2024 · 0 repositories · arXiv:2407.07019
-
Using Pretrained Large Language Model with Prompt Engineering to Answer Biomedical Questions 9 Jul 2024 · 0 repositories · arXiv:2407.06779
-
Vision-and-Language Navigation Today and Tomorrow: A Survey in the Era of Foundation Models 9 Jul 2024 · 1 repository · arXiv:2407.07035
-
A Survey on LoRA of Large Language Models 8 Jul 2024 · 1 repository · arXiv:2407.11046
-
3D Vision and Language Pretraining with Large-Scale Synthetic Data 8 Jul 2024 · 1 repository · arXiv:2407.06084Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
SOLO: A Single Transformer for Scalable Vision-Language Modeling 8 Jul 2024 · 1 repository · arXiv:2407.06438Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
An Empirical Comparison of Vocabulary Expansion and Initialization Approaches for Language Models 8 Jul 2024 · 1 repository · arXiv:2407.05841Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Transfer or Self-Supervised? Bridging the Performance Gap in Medical Imaging 8 Jul 2024 · 0 repositories · arXiv:2407.05592
-
BEVWorld: A Multimodal World Model for Autonomous Driving via Unified BEV Latent Space 8 Jul 2024 · 1 repository · arXiv:2407.05679
-
CharSS: Character-Level Transformer Model for Sanskrit Word Segmentation 8 Jul 2024 · 0 repositories · arXiv:2407.06331
-
CodeUpdateArena: Benchmarking Knowledge Editing on API Updates 8 Jul 2024 · 1 repository · arXiv:2407.06249
-
Cross-domain Few-shot In-context Learning for Enhancing Traffic Sign Recognition 8 Jul 2024 · 0 repositories · arXiv:2407.05814