Methods › General › Attention Mechanisms › Attention › Papers, page 207
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 207 of 316: papers 20,601 to 20,700 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
RSIR Transformer: Hierarchical Vision Transformer using Random Sampling Windows and Important Region Windows 13 Apr 2023 · 0 repositories · arXiv:2304.06250
-
Shall We Pretrain Autoregressive Language Models with Retrieval? A Comprehensive Study 13 Apr 2023 · 1 repository · arXiv:2304.06762
-
Sign Language Translation from Instructional Videos 13 Apr 2023 · 1 repository · arXiv:2304.06371
-
TransHP: Image Classification with Hierarchical Prompting 13 Apr 2023 · 1 repository · arXiv:2304.06385Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
VISION DIFFMASK: Faithful Interpretation of Vision Transformers with Differentiable Patch Masking 13 Apr 2023 · 1 repository · arXiv:2304.06391
-
What does CLIP know about a red circle? Visual prompt engineering for VLMs 13 Apr 2023 · 0 repositories · arXiv:2304.06712
-
An Improved Heart Disease Prediction Using Stacked Ensemble Method 12 Apr 2023 · 0 repositories · arXiv:2304.06015
-
Detection of Fake Generated Scientific Abstracts 12 Apr 2023 · 1 repository · arXiv:2304.06148
-
Distilling Token-Pruned Pose Transformer for 2D Human Pose Estimation 12 Apr 2023 · 0 repositories · arXiv:2304.05548
-
DUFormer: Solving Power Line Detection Task in Aerial Images using Semantic Segmentation 12 Apr 2023 · 0 repositories · arXiv:2304.05821
-
Evaluation of ChatGPT Model for Vulnerability Detection 12 Apr 2023 · 0 repositories · arXiv:2304.07232
-
Galactic ChitChat: Using Large Language Models to Converse with Astronomy Literature 12 Apr 2023 · 0 repositories · arXiv:2304.05406
-
Localizing Model Behavior with Path Patching 12 Apr 2023 · 1 repository · arXiv:2304.05969Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
MED-VT++: Unifying Multimodal Learning with a Multiscale Encoder-Decoder Video Transformer 12 Apr 2023 · 0 repositories · arXiv:2304.05930
-
Multi-scale Geometry-aware Transformer for 3D Point Cloud Classification 12 Apr 2023 · 0 repositories · arXiv:2304.05694
-
PATMAT: Person Aware Tuning of Mask-Aware Transformer for Face Inpainting 12 Apr 2023 · 2 repositories · arXiv:2304.06107
-
Real-time Trajectory-based Social Group Detection 12 Apr 2023 · 1 repository · arXiv:2304.05678
-
RECLIP: Resource-efficient CLIP by Training with Small Images 12 Apr 2023 · 0 repositories · arXiv:2304.06028
-
Towards Evaluating Explanations of Vision Transformers for Medical Imaging 12 Apr 2023 · 1 repository · arXiv:2304.06133
-
A Billion-scale Foundation Model for Remote Sensing Images 11 Apr 2023 · 0 repositories · arXiv:2304.05215
-
Approximating Online Human Evaluation of Social Chatbots with Prompting 11 Apr 2023 · 0 repositories · arXiv:2304.05253
-
Bayesian Optimization of Catalysis With In-Context Learning 11 Apr 2023 · 2 repositories · arXiv:2304.05341
-
chatClimate: Grounding Conversational AI in Climate Science 11 Apr 2023 · 0 repositories · arXiv:2304.05510
-
ChemCrow: Augmenting large-language models with chemistry tools 11 Apr 2023 · 3 repositories · arXiv:2304.05376Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Data-Efficient Image Quality Assessment with Attention-Panel Decoder 11 Apr 2023 · 1 repository · arXiv:2304.04952Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Distinguishing ChatGPT(-3.5, -4)-generated and human-written papers through Japanese stylometric analysis 11 Apr 2023 · 0 repositories · arXiv:2304.05534
-
Exploring the Use of Foundation Models for Named Entity Recognition and Lemmatization Tasks in Slavic Languages 11 Apr 2023 · 0 repositories · arXiv:2304.05336
-
MC-ViViT: Multi-branch Classifier-ViViT to detect Mild Cognitive Impairment in older adults using facial videos 11 Apr 2023 · 0 repositories · arXiv:2304.05292
-
Multi-Graph Convolution Network for Pose Forecasting 11 Apr 2023 · 0 repositories · arXiv:2304.04956
-
Multi-step Jailbreaking Privacy Attacks on ChatGPT 11 Apr 2023 · 1 repository · arXiv:2304.05197Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Panoramic Image-to-Image Translation 11 Apr 2023 · 0 repositories · arXiv:2304.04960
-
Sim-T: Simplify the Transformer Network by Multiplexing Technique for Speech Recognition 11 Apr 2023 · 0 repositories · arXiv:2304.04991
-
Towards preserving word order importance through Forced Invalidation 11 Apr 2023 · 1 repository · arXiv:2304.05221
-
Training Large Language Models Efficiently with Sparsity and Dataflow 11 Apr 2023 · 0 repositories · arXiv:2304.05511
-
Video Event Restoration Based on Keyframes for Video Anomaly Detection 11 Apr 2023 · 0 repositories · arXiv:2304.05112
-
Weakly Supervised Intracranial Hemorrhage Segmentation using Head-Wise Gradient-Infused Self-Attention Maps from a Swin Transformer in Categorical Learning 11 Apr 2023 · 1 repository · arXiv:2304.04902
-
Automated Reading Passage Generation with OpenAI's Large Language Model 10 Apr 2023 · 0 repositories · arXiv:2304.04616
-
Detection Transformer with Stable Matching 10 Apr 2023 · 2 repositories · arXiv:2304.04742Syntology official (archive's flag): 1 ran · 7 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Feature Representation Learning with Adaptive Displacement Generation and Transformer Fusion for Micro-Expression Recognition 10 Apr 2023 · 0 repositories · arXiv:2304.04420
-
High Dynamic Range Imaging with Context-aware Transformer 10 Apr 2023 · 0 repositories · arXiv:2304.04416
-
HST-MRF: Heterogeneous Swin Transformer with Multi-Receptive Field for Medical Image Segmentation 10 Apr 2023 · 0 repositories · arXiv:2304.04614
-
Incorporating Structured Sentences with Time-enhanced BERT for Fully-inductive Temporal Relation Prediction 10 Apr 2023 · 0 repositories · arXiv:2304.04717
-
Is ChatGPT a Good Sentiment Analyzer? A Preliminary Study 10 Apr 2023 · 1 repository · arXiv:2304.04339Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Multilingual Machine Translation with Large Language Models: Empirical Results and Analysis 10 Apr 2023 · 2 repositories · arXiv:2304.04675
-
On the Possibilities of AI-Generated Text Detection 10 Apr 2023 · 0 repositories · arXiv:2304.04736
-
Two Steps Forward and One Behind: Rethinking Time Series Forecasting with Deep Learning 10 Apr 2023 · 0 repositories · arXiv:2304.04553
-
Use the Detection Transformer as a Data Augmenter 10 Apr 2023 · 1 repository · arXiv:2304.04554
-
Are Large Language Models Ready for Healthcare? A Comparative Study on Clinical Language Understanding 9 Apr 2023 · 1 repository · arXiv:2304.05368
-
Learning to Tokenize for Generative Retrieval 9 Apr 2023 · 1 repository · arXiv:2304.04171
-
Slide-Transformer: Hierarchical Vision Transformer with Local Self-Attention 9 Apr 2023 · 1 repository · arXiv:2304.04237Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Sparse Dense Fusion for 3D Object Detection 9 Apr 2023 · 0 repositories · arXiv:2304.04179
-
Transformer Utilization in Medical Image Segmentation Networks 9 Apr 2023 · 0 repositories · arXiv:2304.04225
-
Factify 2: A Multimodal Fake News and Satire News Dataset 8 Apr 2023 · 1 repository · arXiv:2304.03897
-
FlexMoE: Scaling Large-scale Sparse Pre-trained Model Training via Dynamic Device Placement 8 Apr 2023 · 0 repositories · arXiv:2304.03946
-
GPT4Rec: A Generative Framework for Personalized Recommendation and User Interests Interpretation 8 Apr 2023 · 0 repositories · arXiv:2304.03879
-
Interpretable Multi Labeled Bengali Toxic Comments Classification using Deep Learning 8 Apr 2023 · 1 repository · arXiv:2304.04087
-
Multi-class Categorization of Reasons behind Mental Disturbance in Long Texts 8 Apr 2023 · 0 repositories · arXiv:2304.04118
-
Surrogate Lagrangian Relaxation: A Path To Retrain-free Deep Neural Network Pruning 8 Apr 2023 · 0 repositories · arXiv:2304.04120
-
tmn at SemEval-2023 Task 9: Multilingual Tweet Intimacy Detection using XLM-T, Google Translate, and Ensemble Learning 8 Apr 2023 · 1 repository · arXiv:2304.04054
-
A Cross-Scale Hierarchical Transformer with Correspondence-Augmented Attention for inferring Bird's-Eye-View Semantic Segmentation 7 Apr 2023 · 0 repositories · arXiv:2304.03650
-
Cleansing Jewel: A Neural Spelling Correction Model Built On Google OCR-ed Tibetan Manuscripts 7 Apr 2023 · 0 repositories · arXiv:2304.03427
-
Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4 7 Apr 2023 · 1 repository · arXiv:2304.03439
-
Hierarchical Catalogue Generation for Literature Review: A Benchmark 7 Apr 2023 · 1 repository · arXiv:2304.03512
-
PSLT: A Light-weight Vision Transformer with Ladder Self-Attention and Progressive Shift 7 Apr 2023 · 0 repositories · arXiv:2304.03481
-
SparseFormer: Sparse Visual Recognition via Limited Latent Tokens 7 Apr 2023 · 1 repository · arXiv:2304.03768
-
All Keypoints You Need: Detecting Arbitrary Keypoints on the Body of Triple, High, and Long Jump Athletes 6 Apr 2023 · 1 repository · arXiv:2304.02939
-
Can Large Language Models Play Text Games Well? Current State-of-the-Art and Open Questions 6 Apr 2023 · 0 repositories · arXiv:2304.02868
-
ChatGPT-Crawler: Find out if ChatGPT really knows what it's talking about 6 Apr 2023 · 0 repositories · arXiv:2304.03325
-
Continual Detection Transformer for Incremental Object Detection 6 Apr 2023 · 0 repositories · arXiv:2304.03110
-
Deep Learning for Opinion Mining and Topic Classification of Course Reviews 6 Apr 2023 · 0 repositories · arXiv:2304.03394
-
DeLiRa: Self-Supervised Depth, Light, and Radiance Fields 6 Apr 2023 · 0 repositories · arXiv:2304.02797
-
Efficient Audio Captioning Transformer with Patchout and Text Guidance 6 Apr 2023 · 0 repositories · arXiv:2304.02916
-
FengWu: Pushing the Skillful Global Medium-range Weather Forecast beyond 10 Days Lead 6 Apr 2023 · 2 repositories · arXiv:2304.02948
-
From Saliency to DINO: Saliency-guided Vision Transformer for Few-shot Keypoint Detection 6 Apr 2023 · 0 repositories · arXiv:2304.03140
-
GPT detectors are biased against non-native English writers 6 Apr 2023 · 2 repositories · arXiv:2304.02819Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Instruction Tuning with GPT-4 6 Apr 2023 · 2 repositories · arXiv:2304.03277
-
InterFormer: Real-time Interactive Image Segmentation 6 Apr 2023 · 1 repository · arXiv:2304.02942
-
Making AI Less "Thirsty": Uncovering and Addressing the Secret Water Footprint of AI Models 6 Apr 2023 · 1 repository · arXiv:2304.03271Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
MemeFier: Dual-stage Modality Fusion for Image Meme Classification 6 Apr 2023 · 1 repository · arXiv:2304.02906
-
Micron-BERT: BERT-based Facial Micro-Expression Recognition 6 Apr 2023 · 1 repository · arXiv:2304.03195Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
MULLER: Multilayer Laplacian Resizer for Vision 6 Apr 2023 · 1 repository · arXiv:2304.02859
-
Multi-label classification of open-ended questions with BERT 6 Apr 2023 · 0 repositories · arXiv:2304.02945
-
Towards Interpretable Mental Health Analysis with Large Language Models 6 Apr 2023 · 2 repositories · arXiv:2304.03347
-
PointCAT: Cross-Attention Transformer for point cloud 6 Apr 2023 · 1 repository · arXiv:2304.03012
-
R²Former: Unified Retrieval and Reranking Transformer for Place Recognition 6 Apr 2023 · 0 repositories · arXiv:2304.03410
-
Towards an Effective and Efficient Transformer for Rain-by-snow Weather Removal 6 Apr 2023 · 1 repository · arXiv:2304.02860
-
Zero-Shot Next-Item Recommendation using Large Pretrained Language Models 6 Apr 2023 · 1 repository · arXiv:2304.03153Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Conceptual structure coheres in human cognition but not in large language models 5 Apr 2023 · 0 repositories · arXiv:2304.02754
-
Bengali Fake Review Detection using Semi-supervised Generative Adversarial Networks 5 Apr 2023 · 0 repositories · arXiv:2304.02739
-
ChartReader: A Unified Framework for Chart Derendering and Comprehension without Heuristic Rules 5 Apr 2023 · 1 repository · arXiv:2304.02173Syntology official (archive's flag): 16 ran · 16 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 10 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Context-Aware Classification of Legal Document Pages 5 Apr 2023 · 0 repositories · arXiv:2304.02787
-
Document-Level Machine Translation with Large Language Models 5 Apr 2023 · 1 repository · arXiv:2304.02210
-
Face Transformer: Towards High Fidelity and Accurate Face Swapping 5 Apr 2023 · 0 repositories · arXiv:2304.02530
-
Large Language Models as Master Key: Unlocking the Secrets of Materials Science with GPT 5 Apr 2023 · 0 repositories · arXiv:2304.02213
-
ParroT: Translating during Chat using Large Language Models tuned with Human Translation and Feedback 5 Apr 2023 · 1 repository · arXiv:2304.02426
-
Training Strategies for Vision Transformers for Object Detection 5 Apr 2023 · 0 repositories · arXiv:2304.02186
-
Attention Map Guided Transformer Pruning for Edge Device 4 Apr 2023 · 1 repository · arXiv:2304.01452
-
Blockwise Compression of Transformer-based Models without Retraining 4 Apr 2023 · 0 repositories · arXiv:2304.01483
-
Can BERT eat RuCoLA? Topological Data Analysis to Explain 4 Apr 2023 · 2 repositories · arXiv:2304.01680
-
ConvFormer: Parameter Reduction in Transformer Models for 3D Human Pose Estimation by Leveraging Dynamic Multi-Headed Convolutional Attention 4 Apr 2023 · 1 repository · arXiv:2304.02147