Methods › General › Attention Mechanisms › Attention › Papers, page 205
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 205 of 316: papers 20,401 to 20,500 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Towards Better Domain Adaptation for Self-supervised Models: A Case Study of Child ASR 28 Apr 2023 · 1 repository · arXiv:2305.00115
-
Assessing Text Mining and Technical Analyses on Forecasting Financial Time Series 27 Apr 2023 · 0 repositories · arXiv:2304.14544
-
Boosting Big Brother: Attacking Search Engines with Encodings 27 Apr 2023 · 1 repository · arXiv:2304.14031
-
ChatGPT as an Attack Tool: Stealthy Textual Backdoor Attack via Blackbox Generative Model Trigger 27 Apr 2023 · 0 repositories · arXiv:2304.14475
-
CONSCENDI: A Contrastive and Scenario-Guided Distillation Approach to Guardrail Models for Virtual Assistants 27 Apr 2023 · 0 repositories · arXiv:2304.14364
-
DataComp: In search of the next generation of multimodal datasets 27 Apr 2023 · 3 repositories · arXiv:2304.14108Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Deeply-Coupled Convolution-Transformer with Spatial-temporal Complementary Learning for Video-based Person Re-identification 27 Apr 2023 · 1 repository · arXiv:2304.14122
-
Distinguishing a planetary transit from false positives: a Transformer-based classification for planetary transit signals 27 Apr 2023 · 0 repositories · arXiv:2304.14283
-
Exploiting Inductive Bias in Transformer for Point Cloud Classification and Segmentation 27 Apr 2023 · 1 repository · arXiv:2304.14124
-
Framing the News:From Human Perception to Large Language Model Inferences 27 Apr 2023 · 0 repositories · arXiv:2304.14456
-
ICE-Score: Instructing Large Language Models to Evaluate Code 27 Apr 2023 · 2 repositories · arXiv:2304.14317Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Lightweight, Pre-trained Transformers for Remote Sensing Timeseries 27 Apr 2023 · 1 repository · arXiv:2304.14065Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 13 harvested samples)
-
Neural Keyphrase Generation: Analysis and Evaluation 27 Apr 2023 · 0 repositories · arXiv:2304.13883
-
Optimization-Inspired Cross-Attention Transformer for Compressive Sensing 27 Apr 2023 · 1 repository · arXiv:2304.13986
-
Origin Tracing and Detecting of LLMs 27 Apr 2023 · 0 repositories · arXiv:2304.14072
-
pyBibX -- A Python Library for Bibliometric and Scientometric Analysis Powered with Artificial Intelligence Tools 27 Apr 2023 · 1 repository · arXiv:2304.14516
-
SweCTRL-Mini: a data-transparent Transformer-based large language model for controllable text generation in Swedish 27 Apr 2023 · 1 repository · arXiv:2304.13994
-
TempEE: Temporal-Spatial Parallel Transformer for Radar Echo Extrapolation Beyond Auto-Regression 27 Apr 2023 · 0 repositories · arXiv:2304.14131
-
We're Afraid Language Models Aren't Modeling Ambiguity 27 Apr 2023 · 1 repository · arXiv:2304.14399
-
Prompting GPT-3.5 for Text-to-SQL with De-semanticization and Skeleton Retrieval 26 Apr 2023 · 0 repositories · arXiv:2304.13301
-
Evaluation of GPT-3.5 and GPT-4 for supporting real-world information needs in healthcare delivery 26 Apr 2023 · 0 repositories · arXiv:2304.13714
-
Exploring the Curious Case of Code Prompts 26 Apr 2023 · 1 repository · arXiv:2304.13250Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Extracting Structured Seed-Mediated Gold Nanorod Growth Procedures from Literature with GPT-3 26 Apr 2023 · 0 repositories · arXiv:2304.13846
-
Fine Tuning with Abnormal Examples 26 Apr 2023 · 0 repositories · arXiv:2304.13783
-
HausaNLP at SemEval-2023 Task 12: Leveraging African Low Resource TweetData for Sentiment Analysis 26 Apr 2023 · 1 repository · arXiv:2304.13634
-
Technical Report: Impact of Position Bias on Language Models in Token Classification 26 Apr 2023 · 2 repositories · arXiv:2304.13567
-
The Parrot Dilemma: Human-Labeled vs. LLM-augmented Data in Classification Tasks 26 Apr 2023 · 2 repositories · arXiv:2304.13861Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
ScatterFormer: Locally-Invariant Scattering Transformer for Patient-Independent Multispectral Detection of Epileptiform Discharges 26 Apr 2023 · 1 repository · arXiv:2304.14919Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
SIMARA: a database for key-value information extraction from full pages 26 Apr 2023 · 0 repositories · arXiv:2304.13606
-
STIR: Siamese Transformer for Image Retrieval Postprocessing 26 Apr 2023 · 1 repository · arXiv:2304.13393
-
TextDeformer: Geometry Manipulation using Text Guidance 26 Apr 2023 · 1 repository · arXiv:2304.13348Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 16 harvested samples) · 1 pointer-only (licence)
-
The Closeness of In-Context Learning and Weight Shifting for Softmax Regression 26 Apr 2023 · 0 repositories · arXiv:2304.13276
-
Towards Multi-Modal DBMSs for Seamless Querying of Texts and Tables 26 Apr 2023 · 0 repositories · arXiv:2304.13559
-
AI-assisted coding: Experiments with GPT-4 25 Apr 2023 · 1 repository · arXiv:2304.13187
-
Application of Transformers for Nonlinear Channel Compensation in Optical Systems 25 Apr 2023 · 0 repositories · arXiv:2304.13119
-
CompletionFormer: Depth Completion with Convolutions and Vision Transformers 25 Apr 2023 · 1 repository · arXiv:2304.13030
-
Depth-Relative Self Attention for Monocular Depth Estimation 25 Apr 2023 · 0 repositories · arXiv:2304.12849
-
DuETT: Dual Event Time Transformer for Electronic Health Records 25 Apr 2023 · 1 repository · arXiv:2304.13017Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Escaping the sentence-level paradigm in machine translation 25 Apr 2023 · 1 repository · arXiv:2304.12959
-
Introducing MBIB -- the first Media Bias Identification Benchmark Task and Dataset Collection 25 Apr 2023 · 1 repository · arXiv:2304.13148
-
LEMaRT: Label-Efficient Masked Region Transform for Image Harmonization 25 Apr 2023 · 0 repositories · arXiv:2304.13166
-
Measuring Massive Multitask Chinese Understanding 25 Apr 2023 · 2 repositories · arXiv:2304.12986Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
NLP-LTU at SemEval-2023 Task 10: The Impact of Data Augmentation and Semi-Supervised Learning Techniques on Text Classification Performance on an Imbalanced Dataset 25 Apr 2023 · 0 repositories · arXiv:2304.12847
-
SDSC-UNet: Dual Skip Connection ViT-based U-shaped Model for Building Extraction 25 Apr 2023 · 1 repository
-
Semantic Compression With Large Language Models 25 Apr 2023 · 0 repositories · arXiv:2304.12512
-
State Spaces Aren't Enough: Machine Translation Needs Attention 25 Apr 2023 · 0 repositories · arXiv:2304.12776
-
STM-UNet: An Efficient U-shaped Architecture Based on Swin Transformer and Multi-scale MLP for Medical Image Segmentation 25 Apr 2023 · 0 repositories · arXiv:2304.12615
-
SwinFSR: Stereo Image Super-Resolution using SwinIR and Frequency Domain Knowledge 25 Apr 2023 · 0 repositories · arXiv:2304.12556
-
The Potential of Visual ChatGPT For Remote Sensing 25 Apr 2023 · 0 repositories · arXiv:2304.13009
-
Theory of Posterior Concentration for Generalized Bayesian Additive Regression Trees 25 Apr 2023 · 0 repositories · arXiv:2304.12505
-
What does BERT learn about prosody? 25 Apr 2023 · 0 repositories · arXiv:2304.12706
-
AGI: Artificial General Intelligence for Education 24 Apr 2023 · 0 repositories · arXiv:2304.12479
-
Augmentation-based Domain Generalization for Semantic Segmentation 24 Apr 2023 · 0 repositories · arXiv:2304.12122
-
Better Question-Answering Models on a Budget 24 Apr 2023 · 1 repository · arXiv:2304.12370
-
Directed Acyclic Transformer Pre-training for High-quality Non-autoregressive Text Generation 24 Apr 2023 · 1 repository · arXiv:2304.11791Syntology official (archive's flag): 3 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Explicit Correspondence Matching for Generalizable Neural Radiance Fields 24 Apr 2023 · 1 repository · arXiv:2304.12294Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
Extreme Classification for Answer Type Prediction in Question Answering 24 Apr 2023 · 0 repositories · arXiv:2304.12395
-
Generation-driven Contrastive Self-training for Zero-shot Text Classification with Instruction-following LLM 24 Apr 2023 · 1 repository · arXiv:2304.11872
-
IRNeXt: Rethinking Convolutional Network Design for Image Restoration 24 Apr 2023 · 1 repository
-
Master: Meta Style Transformer for Controllable Zero-Shot and Few-Shot Artistic Style Transfer 24 Apr 2023 · 0 repositories · arXiv:2304.11818
-
MixPro: Data Augmentation with MaskMix and Progressive Attention Labeling for Vision Transformer 24 Apr 2023 · 1 repository · arXiv:2304.12043Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 7 where Syntology's instrument failed) · 2 unverified (of 18 harvested samples) · 8 pointer-only (licence)
-
NoiseTrans: Point Cloud Denoising with Transformers 24 Apr 2023 · 0 repositories · arXiv:2304.11812
-
Once Detected, Never Lost: Surpassing Human Performance in Offline LiDAR based 3D Object Detection 24 Apr 2023 · 2 repositories · arXiv:2304.12315Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
PARAGRAPH2GRAPH: A GNN-based framework for layout paragraph analysis 24 Apr 2023 · 1 repository · arXiv:2304.11810
-
Pre-trained Embeddings for Entity Resolution: An Experimental Analysis [Experiment, Analysis & Benchmark] 24 Apr 2023 · 1 repository · arXiv:2304.12329
-
Rank Flow Embedding for Unsupervised and Semi-Supervised Manifold Learning 24 Apr 2023 · 1 repository · arXiv:2304.12448
-
Vision-based Estimation of Fatigue and Engagement in Cognitive Training Sessions 24 Apr 2023 · 1 repository · arXiv:2304.12470
-
Self-regularised Minimum Latency Training for Streaming Transformer-based Speech Recognition 24 Apr 2023 · 0 repositories · arXiv:2304.11985
-
SocialDial: A Benchmark for Socially-Aware Dialogue Systems 24 Apr 2023 · 1 repository · arXiv:2304.12026
-
Text-to-Audio Generation using Instruction-Tuned LLM and Latent Diffusion Model 24 Apr 2023 · 1 repository · arXiv:2304.13731
-
Transformer-based stereo-aware 3D object detection from binocular images 24 Apr 2023 · 0 repositories · arXiv:2304.11906
-
Universal Domain Adaptation via Compressive Attention Matching 24 Apr 2023 · 0 repositories · arXiv:2304.11862
-
WizardLM: Empowering Large Language Models to Follow Complex Instructions 24 Apr 2023 · 4 repositories · arXiv:2304.12244Syntology official (archive's flag): 1 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Processing Natural Language on Embedded Devices: How Well Do Transformer Models Perform? 23 Apr 2023 · 2 repositories · arXiv:2304.11520
-
Vision Transformer for Efficient Chest X-ray and Gastrointestinal Image Classification 23 Apr 2023 · 0 repositories · arXiv:2304.11529
-
Boosting Theory-of-Mind Performance in Large Language Models via Prompting 22 Apr 2023 · 1 repository · arXiv:2304.11490
-
Dilated-UNet: A Fast and Accurate Medical Image Segmentation Approach using a Dilated Transformer and U-Net Architecture 22 Apr 2023 · 1 repository · arXiv:2304.11450
-
Incomplete Multimodal Learning for Remote Sensing Data Fusion 22 Apr 2023 · 0 repositories · arXiv:2304.11381
-
L3Cube-IndicSBERT: A simple approach for learning cross-lingual sentence representations using multilingual BERT 22 Apr 2023 · 0 repositories · arXiv:2304.11434
-
Vision Transformers, a new approach for high-resolution and large-scale mapping of canopy heights 22 Apr 2023 · 0 repositories · arXiv:2304.11487
-
A Group-Specific Approach to NLP for Hate Speech Detection 21 Apr 2023 · 1 repository · arXiv:2304.11223
-
BERT Based Clinical Knowledge Extraction for Biomedical Knowledge Graph Construction and Analysis 21 Apr 2023 · 0 repositories · arXiv:2304.10996
-
Building Multimodal AI Chatbots 21 Apr 2023 · 1 repository · arXiv:2305.03512
-
Can GPT-4 Perform Neural Architecture Search? 21 Apr 2023 · 1 repository · arXiv:2304.10970
-
DeformableFormer: Classification of Endoscopic Ultrasound Guided Fine Needle Biopsy in Pancreatic Diseases 21 Apr 2023 · 0 repositories · arXiv:2304.10791
-
Evaluating Transformer Language Models on Arithmetic Operations Using Number Decomposition 21 Apr 2023 · 1 repository · arXiv:2304.10977
-
Inducing anxiety in large language models can induce bias 21 Apr 2023 · 0 repositories · arXiv:2304.11111
-
Multi-Modal Deep Learning for Credit Rating Prediction Using Text and Numerical Data Streams 21 Apr 2023 · 1 repository · arXiv:2304.10740
-
Self-Attention in Colors: Another Take on Encoding Graph Structure in Transformers 21 Apr 2023 · 1 repository · arXiv:2304.10933
-
Text2Time: Transformer-based Article Time Period Prediction 21 Apr 2023 · 0 repositories · arXiv:2304.10859
-
The Dark Side of ChatGPT: Legal and Ethical Challenges from Stochastic Parrots and Hallucination 21 Apr 2023 · 0 repositories · arXiv:2304.14347
-
Transformer-based models and hardware acceleration analysis in autonomous driving: A survey 21 Apr 2023 · 0 repositories · arXiv:2304.10891
-
Who's the Best Detective? LLMs vs. MLs in Detecting Incoherent Fourth Grade Math Answers 21 Apr 2023 · 0 repositories · arXiv:2304.11257
-
Analyzing FOMC Minutes: Accuracy and Constraints of Language Models 20 Apr 2023 · 0 repositories · arXiv:2304.10164
-
Contrastive Tuning: A Little Help to Make Masked Autoencoders Forget 20 Apr 2023 · 1 repository · arXiv:2304.10520
-
Domain-specific Continued Pretraining of Language Models for Capturing Long Context in Mental Health 20 Apr 2023 · 0 repositories · arXiv:2304.10447
-
HM-ViT: Hetero-modal Vehicle-to-Vehicle Cooperative perception with vision transformer 20 Apr 2023 · 1 repository · arXiv:2304.10628Syntology 17 ran (of which 7 constructed an object rather than computing a result; 9 with no instrument failure: 2 honoured, 0 violated, 7 with no contract checked; 8 where Syntology's instrument failed) · 4 unverified (of 21 harvested samples) · 21 pointer-only (licence)
-
Is Cross-modal Information Retrieval Possible without Training? 20 Apr 2023 · 0 repositories · arXiv:2304.11095
-
Learning CLIP Guided Visual-Text Fusion Transformer for Video-based Pedestrian Attribute Recognition 20 Apr 2023 · 1 repository · arXiv:2304.10091
-
Meta Semantics: Towards better natural language understanding and reasoning 20 Apr 2023 · 0 repositories · arXiv:2304.10663