Methods › General › Attention Mechanisms › Attention › Papers, page 164
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 164 of 316: papers 16,301 to 16,400 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Evaluating Large Language Models on the GMAT: Implications for the Future of Business Education 2 Jan 2024 · 0 repositories · arXiv:2401.02985
-
Identification of Regulatory Requirements Relevant to Business Processes: A Comparative Study on Generative AI, Embedding-based Ranking, Crowd and Expert-driven Methods 2 Jan 2024 · 0 repositories · arXiv:2401.02986
-
Joint Generative Modeling of Scene Graphs and Images via Diffusion Models 2 Jan 2024 · 0 repositories · arXiv:2401.01130
-
MOC-RVQ: Multilevel Codebook-Assisted Digital Generative Semantic Communication 2 Jan 2024 · 1 repository · arXiv:2401.01272
-
Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models 2 Jan 2024 · 2 repositories · arXiv:2401.01335Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Uncertainty Resolution in Misinformation Detection 2 Jan 2024 · 0 repositories · arXiv:2401.01197
-
Unifying Structured Data as Graph for Data-to-Text Pre-Training 2 Jan 2024 · 1 repository · arXiv:2401.01183
-
Vietnamese Poem Generation & The Prospect Of Cross-Language Poem-To-Poem Translation 2 Jan 2024 · 1 repository · arXiv:2401.01078
-
1st Place Solution for 5th LSVOS Challenge: Referring Video Object Segmentation 1 Jan 2024 · 1 repository · arXiv:2401.00663
-
LogicAsker: Evaluating and Improving the Logical Reasoning Ability of Large Language Models 1 Jan 2024 · 1 repository · arXiv:2401.00757Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Computational Framework for Behavioral Assessment of LLM Therapists 1 Jan 2024 · 1 repository · arXiv:2401.00820
-
Accurate Leukocyte Detection Based on Deformable-DETR and Multi-Level Feature Fusion for Aiding Diagnosis of Blood Diseases 1 Jan 2024 · 1 repository · arXiv:2401.00926
-
Adapt or Perish: Adaptive Sparse Transformer with Attentive Feature Refinement for Image Restoration 1 Jan 2024 · 1 repository
-
AdaShift: Learning Discriminative Self-Gated Neural Feature Activation With an Adaptive Shift Factor 1 Jan 2024 · 0 repositories
-
Beyond Subspace Isolation: Many-to-Many Transformer for Light Field Image Super-resolution 1 Jan 2024 · 1 repository · arXiv:2401.00740
-
Boosting Image Quality Assessment through Efficient Transformer Adaptation with Local Feature Enhancement 1 Jan 2024 · 1 repository
-
Capturing Closely Interacted Two-Person Motions with Reaction Priors 1 Jan 2024 · 0 repositories
-
CDFormer: When Degradation Prediction Embraces Diffusion Model for Blind Image Super-Resolution 1 Jan 2024 · 1 repository
-
CFAT: Unleashing Triangular Windows for Image Super-resolution 1 Jan 2024 · 1 repository
-
Class Tokens Infusion for Weakly Supervised Semantic Segmentation 1 Jan 2024 · 1 repository
-
DeiT-LT: Distillation Strikes Back for Vision Transformer Training on Long-Tailed Datasets 1 Jan 2024 · 0 repositories
-
Dynamic Cues-Assisted Transformer for Robust Point Cloud Registration 1 Jan 2024 · 0 repositories
-
Enhanced Motion-Text Alignment for Image-to-Video Transfer Learning 1 Jan 2024 · 0 repositories
-
FAR: Flexible Accurate and Robust 6DoF Relative Camera Pose Estimation 1 Jan 2024 · 0 repositories
-
Flexible Biometrics Recognition: Bridging the Multimodality Gap through Attention Alignment and Prompt Tuning 1 Jan 2024 · 1 repository
-
H-ViT: A Hierarchical Vision Transformer for Deformable Image Registration 1 Jan 2024 · 1 repository
-
Harmonizing SO(3)-Equivariance with Neural Expressiveness: a Hybrid Deep Learning Framework Oriented to the Prediction of Electronic Structure Hamiltonian 1 Jan 2024 · 0 repositories · arXiv:2401.00744
-
Hybrid Proposal Refiner: Revisiting DETR Series from the Faster R-CNN Perspective 1 Jan 2024 · 1 repository
-
JointSQ: Joint Sparsification-Quantization for Distributed Learning 1 Jan 2024 · 1 repository
-
KD-DETR: Knowledge Distillation for Detection Transformer with Consistent Distillation Points Sampling 1 Jan 2024 · 0 repositories
-
Large Language Models aren't all that you need 1 Jan 2024 · 0 repositories · arXiv:2401.00698
-
Large Language Models in Mental Health Care: a Scoping Review 1 Jan 2024 · 0 repositories · arXiv:2401.02984
-
Learn to Rectify the Bias of CLIP for Unsupervised Semantic Segmentation 1 Jan 2024 · 1 repository
-
Linguistic-Aware Patch Slimming Framework for Fine-grained Cross-Modal Alignment 1 Jan 2024 · 1 repository
-
Mean-Shift Feature Transformer 1 Jan 2024 · 1 repository
-
Multi-Attribute Interactions Matter for 3D Visual Grounding 1 Jan 2024 · 1 repository
-
PairDETR : Joint Detection and Association of Human Bodies and Faces 1 Jan 2024 · 1 repository
-
ParameterNet: Parameters Are All You Need for Large-scale Visual Pretraining of Mobile Networks 1 Jan 2024 · 0 repositories
-
Person-in-WiFi 3D: End-to-End Multi-Person 3D Pose Estimation with Wi-Fi 1 Jan 2024 · 0 repositories
-
Point Transformer V3: Simpler Faster Stronger 1 Jan 2024 · 1 repository
-
Pre-training Vision Models with Mandelbulb Variations 1 Jan 2024 · 1 repository
-
Random Entangled Tokens for Adversarially Robust Vision Transformer 1 Jan 2024 · 0 repositories
-
Revisiting Counterfactual Problems in Referring Expression Comprehension 1 Jan 2024 · 1 repository
-
SecFormer: Fast and Accurate Privacy-Preserving Inference for Transformer Models via SMPC 1 Jan 2024 · 1 repository · arXiv:2401.00793Syntology official: harvested, nothing ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
SEED-Bench: Benchmarking Multimodal Large Language Models 1 Jan 2024 · 1 repository
-
SlowFormer: Adversarial Attack on Compute and Energy Consumption of Efficient Vision Transformers 1 Jan 2024 · 1 repository
-
SPECAT: SPatial-spEctral Cumulative-Attention Transformer for High-Resolution Hyperspectral Image Reconstruction 1 Jan 2024 · 1 repository
-
Taking the Next Step with Generative Artificial Intelligence: The Transformative Role of Multimodal Large Language Models in Science Education 1 Jan 2024 · 0 repositories · arXiv:2401.00832
-
Time- Memory- and Parameter-Efficient Visual Adaptation 1 Jan 2024 · 0 repositories
-
Training Vision Transformers for Semi-Supervised Semantic Segmentation 1 Jan 2024 · 1 repository
-
TransLoc4D: Transformer-based 4D Radar Place Recognition 1 Jan 2024 · 1 repository
-
UFC-Net: Unrolling Fixed-point Continuous Network for Deep Compressive Sensing 1 Jan 2024 · 0 repositories
-
Uncertainty-aware Action Decoupling Transformer for Action Anticipation 1 Jan 2024 · 0 repositories
-
Unlocking the Potential of Pre-trained Vision Transformers for Few-Shot Semantic Segmentation through Relationship Descriptors 1 Jan 2024 · 1 repository
-
Video Harmonization with Triplet Spatio-Temporal Variation Patterns 1 Jan 2024 · 1 repository
-
A Two-stream Hybrid CNN-Transformer Network for Skeleton-based Human Interaction Recognition 31 Dec 2023 · 0 repositories · arXiv:2401.00409
-
An Analysis of Embedding Layers and Similarity Scores using Siamese Neural Networks 31 Dec 2023 · 0 repositories · arXiv:2401.00582
-
Analyzing Local Representations of Self-supervised Vision Transformers 31 Dec 2023 · 0 repositories · arXiv:2401.00463
-
EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling 31 Dec 2023 · 1 repository · arXiv:2401.00374Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Generative Model-Driven Synthetic Training Image Generation: An Approach to Cognition in Rail Defect Detection 31 Dec 2023 · 1 repository · arXiv:2401.00393
-
RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models 31 Dec 2023 · 3 repositories · arXiv:2401.00396Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
SVFAP: Self-supervised Video Facial Affect Perceiver 31 Dec 2023 · 1 repository · arXiv:2401.00416
-
Advancing TTP Analysis: Harnessing the Power of Large Language Models with Retrieval Augmented Generation 30 Dec 2023 · 1 repository · arXiv:2401.00280
-
HybridGait: A Benchmark for Spatial-Temporal Cloth-Changing Gait Recognition with Hybrid Explorations 30 Dec 2023 · 1 repository · arXiv:2401.00271
-
Image Super-resolution Reconstruction Network based on Enhanced Swin Transformer via Alternating Aggregation of Local-Global Features 30 Dec 2023 · 0 repositories · arXiv:2401.00241
-
L3Cube-MahaSocialNER: A Social Media based Marathi NER Dataset and BERT models 30 Dec 2023 · 1 repository · arXiv:2401.00170
-
Trace and Edit Relation Associations in GPT 30 Dec 2023 · 0 repositories · arXiv:2401.02976
-
Why is the User Interface a Dark Pattern? : Explainable Auto-Detection and its Analysis 30 Dec 2023 · 1 repository · arXiv:2401.04119
-
A Fully Automated Pipeline Using Swin Transformers for Deep Learning-Based Blood Segmentation on Head CT Scans After Aneurysmal Subarachnoid Hemorrhage 29 Dec 2023 · 0 repositories · arXiv:2312.17553
-
Action-Item-Driven Summarization of Long Meeting Transcripts 29 Dec 2023 · 1 repository · arXiv:2312.17581
-
Adaptive Control Strategy for Quadruped Robots in Actuator Degradation Scenarios 29 Dec 2023 · 1 repository · arXiv:2312.17606
-
Efficacy of Utilizing Large Language Models to Detect Public Threat Posted Online 29 Dec 2023 · 0 repositories · arXiv:2401.02974
-
Enhancing Quantitative Reasoning Skills of Large Language Models through Dimension Perception 29 Dec 2023 · 0 repositories · arXiv:2312.17532
-
Gemini in Reasoning: Unveiling Commonsense in Multimodal Large Language Models 29 Dec 2023 · 1 repository · arXiv:2312.17661
-
Jatmo: Prompt Injection Defense by Task-Specific Finetuning 29 Dec 2023 · 1 repository · arXiv:2312.17673Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
MosaicBERT: A Bidirectional Encoder Optimized for Fast Pretraining 29 Dec 2023 · 1 repository · arXiv:2312.17482
-
Multiscale Vision Transformers meet Bipartite Matching for efficient single-stage Action Localization 29 Dec 2023 · 1 repository · arXiv:2312.17686
-
TuPy-E: detecting hate speech in Brazilian Portuguese social media with a novel dataset and comprehensive analysis of models 29 Dec 2023 · 1 repository · arXiv:2312.17704
-
XAI for In-hospital Mortality Prediction via Multimodal ICU Data 29 Dec 2023 · 1 repository · arXiv:2312.17624
-
MR-GSM8K: A Meta-Reasoning Benchmark for Large Language Model Evaluation 28 Dec 2023 · 2 repositories · arXiv:2312.17080
-
Evaluating the Performance of Large Language Models for Spanish Language in Undergraduate Admissions Exams 28 Dec 2023 · 0 repositories · arXiv:2312.16845
-
Geometry-Biased Transformer for Robust Multi-View 3D Human Pose Reconstruction 28 Dec 2023 · 0 repositories · arXiv:2312.17106
-
Enhancing Open-Domain Task-Solving Capability of LLMs via Autonomous Tool Integration from GitHub 28 Dec 2023 · 1 repository · arXiv:2312.17294
-
Language Model as an Annotator: Unsupervised Context-aware Quality Phrase Generation 28 Dec 2023 · 0 repositories · arXiv:2312.17349
-
LLM4Causal: Democratized Causal Tools for Everyone via Large Language Model 28 Dec 2023 · 0 repositories · arXiv:2312.17122
-
Learning Multi-axis Representation in Frequency Domain for Medical Image Segmentation 28 Dec 2023 · 1 repository · arXiv:2312.17030
-
Learning Vision from Models Rivals Learning Vision from Data 28 Dec 2023 · 2 repositories · arXiv:2312.17742Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Length Extrapolation of Transformers: A Survey from the Perspective of Positional Encoding 28 Dec 2023 · 0 repositories · arXiv:2312.17044
-
Replication-proof Bandit Mechanism Design with Bayesian Agents 28 Dec 2023 · 0 repositories · arXiv:2312.16896
-
ROI-Aware Multiscale Cross-Attention Vision Transformer for Pest Image Identification 28 Dec 2023 · 0 repositories · arXiv:2312.16914
-
SentinelLMs: Encrypted Input Adaptation and Fine-tuning of Language Models for Private and Secure Inference 28 Dec 2023 · 1 repository · arXiv:2312.17342
-
The LLM Surgeon 28 Dec 2023 · 1 repository · arXiv:2312.17244Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
A New Perspective on Speaker Verification: Joint Modeling with DFSMN and Transformer 28 Dec 2023 · 1 repository · arXiv:2312.16826
-
Weed mapping in multispectral drone imagery using lightweight vision transformers 28 Dec 2023 · 1 repository
-
A Non-Uniform Low-Light Image Enhancement Method with Multi-Scale Attention Transformer and Luminance Consistency Loss 27 Dec 2023 · 1 repository · arXiv:2312.16498
-
Group Multi-View Transformer for 3D Shape Analysis with Spatial Encoding 27 Dec 2023 · 1 repository · arXiv:2312.16477
-
Learn From Orientation Prior for Radiograph Super-Resolution: Orientation Operator Transformer 27 Dec 2023 · 0 repositories · arXiv:2312.16455
-
PanGu-π: Enhancing Language Model Architectures via Nonlinearity Compensation 27 Dec 2023 · 0 repositories · arXiv:2312.17276
-
RefineNet: Enhancing Text-to-Image Conversion with High-Resolution and Detail Accuracy through Hierarchical Transformers and Progressive Refinement 27 Dec 2023 · 0 repositories · arXiv:2312.17274
-
Relationship between auditory and semantic entrainment using Deep Neural Networks (DNN) 27 Dec 2023 · 0 repositories · arXiv:2312.16599