Methods › General › Attention Mechanisms › Attention › Papers, page 165
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 165 of 316: papers 16,401 to 16,500 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
scRNA-seq Data Clustering by Cluster-aware Iterative Contrastive Learning 27 Dec 2023 · 1 repository · arXiv:2312.16600
-
Spatial-Related Sensors Matters: 3D Human Motion Reconstruction Assisted with Textual Semantics 27 Dec 2023 · 0 repositories · arXiv:2401.05412
-
Attention-aware Social Graph Transformer Networks for Stochastic Trajectory Prediction 26 Dec 2023 · 0 repositories · arXiv:2312.15881
-
C2T-Net: Channel-Aware Cross-Fused Transformer-Style Networks for Pedestrian Attribute Recognition 26 Dec 2023 · 1 repository
-
ChartBench: A Benchmark for Complex Visual Reasoning in Charts 26 Dec 2023 · 0 repositories · arXiv:2312.15915
-
Graph Context Transformation Learning for Progressive Correspondence Pruning 26 Dec 2023 · 1 repository · arXiv:2312.15971
-
Heterogeneous Encoders Scaling In The Transformer For Neural Machine Translation 26 Dec 2023 · 0 repositories · arXiv:2312.15872
-
LangSplat: 3D Language Gaussian Splatting 26 Dec 2023 · 1 repository · arXiv:2312.16084Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Modality-Collaborative Transformer with Hybrid Feature Reconstruction for Robust Emotion Recognition 26 Dec 2023 · 1 repository · arXiv:2312.15848
-
PDiT: Interleaving Perception and Decision-making Transformers for Deep Reinforcement Learning 26 Dec 2023 · 2 repositories · arXiv:2312.15863
-
Principled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4 26 Dec 2023 · 2 repositories · arXiv:2312.16171Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
RoleEval: A Bilingual Role Evaluation Benchmark for Large Language Models 26 Dec 2023 · 1 repository · arXiv:2312.16132
-
Scaling Down, LiTting Up: Efficient Zero-Shot Listwise Reranking with Seq2seq Encoder-Decoder Models 26 Dec 2023 · 2 repositories · arXiv:2312.16098
-
SecQA: A Concise Question-Answering Dataset for Evaluating Large Language Models in Computer Security 26 Dec 2023 · 1 repository · arXiv:2312.15838Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Task Contamination: Language Models May Not Be Few-Shot Anymore 26 Dec 2023 · 0 repositories · arXiv:2312.16337
-
Compositional Generalization in Spoken Language Understanding 25 Dec 2023 · 0 repositories · arXiv:2312.15815
-
Deep Structure and Attention Aware Subspace Clustering 25 Dec 2023 · 1 repository · arXiv:2312.15577
-
ESGReveal: An LLM-based approach for extracting structured data from ESG reports 25 Dec 2023 · 0 repositories · arXiv:2312.17264
-
IQAGPT: Image Quality Assessment with Vision-language and ChatGPT Models 25 Dec 2023 · 0 repositories · arXiv:2312.15663
-
Lifting by Image -- Leveraging Image Cues for Accurate 3D Human Pose Estimation 25 Dec 2023 · 0 repositories · arXiv:2312.15636
-
Nighttime Person Re-Identification via Collaborative Enhancement Network with Multi-domain Learning 25 Dec 2023 · 1 repository · arXiv:2312.16246
-
Partial Fine-Tuning: A Successor to Full Fine-Tuning for Vision Transformers 25 Dec 2023 · 0 repositories · arXiv:2312.15681
-
Proximal Gradient Descent Unfolding Dense-spatial Spectral-attention Transformer for Compressive Spectral Imaging 25 Dec 2023 · 0 repositories · arXiv:2312.16237
-
UniRef++: Segment Every Reference Object in Spatial and Temporal Spaces 25 Dec 2023 · 2 repositories · arXiv:2312.15715Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Word length-aware text spotting: Enhancing detection and recognition in dense text image 25 Dec 2023 · 0 repositories · arXiv:2312.15690
-
DEAP: Design Space Exploration for DNN Accelerator Parallelism 24 Dec 2023 · 0 repositories · arXiv:2312.15388
-
Deformable Audio Transformer for Audio Event Detection 24 Dec 2023 · 0 repositories · arXiv:2312.16228
-
Diffusion-EXR: Controllable Review Generation for Explainable Recommendation via Diffusion Models 24 Dec 2023 · 0 repositories · arXiv:2312.15490
-
Fairness-Aware Structured Pruning in Transformers 24 Dec 2023 · 1 repository · arXiv:2312.15398Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Multi-level biomedical NER through multi-granularity embeddings and enhanced labeling 24 Dec 2023 · 0 repositories · arXiv:2312.15550
-
PointCT: Point Central Transformer Network for Weakly-supervised Point Cloud Semantic Segmentation 24 Dec 2023 · 1 repository
-
Do LLM Agents Exhibit Social Behavior? 23 Dec 2023 · 0 repositories · arXiv:2312.15198
-
Enhancing User Intent Capture in Session-Based Recommendation with Attribute Patterns 23 Dec 2023 · 1 repository · arXiv:2312.16199Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
GestaltMML: Enhancing Rare Genetic Disease Diagnosis through Multimodal Machine Learning Combining Facial Images and Clinical Texts 23 Dec 2023 · 2 repositories · arXiv:2312.15320
-
Narrowing the semantic gaps in U-Net with learnable skip connections: The case of medical image segmentation 23 Dec 2023 · 3 repositories · arXiv:2312.15182
-
Paralinguistics-Enhanced Large Language Modeling of Spoken Dialogue 23 Dec 2023 · 0 repositories · arXiv:2312.15316
-
Understanding the Potential of FPGA-Based Spatial Acceleration for Large Language Model Inference 23 Dec 2023 · 1 repository · arXiv:2312.15159
-
Context Enhanced Transformer for Single Image Object Detection 22 Dec 2023 · 0 repositories · arXiv:2312.14492
-
Efficacy of Machine-Generated Instructions 22 Dec 2023 · 0 repositories · arXiv:2312.14423
-
Personalized Large Language Model Assistant with Evolving Conditional Memory 22 Dec 2023 · 0 repositories · arXiv:2312.17257
-
FM-OV3D: Foundation Model-based Cross-modal Knowledge Blending for Open-Vocabulary 3D Detection 22 Dec 2023 · 0 repositories · arXiv:2312.14465
-
Generative Pretraining at Scale: Transformer-Based Encoding of Transactional Behavior for Fraud Detection 22 Dec 2023 · 0 repositories · arXiv:2312.14406
-
Global Occlusion-Aware Transformer for Robust Stereo Matching 22 Dec 2023 · 1 repository · arXiv:2312.14650
-
Large Language Model (LLM) Bias Index -- LLMBI 22 Dec 2023 · 0 repositories · arXiv:2312.14769
-
MMGPL: Multimodal Medical Data Analysis with Graph Prompt Learning 22 Dec 2023 · 0 repositories · arXiv:2312.14574
-
Numerical Reasoning for Financial Reports 22 Dec 2023 · 1 repository · arXiv:2312.14870
-
QUAR-VLA: Vision-Language-Action Model for Quadruped Robots 22 Dec 2023 · 0 repositories · arXiv:2312.14457
-
Refining GPT-3 Embeddings with a Siamese Structure for Technical Post Duplicate Detection 22 Dec 2023 · 1 repository · arXiv:2312.15068
-
SCUNet++: Swin-UNet and CNN Bottleneck Hybrid Architecture with Multi-Fusion Dense Skip Connection for Pulmonary Embolism CT Image Segmentation 22 Dec 2023 · 1 repository · arXiv:2312.14705
-
Spatiotemporal-Linear: Towards Universal Multivariate Time Series Forecasting 22 Dec 2023 · 0 repositories · arXiv:2312.14869
-
Theory of Hallucinations based on Equivariance 22 Dec 2023 · 0 repositories · arXiv:2312.14504
-
Towards a Unified Multimodal Reasoning Framework 22 Dec 2023 · 1 repository · arXiv:2312.15021
-
Towards Detecting Cascades of Biased Medical Claims on Twitter 22 Dec 2023 · 0 repositories · arXiv:2312.15040
-
TPTNet: A Data-Driven Temperature Prediction Model Based on Turbulent Potential Temperature 22 Dec 2023 · 0 repositories · arXiv:2312.14980
-
How Smooth Is Attention? 22 Dec 2023 · 0 repositories · arXiv:2312.14820
-
Unsupervised Auditory and Semantic Entrainment Models with Deep Neural Networks 22 Dec 2023 · 0 repositories · arXiv:2312.15098
-
ViStripformer: A Token-Efficient Transformer for Versatile Video Restoration 22 Dec 2023 · 1 repository · arXiv:2312.14502
-
Voila-A: Aligning Vision-Language Models with User's Gaze Attention 22 Dec 2023 · 0 repositories · arXiv:2401.09454
-
Anchoring Path for Inductive Relation Prediction in Knowledge Graphs 21 Dec 2023 · 1 repository · arXiv:2312.13596
-
Argue with Me Tersely: Towards Sentence-Level Counter-Argument Generation 21 Dec 2023 · 1 repository · arXiv:2312.13608
-
ChatGPT as a commenter to the news: can LLMs generate human-like opinions? 21 Dec 2023 · 1 repository · arXiv:2312.13961
-
CR-SAM: Curvature Regularized Sharpness-Aware Minimization 21 Dec 2023 · 1 repository · arXiv:2312.13555Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Critic-Guided Decision Transformer for Offline Reinforcement Learning 21 Dec 2023 · 1 repository · arXiv:2312.13716Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
D-STGCNT: A Dense Spatio-Temporal Graph Conv-GRU Network based on transformer for assessment of patient physical rehabilitation 21 Dec 2023 · 0 repositories · arXiv:2401.06150
-
De novo Drug Design using Reinforcement Learning with Multiple GPT Agents 21 Dec 2023 · 2 repositories · arXiv:2401.06155Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
DECO: Query-Based End-to-End Object Detection with ConvNets 21 Dec 2023 · 3 repositories · arXiv:2312.13735Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
DUSt3R: Geometric 3D Vision Made Easy 21 Dec 2023 · 2 repositories · arXiv:2312.14132Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Exploiting Novel GPT-4 APIs 21 Dec 2023 · 1 repository · arXiv:2312.14302
-
Free-Editor: Zero-shot Text-driven 3D Scene Editing 21 Dec 2023 · 1 repository · arXiv:2312.13663
-
How to Prune Your Language Model: Recovering Accuracy on the "Sparsity May Cry'' Benchmark 21 Dec 2023 · 0 repositories · arXiv:2312.13547
-
HW-V2W-Map: Hardware Vulnerability to Weakness Mapping Framework for Root Cause Analysis with GPT-assisted Mitigation Suggestion 21 Dec 2023 · 1 repository · arXiv:2312.13530
-
InfoVisDial: An Informative Visual Dialogue Dataset by Bridging Large Multimodal and Language Models 21 Dec 2023 · 0 repositories · arXiv:2312.13503
-
LiDAR-LLM: Exploring the Potential of Large Language Models for 3D LiDAR Understanding 21 Dec 2023 · 0 repositories · arXiv:2312.14074
-
LingoQA: Visual Question Answering for Autonomous Driving 21 Dec 2023 · 2 repositories · arXiv:2312.14115Syntology community repositories only · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
TraceFL: Interpretability-Driven Debugging in Federated Learning via Neuron Provenance 21 Dec 2023 · 2 repositories · arXiv:2312.13632
-
Shai: A large language model for asset management 21 Dec 2023 · 0 repositories · arXiv:2312.14203
-
Team Irisapu Project Description for DRC2023 21 Dec 2023 · 0 repositories · arXiv:2312.13765
-
Typhoon: Thai Large Language Models 21 Dec 2023 · 0 repositories · arXiv:2312.13951
-
Understanding Inter-Session Intentions via Complex Logical Reasoning 21 Dec 2023 · 1 repository · arXiv:2312.13866
-
Preparing to Integrate Generative Pretrained Transformer Series 4 models into Genetic Variant Assessment Workflows: Assessing Performance, Drift, and Nondeterminism Characteristics Relative to Classifying Functional Evidence in Literature 21 Dec 2023 · 0 repositories · arXiv:2312.13521
-
VideoPoet: A Large Language Model for Zero-Shot Video Generation 21 Dec 2023 · 0 repositories · arXiv:2312.14125
-
AgentCoder: Multi-Agent-based Code Generation with Iterative Testing and Optimisation 20 Dec 2023 · 1 repository · arXiv:2312.13010Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Benchmarking and Analyzing In-context Learning, Fine-tuning and Supervised Learning for Biomedical Knowledge Curation: a focused study on chemical entities of biological interest 20 Dec 2023 · 0 repositories · arXiv:2312.12989
-
Cached Transformers: Improving Transformers with Differentiable Memory Cache 20 Dec 2023 · 1 repository · arXiv:2312.12742
-
CST-former: Transformer with Channel-Spectro-Temporal Attention for Sound Event Localization and Detection 20 Dec 2023 · 0 repositories · arXiv:2312.12821
-
MonoCoder: Domain-Specific Code Language Model for HPC Codes and Tasks 20 Dec 2023 · 3 repositories · arXiv:2312.13322
-
EPNet: An Efficient Pyramid Network for Enhanced Single-Image Super-Resolution with Reduced Computational Requirements 20 Dec 2023 · 0 repositories · arXiv:2312.13396
-
Exploring the potential of channel interactions for image restoration 20 Dec 2023 · 1 repository
-
In2SET: Intra-Inter Similarity Exploiting Transformer for Dual-Camera Compressive Hyperspectral Imaging 20 Dec 2023 · 1 repository · arXiv:2312.13319
-
Learning Exhaustive Correlation for Spectral Super-Resolution: Where Spatial-Spectral Attention Meets Linear Dependence 20 Dec 2023 · 0 repositories · arXiv:2312.12833
-
Lookahead: An Inference Acceleration Framework for Large Language Model with Lossless Generation Accuracy 20 Dec 2023 · 1 repository · arXiv:2312.12728Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Pixel-to-Abundance Translation: Conditional Generative Adversarial Networks Based on Patch Transformer for Hyperspectral Unmixing 20 Dec 2023 · 0 repositories · arXiv:2312.13127
-
Scaling Down to Scale Up: A Cost-Benefit Analysis of Replacing OpenAI's LLM with Open Source SLMs in Production 20 Dec 2023 · 1 repository · arXiv:2312.14972
-
An Empirical study of Unsupervised Neural Machine Translation: analyzing NMT output, model's behavior and sentences' contribution 19 Dec 2023 · 0 repositories · arXiv:2312.12588
-
Can ChatGPT be Your Personal Medical Assistant? 19 Dec 2023 · 0 repositories · arXiv:2312.12006
-
Can Transformers Learn Sequential Function Classes In Context? 19 Dec 2023 · 0 repositories · arXiv:2312.12655
-
Context Disentangling and Prototype Inheriting for Robust Visual Grounding 19 Dec 2023 · 1 repository · arXiv:2312.11967
-
Self-Admitted Technical Debt Detection Approaches: A Decade Systematic Review 19 Dec 2023 · 1 repository · arXiv:2312.15020
-
Efficient Title Reranker for Fast and Improved Knowledge-Intense NLP 19 Dec 2023 · 0 repositories · arXiv:2312.12430
-
External Knowledge Augmented Polyphone Disambiguation Using Large Language Model 19 Dec 2023 · 0 repositories · arXiv:2312.11920