Methods › General › Attention Mechanisms › Attention › Papers, page 107
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 107 of 316: papers 10,601 to 10,700 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
MambaEVT: Event Stream based Visual Object Tracking using State Space Model 20 Aug 2024 · 1 repository · arXiv:2408.10487Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples)
-
Multichannel Attention Networks with Ensembled Transfer Learning to Recognize Bangla Handwritten Charecter 20 Aug 2024 · 0 repositories · arXiv:2408.10955
-
Navigating Spatio-Temporal Heterogeneity: A Graph Transformer Approach for Traffic Forecasting 20 Aug 2024 · 1 repository · arXiv:2408.10822
-
Hierarchical Attention Diffusion Networks with Object Priors for Video Change Detection 20 Aug 2024 · 0 repositories · arXiv:2408.10619
-
On the Potential of Open-Vocabulary Models for Object Detection in Unusual Street Scenes 20 Aug 2024 · 0 repositories · arXiv:2408.11221
-
Open-FinLLMs: Open Multimodal Large Language Models for Financial Applications 20 Aug 2024 · 0 repositories · arXiv:2408.11878
-
Optimization of Multi-Agent Flying Sidekick Traveling Salesman Problem over Road Networks 20 Aug 2024 · 0 repositories · arXiv:2408.11187
-
Out-of-Distribution Detection with Attention Head Masking for Multimodal Document Classification 20 Aug 2024 · 1 repository · arXiv:2408.11237
-
Perception-guided Jailbreak against Text-to-Image Models 20 Aug 2024 · 0 repositories · arXiv:2408.10848
-
PRformer: Pyramidal Recurrent Transformer for Multivariate Time Series Forecasting 20 Aug 2024 · 1 repository · arXiv:2408.10483Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Prompt Your Brain: Scaffold Prompt Tuning for Efficient Adaptation of fMRI Pre-trained Model 20 Aug 2024 · 0 repositories · arXiv:2408.10567
-
Quantum Inverse Contextual Vision Transformers (Q-ICVT): A New Frontier in 3D Object Detection for AVs 20 Aug 2024 · 1 repository · arXiv:2408.11207
-
Reading with Intent 20 Aug 2024 · 0 repositories · arXiv:2408.11189
-
Reconciling Methodological Paradigms: Employing Large Language Models as Novice Qualitative Research Assistants in Talent Management Research 20 Aug 2024 · 0 repositories · arXiv:2408.11043
-
Rethinking Video Segmentation with Masked Video Consistency: Did the Model Learn as Intended? 20 Aug 2024 · 0 repositories · arXiv:2408.10627
-
Revisiting VerilogEval: A Year of Improvements in Large-Language Models for Hardware Code Generation 20 Aug 2024 · 1 repository · arXiv:2408.11053Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SDI-Net: Toward Sufficient Dual-View Interaction for Low-light Stereo Image Enhancement 20 Aug 2024 · 0 repositories · arXiv:2408.10934
-
Soda-Eval: Open-Domain Dialogue Evaluation in the age of LLMs 20 Aug 2024 · 1 repository · arXiv:2408.10902
-
GACL: Graph Attention Collaborative Learning for Temporal QoS Prediction 20 Aug 2024 · 0 repositories · arXiv:2408.10555
-
TDS-CLIP: Temporal Difference Side Network for Image-to-Video Transfer Learning 20 Aug 2024 · 1 repository · arXiv:2408.10688
-
Towardseffective teaching assistants: From intent-based chatbots to LLM-poweredteachingassistants 20 Aug 2024 · 0 repositories
-
Tracing Privacy Leakage of Language Models to Training Data via Adjusted Influence Functions 20 Aug 2024 · 0 repositories · arXiv:2408.10468
-
Trustworthy Compression? Impact of AI-based Codecs on Biometrics for Law Enforcement 20 Aug 2024 · 0 repositories · arXiv:2408.10823
-
UIE-UnFold: Deep Unfolding Network with Color Priors and Vision Transformer for Underwater Image Enhancement 20 Aug 2024 · 1 repository · arXiv:2408.10653
-
Security Attacks on LLM-based Code Completion Tools 20 Aug 2024 · 1 repository · arXiv:2408.11006
-
A More Accurate Approximation of Activation Function with Few Spikes Neurons 19 Aug 2024 · 0 repositories · arXiv:2409.00044
-
A Strategy to Combine 1stGen Transformers and Open LLMs for Automatic Text Classification 19 Aug 2024 · 0 repositories · arXiv:2408.09629
-
Acquiring Bidirectionality via Large and Small Language Models 19 Aug 2024 · 1 repository · arXiv:2408.09640
-
Active Learning for Identifying Disaster-Related Tweets: A Comparison with Keyword Filtering and Generic Fine-Tuning 19 Aug 2024 · 0 repositories · arXiv:2408.09914
-
Attention is a smoothed cubic spline 19 Aug 2024 · 0 repositories · arXiv:2408.09624
-
Large Language Models for Classical Chinese Poetry Translation: Benchmarking, Evaluating, and Improving 19 Aug 2024 · 0 repositories · arXiv:2408.09945
-
Caption-Driven Explorations: Aligning Image and Text Embeddings through Human-Inspired Foveated Vision 19 Aug 2024 · 0 repositories · arXiv:2408.09948
-
Carbon Footprint Accounting Driven by Large Language Models and Retrieval-augmented Generation 19 Aug 2024 · 0 repositories · arXiv:2408.09713
-
Coarse-Fine View Attention Alignment-Based GAN for CT Reconstruction from Biplanar X-Rays 19 Aug 2024 · 0 repositories · arXiv:2408.09736
-
Edge-Cloud Collaborative Motion Planning for Autonomous Driving with Large Language Models 19 Aug 2024 · 0 repositories · arXiv:2408.09972
-
ELDER: Enhancing Lifelong Model Editing with Mixture-of-LoRA 19 Aug 2024 · 1 repository · arXiv:2408.11869
-
Enhanced document retrieval with topic embeddings 19 Aug 2024 · 0 repositories · arXiv:2408.10435
-
Enhancing Reinforcement Learning Through Guided Search 19 Aug 2024 · 0 repositories · arXiv:2408.10113
-
Event Stream based Human Action Recognition: A High-Definition Benchmark Dataset and Algorithms 19 Aug 2024 · 1 repository · arXiv:2408.09764
-
Exploiting Fine-Grained Prototype Distribution for Boosting Unsupervised Class Incremental Learning 19 Aug 2024 · 0 repositories · arXiv:2408.10046
-
Factorized-Dreamer: Training A High-Quality Video Generator with Limited and Low-Quality Data 19 Aug 2024 · 0 repositories · arXiv:2408.10119
-
Faster Adaptive Decentralized Learning Algorithms 19 Aug 2024 · 0 repositories · arXiv:2408.09775
-
Federated Frank-Wolfe Algorithm 19 Aug 2024 · 1 repository · arXiv:2408.10090
-
Goldfish: Monolingual Language Models for 350 Languages 19 Aug 2024 · 1 repository · arXiv:2408.10441Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
GARLIC: GPT-Augmented Reinforcement Learning with Intelligent Control for Vehicle Dispatching 19 Aug 2024 · 0 repositories · arXiv:2408.10286
-
Harmonizing Attention: Training-free Texture-aware Geometry Transfer 19 Aug 2024 · 0 repositories · arXiv:2408.10846
-
Image-based Freeform Handwriting Authentication with Energy-oriented Self-Supervised Learning 19 Aug 2024 · 0 repositories · arXiv:2408.09676
-
Instruction-Based Molecular Graph Generation with Unified Text-Graph Diffusion Model 19 Aug 2024 · 1 repository · arXiv:2408.09896
-
Learning Precise Affordances from Egocentric Videos for Robotic Manipulation 19 Aug 2024 · 0 repositories · arXiv:2408.10123
-
LegalBench-RAG: A Benchmark for Retrieval-Augmented Generation in the Legal Domain 19 Aug 2024 · 1 repository · arXiv:2408.10343Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
LightWeather: Harnessing Absolute Positional Encoding to Efficient and Scalable Global Weather Forecasting 19 Aug 2024 · 0 repositories · arXiv:2408.09695
-
MoDeGPT: Modular Decomposition for Large Language Model Compression 19 Aug 2024 · 0 repositories · arXiv:2408.09632
-
Multi-Scale Representation Learning for Image Restoration with State-Space Model 19 Aug 2024 · 0 repositories · arXiv:2408.10145
-
Mutually-Aware Feature Learning for Few-Shot Object Counting 19 Aug 2024 · 0 repositories · arXiv:2408.09734
-
Pedestrian Attribute Recognition: A New Benchmark Dataset and A Large Language Model Augmented Framework 19 Aug 2024 · 2 repositories · arXiv:2408.09720
-
PLUTUS: A Well Pre-trained Large Unified Transformer can Unveil Financial Time Series Regularities 19 Aug 2024 · 0 repositories · arXiv:2408.10111
-
Preoperative Rotator Cuff Tear Prediction from Shoulder Radiographs using a Convolutional Block Attention Module-Integrated Neural Network 19 Aug 2024 · 0 repositories · arXiv:2408.09894
-
Privacy Checklist: Privacy Violation Detection Grounding on Contextual Integrity Theory 19 Aug 2024 · 0 repositories · arXiv:2408.10053
-
Propagating the prior from shallow to deep with a pre-trained velocity-model Generative Transformer network 19 Aug 2024 · 0 repositories · arXiv:2408.09767
-
R2GenCSR: Retrieving Context Samples for Large Language Model based X-ray Medical Report Generation 19 Aug 2024 · 1 repository · arXiv:2408.09743
-
Revisiting Reciprocal Recommender Systems: Metrics, Formulation, and Method 19 Aug 2024 · 1 repository · arXiv:2408.09748
-
Rhyme-aware Chinese lyric generator based on GPT 19 Aug 2024 · 0 repositories · arXiv:2408.10130
-
SAM-UNet:Enhancing Zero-Shot Segmentation of SAM for Universal Medical Images 19 Aug 2024 · 1 repository · arXiv:2408.09886
-
Self-Directed Turing Test for Large Language Models 19 Aug 2024 · 0 repositories · arXiv:2408.09853
-
Sliced Maximal Information Coefficient: A Training-Free Approach for Image Quality Assessment Enhancement 19 Aug 2024 · 1 repository · arXiv:2408.09920
-
Characteristic Performance Study on Solving Oscillator ODEs via Soft-constrained Physics-informed Neural Network with Small Data 19 Aug 2024 · 1 repository · arXiv:2408.11077
-
sTransformer: A Modular Approach for Extracting Inter-Sequential and Temporal Information for Time-Series Forecasting 19 Aug 2024 · 0 repositories · arXiv:2408.09723
-
SSDTrain: An Activation Offloading Framework to SSDs for Faster Large Language Model Training 19 Aug 2024 · 0 repositories · arXiv:2408.10013
-
Toward Large-scale Spiking Neural Networks: A Comprehensive Survey and Future Directions 19 Aug 2024 · 0 repositories · arXiv:2409.02111
-
Transformers to SSMs: Distilling Quadratic Knowledge to Subquadratic Models 19 Aug 2024 · 1 repository · arXiv:2408.10189
-
A Unified Framework for Interpretable Transformers Using PDEs and Information Theory 18 Aug 2024 · 0 repositories · arXiv:2408.09523
-
Advances in Multiple Instance Learning for Whole Slide Image Analysis: Techniques, Challenges, and Future Directions 18 Aug 2024 · 0 repositories · arXiv:2408.09476
-
Agentic Retrieval-Augmented Generation for Time Series Analysis 18 Aug 2024 · 0 repositories · arXiv:2408.14484
-
An Introduction to Cognidynamics 18 Aug 2024 · 0 repositories · arXiv:2408.13112
-
Comparison between the Structures of Word Co-occurrence and Word Similarity Networks for Ill-formed and Well-formed Texts in Taiwan Mandarin 18 Aug 2024 · 0 repositories · arXiv:2408.09404
-
Deep Limit Model-free Prediction in Regression 18 Aug 2024 · 0 repositories · arXiv:2408.09532
-
ELASTIC: Efficient Linear Attention for Sequential Interest Compression 18 Aug 2024 · 0 repositories · arXiv:2408.09380
-
FD2Talk: Towards Generalized Talking Head Generation with Facial Decoupled Diffusion Model 18 Aug 2024 · 0 repositories · arXiv:2408.09384
-
Enhanced BPINN Training Convergence in Solving General and Multi-scale Elliptic PDEs with Noise 18 Aug 2024 · 0 repositories · arXiv:2408.09340
-
OU-CoViT: Copula-Enhanced Bi-Channel Multi-Task Vision Transformers with Dual Adaptation for OU-UWF Images 18 Aug 2024 · 0 repositories · arXiv:2408.09395
-
Out-of-distribution generalization via composition: a lens through induction heads in Transformers 18 Aug 2024 · 1 repository · arXiv:2408.09503Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Threshold Filtering Packing for Supervised Fine-Tuning: Training Related Samples within Packs 18 Aug 2024 · 0 repositories · arXiv:2408.09327
-
ConVerSum: A Contrastive Learning-based Approach for Data-Scarce Solution of Cross-Lingual Summarization Beyond Direct Equivalents 17 Aug 2024 · 0 repositories · arXiv:2408.09273
-
Cross-Species Data Integration for Enhanced Layer Segmentation in Kidney Pathology 17 Aug 2024 · 1 repository · arXiv:2408.09278
-
EEG-SCMM: Soft Contrastive Masked Modeling for Cross-Corpus EEG-Based Emotion Recognition 17 Aug 2024 · 0 repositories · arXiv:2408.09186
-
Graph Classification with GNNs: Optimisation, Representation and Inductive Bias 17 Aug 2024 · 1 repository · arXiv:2408.09266
-
HybridOcc: NeRF Enhanced Transformer-based Multi-Camera 3D Occupancy Prediction 17 Aug 2024 · 0 repositories · arXiv:2408.09104
-
Improving Rare Word Translation With Dictionaries and Attention Masking 17 Aug 2024 · 1 repository · arXiv:2408.09075
-
Linear Attention is Enough in Spatial-Temporal Forecasting 17 Aug 2024 · 1 repository · arXiv:2408.09158
-
MagicID: Flexible ID Fidelity Generation System 17 Aug 2024 · 0 repositories · arXiv:2408.09248
-
MaskBEV: Towards A Unified Framework for BEV Detection and Map Segmentation 17 Aug 2024 · 0 repositories · arXiv:2408.09122
-
On the Improvement of Generalization and Stability of Forward-Only Learning via Neural Polarization 17 Aug 2024 · 1 repository · arXiv:2408.09210
-
On the KL-Divergence-based Robust Satisficing Model 17 Aug 2024 · 0 repositories · arXiv:2408.09157
-
PADetBench: Towards Benchmarking Physical Attacks against Object Detection 17 Aug 2024 · 2 repositories · arXiv:2408.09181
-
Quality Assessment in the Era of Large Models: A Survey 17 Aug 2024 · 0 repositories · arXiv:2409.00031
-
Selective Prompt Anchoring for Code Generation 17 Aug 2024 · 1 repository · arXiv:2408.09121Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 1 pointer-only (licence)
-
Sentiment analysis of preservice teachers' reflections using a large language model 17 Aug 2024 · 0 repositories · arXiv:2408.11862
-
Siamese Multiple Attention Temporal Convolution Networks for Human Mobility Signature Identification 17 Aug 2024 · 0 repositories · arXiv:2408.09230
-
TableBench: A Comprehensive and Complex Benchmark for Table Question Answering 17 Aug 2024 · 0 repositories · arXiv:2408.09174
-
TC-RAG:Turing-Complete RAG's Case study on Medical LLM Systems 17 Aug 2024 · 2 repositories · arXiv:2408.09199