Methods › General › Attention Mechanisms › Attention › Papers, page 60
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 60 of 316: papers 5,901 to 6,000 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Reversed in Time: A Novel Temporal-Emphasized Benchmark for Cross-Modal Video-Text Retrieval 26 Dec 2024 · 1 repository · arXiv:2412.19178
-
Sentiment trading with large language models 26 Dec 2024 · 0 repositories · arXiv:2412.19245
-
SpectralKD: A Unified Framework for Interpreting and Distilling Vision Transformers via Spectral Analysis 26 Dec 2024 · 1 repository · arXiv:2412.19055
-
To Predict or Not To Predict? Proportionally Masked Autoencoders for Tabular Data Imputation 26 Dec 2024 · 1 repository · arXiv:2412.19152
-
Transformer-Based Wireless Capsule Endoscopy Bleeding Tissue Detection and Classification 26 Dec 2024 · 1 repository · arXiv:2412.19218
-
Accelerating Diffusion Transformers with Dual Feature Caching 25 Dec 2024 · 3 repositories · arXiv:2412.18911Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Automatic Self-supervised Learning for Social Recommendations 25 Dec 2024 · 0 repositories · arXiv:2412.18735
-
Adopting Trustworthy AI for Sleep Disorder Prediction: Deep Time Series Analysis with Temporal Attention Mechanism and Counterfactual Explanations 25 Dec 2024 · 0 repositories · arXiv:2412.18971
-
An Attentive Dual-Encoder Framework Leveraging Multimodal Visual and Semantic Information for Automatic OSAHS Diagnosis 25 Dec 2024 · 1 repository · arXiv:2412.18919
-
BCR-Net: Boundary-Category Refinement Network for Weakly Semi-Supervised X-Ray Prohibited Item Detection with Points 25 Dec 2024 · 0 repositories · arXiv:2412.18918
-
Conditional Balance: Improving Multi-Conditioning Trade-Offs in Image Generation 25 Dec 2024 · 0 repositories · arXiv:2412.19853
-
DCIS: Efficient Length Extrapolation of LLMs via Divide-and-Conquer Scaling Factor Search 25 Dec 2024 · 1 repository · arXiv:2412.18811
-
Distortion-Aware Adversarial Attacks on Bounding Boxes of Object Detectors 25 Dec 2024 · 1 repository · arXiv:2412.18815
-
Don't Lose Yourself: Boosting Multimodal Recommendation via Reducing Node-neighbor Discrepancy in Graph Convolutional Network 25 Dec 2024 · 0 repositories · arXiv:2412.18962
-
EC-Diffuser: Multi-Object Manipulation via Entity-Centric Behavior Generation 25 Dec 2024 · 0 repositories · arXiv:2412.18907
-
Evaluating deep learning models for fault diagnosis of a rotating machinery with epistemic and aleatoric uncertainty 25 Dec 2024 · 0 repositories · arXiv:2412.18980
-
Evaluating the Adversarial Robustness of Detection Transformers 25 Dec 2024 · 0 repositories · arXiv:2412.18718
-
HAND: Hierarchical Attention Network for Multi-Scale Handwritten Document Recognition and Layout Analysis 25 Dec 2024 · 0 repositories · arXiv:2412.18981
-
Implicit factorized transformer approach to fast prediction of turbulent channel flows 25 Dec 2024 · 1 repository · arXiv:2412.18840
-
Injecting Bias into Text Classification Models using Backdoor Attacks 25 Dec 2024 · 0 repositories · arXiv:2412.18975
-
Ister: Inverted Seasonal-Trend Decomposition Transformer for Explainable Multivariate Time Series Forecasting 25 Dec 2024 · 0 repositories · arXiv:2412.18798
-
ModelGrow: Continual Text-to-Video Pre-training with Model Expansion and Language Understanding Enhancement 25 Dec 2024 · 0 repositories · arXiv:2412.18966
-
MTCAE-DFER: Multi-Task Cascaded Autoencoder for Dynamic Facial Expression Recognition 25 Dec 2024 · 1 repository · arXiv:2412.18988
-
ObitoNet: Multimodal High-Resolution Point Cloud Reconstruction 25 Dec 2024 · 1 repository · arXiv:2412.18775
-
On the Robustness of Generative Information Retrieval Models 25 Dec 2024 · 1 repository · arXiv:2412.18768
-
Open-Vocabulary Panoptic Segmentation Using BERT Pre-Training of Vision-Language Multiway Transformer Model 25 Dec 2024 · 1 repository · arXiv:2412.18917
-
Optimization and Scalability of Collaborative Filtering Algorithms in Large Language Models 25 Dec 2024 · 0 repositories · arXiv:2412.18715
-
Optimizing Large Language Models with an Enhanced LoRA Fine-Tuning Algorithm for Efficiency and Robustness in NLP Tasks 25 Dec 2024 · 0 repositories · arXiv:2412.18729
-
Position-aware Graph Transformer for Recommendation 25 Dec 2024 · 0 repositories · arXiv:2412.18731
-
Predicting Time Series of Networked Dynamical Systems without Knowing Topology 25 Dec 2024 · 1 repository · arXiv:2412.18734
-
Resource-Efficient Transformer Architecture: Optimizing Memory and Execution Time for Real-Time Applications 25 Dec 2024 · 0 repositories · arXiv:2501.00042
-
SAFLITE: Fuzzing Autonomous Systems via Large Language Models 25 Dec 2024 · 0 repositories · arXiv:2412.18727
-
TopoBDA: Towards Bezier Deformable Attention for Road Topology Understanding 25 Dec 2024 · 0 repositories · arXiv:2412.18951
-
Towards Expressive Video Dubbing with Multiscale Multimodal Context Interaction 25 Dec 2024 · 0 repositories · arXiv:2412.18748
-
TPCH: Tensor-interacted Projection and Cooperative Hashing for Multi-view Clustering 25 Dec 2024 · 1 repository · arXiv:2412.18847
-
UNIC-Adapter: Unified Image-instruction Adapter with Multi-modal Transformer for Image Generation 25 Dec 2024 · 0 repositories · arXiv:2412.18928
-
Unified Local and Global Attention Interaction Modeling for Vision Transformers 25 Dec 2024 · 0 repositories · arXiv:2412.18778
-
Using Large Language Models for Automated Grading of Student Writing about Science 25 Dec 2024 · 0 repositories · arXiv:2412.18719
-
WeatherGS: 3D Scene Reconstruction in Adverse Weather Conditions via Gaussian Splatting 25 Dec 2024 · 1 repository · arXiv:2412.18862
-
Whose Morality Do They Speak? Unraveling Cultural Bias in Multilingual Language Models 25 Dec 2024 · 0 repositories · arXiv:2412.18863
-
3DEnhancer: Consistent Multi-View Diffusion for 3D Enhancement 24 Dec 2024 · 0 repositories · arXiv:2412.18565
-
Advancing Deformable Medical Image Registration with Multi-axis Cross-covariance Attention 24 Dec 2024 · 1 repository · arXiv:2412.18545
-
Advancing Explainability in Neural Machine Translation: Analytical Metrics for Attention and Alignment Consistency 24 Dec 2024 · 0 repositories · arXiv:2412.18669
-
AgreeMate: Teaching LLMs to Haggle 24 Dec 2024 · 1 repository · arXiv:2412.18690
-
AutoSculpt: A Pattern-based Model Auto-pruning Framework Using Reinforcement Learning and Graph Learning 24 Dec 2024 · 0 repositories · arXiv:2412.18091
-
ClassifyViStA:WCE Classification with Visual understanding through Segmentation and Attention 24 Dec 2024 · 1 repository · arXiv:2412.18591
-
Comprehensive Assessment of BERT-Based Methods for Predicting Antimicrobial Peptides 24 Dec 2024 · 1 repository
-
Contrastive Representation for Interactive Recommendation 24 Dec 2024 · 1 repository · arXiv:2412.18396
-
Decentralized Intelligence in GameFi: Embodied AI Agents and the Convergence of DeFi and Virtual Ecosystems 24 Dec 2024 · 1 repository · arXiv:2412.18601
-
DiTCtrl: Exploring Attention Control in Multi-Modal Diffusion Transformer for Tuning-Free Multi-Prompt Longer Video Generation 24 Dec 2024 · 1 repository · arXiv:2412.18597Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Do Language Models Understand the Cognitive Tasks Given to Them? Investigations with the N-Back Paradigm 24 Dec 2024 · 0 repositories · arXiv:2412.18120
-
ERVD: An Efficient and Robust ViT-Based Distillation Framework for Remote Sensing Image Retrieval 24 Dec 2024 · 1 repository · arXiv:2412.18136
-
EvoPat: A Multi-LLM-based Patents Summarization and Analysis Agent 24 Dec 2024 · 0 repositories · arXiv:2412.18100
-
Fashionability-Enhancing Outfit Image Editing with Conditional Diffusion Models 24 Dec 2024 · 0 repositories · arXiv:2412.18421
-
GeAR: Graph-enhanced Agent for Retrieval-augmented Generation 24 Dec 2024 · 0 repositories · arXiv:2412.18431
-
Harnessing Large Language Models for Knowledge Graph Question Answering via Adaptive Multi-Aspect Retrieval-Augmentation 24 Dec 2024 · 1 repository · arXiv:2412.18537Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
HTR-JAND: Handwritten Text Recognition with Joint Attention Network and Knowledge Distillation 24 Dec 2024 · 0 repositories · arXiv:2412.18524
-
Improving Factuality with Explicit Working Memory 24 Dec 2024 · 0 repositories · arXiv:2412.18069
-
Leveraging Convolutional Neural Network-Transformer Synergy for Predictive Modeling in Risk-Based Applications 24 Dec 2024 · 0 repositories · arXiv:2412.18222
-
Leveraging Deep Learning with Multi-Head Attention for Accurate Extraction of Medicine from Handwritten Prescriptions 24 Dec 2024 · 0 repositories · arXiv:2412.18199
-
Molly: Making Large Language Model Agents Solve Python Problem More Logically 24 Dec 2024 · 0 repositories · arXiv:2412.18093
-
Multi-View Fusion Neural Network for Traffic Demand Prediction 24 Dec 2024 · 0 repositories · arXiv:2412.19839
-
Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English 24 Dec 2024 · 1 repository · arXiv:2412.18415
-
Multimodal joint prediction of traffic spatial-temporal data with graph sparse attention mechanism and bidirectional temporal convolutional network 24 Dec 2024 · 0 repositories · arXiv:2412.19842
-
Pirates of the RAG: Adaptively Attacking LLMs to Leak Knowledge Bases 24 Dec 2024 · 0 repositories · arXiv:2412.18295
-
Re-assessing ImageNet: How aligned is its single-label assumption with its multi-label nature? 24 Dec 2024 · 0 repositories · arXiv:2412.18409
-
Research on the Proximity Relationships of Psychosomatic Disease Knowledge Graph Modules Extracted by Large Language Models 24 Dec 2024 · 0 repositories · arXiv:2412.18419
-
SDM-Car: A Dataset for Small and Dim Moving Vehicles Detection in Satellite Videos 24 Dec 2024 · 1 repository · arXiv:2412.18214
-
Segment-Based Attention Masking for GPTs 24 Dec 2024 · 1 repository · arXiv:2412.18487
-
Semi-supervised Credit Card Fraud Detection via Attribute-Driven Graph Representation 24 Dec 2024 · 2 repositories · arXiv:2412.18287
-
SlimGPT: Layer-wise Structured Pruning for Large Language Models 24 Dec 2024 · 0 repositories · arXiv:2412.18110
-
Smooth-Foley: Creating Continuous Sound for Video-to-Audio Generation Under Semantic Guidance 24 Dec 2024 · 0 repositories · arXiv:2412.18157
-
TAB: Transformer Attention Bottlenecks enable User Intervention and Debugging in Vision-Language Models 24 Dec 2024 · 1 repository · arXiv:2412.18675
-
Tackling the Dynamicity in a Production LLM Serving System with SOTA Optimizations via Hybrid Prefill/Decode/Verify Scheduling on Efficient Meta-kernels 24 Dec 2024 · 0 repositories · arXiv:2412.18106
-
TimelyLLM: Segmented LLM Serving System for Time-sensitive Robotic Applications 24 Dec 2024 · 0 repositories · arXiv:2412.18695
-
Towards understanding how attention mechanism works in deep learning 24 Dec 2024 · 0 repositories · arXiv:2412.18288
-
Underwater Image Restoration via Polymorphic Large Kernel CNNs 24 Dec 2024 · 1 repository · arXiv:2412.18459
-
Unlocking the Hidden Treasures: Enhancing Recommendations with Unlabeled Data 24 Dec 2024 · 1 repository · arXiv:2412.18170
-
Unlocking the Potential of Multiple BERT Models for Bangla Question Answering in NCTB Textbooks 24 Dec 2024 · 0 repositories · arXiv:2412.18440
-
Unveiling Visual Perception in Language Models: An Attention Head Analysis Approach 24 Dec 2024 · 0 repositories · arXiv:2412.18108
-
Video-Panda: Parameter-efficient Alignment for Encoder-free Video-Language Models 24 Dec 2024 · 1 repository · arXiv:2412.18609
-
VisionGRU: A Linear-Complexity RNN Model for Efficient Image Analysis 24 Dec 2024 · 1 repository · arXiv:2412.18178
-
A Bias-Free Training Paradigm for More General AI-generated Image Detection 23 Dec 2024 · 0 repositories · arXiv:2412.17671
-
A Coalition Game for On-demand Multi-modal 3D Automated Delivery System 23 Dec 2024 · 0 repositories · arXiv:2412.17252
-
A Silver Bullet or a Compromise for Full Attention? A Comprehensive Study of Gist Token-based Context Compression 23 Dec 2024 · 0 repositories · arXiv:2412.17483
-
A Survey of Query Optimization in Large Language Models 23 Dec 2024 · 0 repositories · arXiv:2412.17558
-
Balanced 3DGS: Gaussian-wise Parallelism Rendering with Fine-Grained Tiling 23 Dec 2024 · 0 repositories · arXiv:2412.17378
-
Cech Complex Generation with Homotopy Equivalence Framework for Myocardial Infarction Diagnosis using Electrocardiogram Signals 23 Dec 2024 · 0 repositories · arXiv:2412.17370
-
CiteBART: Learning to Generate Citations for Local Citation Recommendation 23 Dec 2024 · 1 repository · arXiv:2412.17534Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 16 harvested samples)
-
Comparative Analysis of Document-Level Embedding Methods for Similarity Scoring on Shakespeare Sonnets and Taylor Swift Lyrics 23 Dec 2024 · 0 repositories · arXiv:2412.17552
-
Comprehensive Multi-Modal Prototypes are Simple and Effective Classifiers for Vast-Vocabulary Object Detection 23 Dec 2024 · 1 repository · arXiv:2412.17800
-
Contemporary implementations of spiking bio-inspired neural networks 23 Dec 2024 · 0 repositories · arXiv:2412.17926
-
DiffFormer: a Differential Spatial-Spectral Transformer for Hyperspectral Image Classification 23 Dec 2024 · 1 repository · arXiv:2412.17350
-
Dora: Sampling and Benchmarking for 3D Shape Variational Auto-Encoders 23 Dec 2024 · 1 repository · arXiv:2412.17808Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
DreamFit: Garment-Centric Human Generation via a Lightweight Anything-Dressing Encoder 23 Dec 2024 · 1 repository · arXiv:2412.17644
-
Edge-AI for Agriculture: Lightweight Vision Models for Disease Detection in Resource-Limited Settings 23 Dec 2024 · 0 repositories · arXiv:2412.18635
-
Efficient fine-tuning methodology of text embedding models for information retrieval: contrastive learning penalty (clp) 23 Dec 2024 · 1 repository · arXiv:2412.17364
-
Enhancing Multi-Text Long Video Generation Consistency without Tuning: Time-Frequency Analysis, Prompt Alignment, and Theory 23 Dec 2024 · 0 repositories · arXiv:2412.17254
-
Fast Gradient Computation for RoPE Attention in Almost Linear Time 23 Dec 2024 · 0 repositories · arXiv:2412.17316
-
Feature Based Methods in Domain Adaptation for Object Detection: A Review Paper 23 Dec 2024 · 0 repositories · arXiv:2412.17325