Methods › General › Attention Mechanisms › Attention › Papers, page 59
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 59 of 316: papers 5,801 to 5,900 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Movable Antennas Enabled ISAC Systems: Fundamentals, Opportunities, and Future Directions 30 Dec 2024 · 0 repositories · arXiv:2412.20819
-
Open-Book Neural Algorithmic Reasoning 30 Dec 2024 · 1 repository · arXiv:2501.00072
-
Plancraft: an evaluation dataset for planning with LLM agents 30 Dec 2024 · 1 repository · arXiv:2412.21033
-
Position Information Emerges in Causal Transformers Without Positional Encodings via Similarity of Nearby Embeddings 30 Dec 2024 · 0 repositories · arXiv:2501.00073
-
RobustBlack: Challenging Black-Box Adversarial Attacks on State-of-the-Art Defenses 30 Dec 2024 · 0 repositories · arXiv:2412.20987
-
Sample Correlation for Fingerprinting Deep Face Recognition 30 Dec 2024 · 2 repositories · arXiv:2412.20768
-
Text Classification: Neural Networks VS Machine Learning Models VS Pre-trained Models 30 Dec 2024 · 0 repositories · arXiv:2412.21022
-
Towards Compatible Fine-tuning for Vision-Language Model Updates 30 Dec 2024 · 0 repositories · arXiv:2412.20895
-
UniRS: Unifying Multi-temporal Remote Sensing Tasks through Vision Language Models 30 Dec 2024 · 1 repository · arXiv:2412.20742
-
Bringing Objects to Life: 4D generation from 3D objects 29 Dec 2024 · 0 repositories · arXiv:2412.20422
-
Comparative Performance of Advanced NLP Models and LLMs in Multilingual Geo-Entity Detection 29 Dec 2024 · 0 repositories · arXiv:2412.20414
-
Cross-Modal Fusion and Attention Mechanism for Weakly Supervised Video Anomaly Detection 29 Dec 2024 · 0 repositories · arXiv:2412.20455
-
ELECTRA and GPT-4o: Cost-Effective Partners for Sentiment Analysis 29 Dec 2024 · 1 repository · arXiv:2501.00062
-
EraseAnything: Enabling Concept Erasure in Rectified Flow Transformers 29 Dec 2024 · 1 repository · arXiv:2412.20413Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 16 harvested samples) · 1 pointer-only (licence)
-
Exploiting Aggregation and Segregation of Representations for Domain Adaptive Human Pose Estimation 29 Dec 2024 · 1 repository · arXiv:2412.20538
-
FreqMixFormerV2: Lightweight Frequency-aware Mixed Transformer for Human Skeleton Action Recognition 29 Dec 2024 · 1 repository · arXiv:2412.20621
-
MATEY: multiscale adaptive foundation models for spatiotemporal physical systems 29 Dec 2024 · 0 repositories · arXiv:2412.20601
-
MR-Occ: Efficient Camera-LiDAR 3D Semantic Occupancy Prediction Using Hierarchical Multi-Resolution Voxel Representation 29 Dec 2024 · 0 repositories · arXiv:2412.20480
-
Multi-Objective Large Language Model Unlearning 29 Dec 2024 · 1 repository · arXiv:2412.20412
-
NLP-based Regulatory Compliance -- Using GPT 4.0 to Decode Regulatory Documents 29 Dec 2024 · 0 repositories · arXiv:2412.20602
-
On Adversarial Robustness of Language Models in Transfer Learning 29 Dec 2024 · 0 repositories · arXiv:2501.00066
-
Open-Sora: Democratizing Efficient Video Production for All 29 Dec 2024 · 2 repositories · arXiv:2412.20404Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Training-free Heterogeneous Model Merging 29 Dec 2024 · 1 repository · arXiv:2501.00061
-
Understanding the Impact of Confidence in Retrieval Augmented Generation: A Case Study in the Medical Domain 29 Dec 2024 · 1 repository · arXiv:2412.20309
-
Zeroth-Order Methods for Nonconvex Stochastic Problems with Decision-Dependent Distributions 29 Dec 2024 · 0 repositories · arXiv:2412.20330
-
A fuzzy rank-based ensemble of CNN models for MRI segmentation 28 Dec 2024 · 1 repository
-
Adversarial Robustness for Deep Learning-based Wildfire Prediction Models 28 Dec 2024 · 0 repositories · arXiv:2412.20006
-
An analytic theory of creativity in convolutional diffusion models 28 Dec 2024 · 0 repositories · arXiv:2412.20292
-
DDD-GenDT: Dynamic Data-driven Generative Digital Twin Framework 28 Dec 2024 · 0 repositories · arXiv:2501.00051
-
Distilled Transformers with Locally Enhanced Global Representations for Face Forgery Detection 28 Dec 2024 · 0 repositories · arXiv:2412.20156
-
Efficient Multi-Agent Collaboration with Tool Use for Online Planning in Complex Table Question Answering 28 Dec 2024 · 0 repositories · arXiv:2412.20145
-
Federated Unlearning with Gradient Descent and Conflict Mitigation 28 Dec 2024 · 1 repository · arXiv:2412.20200
-
INFELM: In-depth Fairness Evaluation of Large Text-To-Image Models 28 Dec 2024 · 0 repositories · arXiv:2501.01973
-
LoL-PIM: Long-Context LLM Decoding with Scalable DRAM-PIM System 28 Dec 2024 · 0 repositories · arXiv:2412.20166
-
MADiff: Text-Guided Fashion Image Editing with Mask Prediction and Attention-Enhanced Diffusion 28 Dec 2024 · 0 repositories · arXiv:2412.20062
-
MaIR: A Locality- and Continuity-Preserving Mamba for Image Restoration 28 Dec 2024 · 1 repository · arXiv:2412.20066
-
MAKIMA: Tuning-free Multi-Attribute Open-domain Video Editing via Mask-Guided Attention Modulation 28 Dec 2024 · 0 repositories · arXiv:2412.19978
-
On Random Sampling of Diffused Graph Signals with Sparse Inputs on Vertex Domain 28 Dec 2024 · 0 repositories · arXiv:2412.20041
-
On the Validity of Traditional Vulnerability Scoring Systems for Adversarial Attacks against LLMs 28 Dec 2024 · 0 repositories · arXiv:2412.20087
-
Real-time Calibration Model for Low-cost Sensor in Fine-grained Time series 28 Dec 2024 · 0 repositories · arXiv:2412.20170
-
SegKAN: High-Resolution Medical Image Segmentation with Long-Distance Dependencies 28 Dec 2024 · 1 repository · arXiv:2412.19990
-
ST³: Accelerating Multimodal Large Language Model by Spatial-Temporal Visual Token Trimming 28 Dec 2024 · 0 repositories · arXiv:2412.20105
-
Stable-TTS: Stable Speaker-Adaptive Text-to-Speech Synthesis via Prosody Prompting 28 Dec 2024 · 0 repositories · arXiv:2412.20155
-
Towards Ideal Temporal Graph Neural Networks: Evaluations and Conclusions after 10,000 GPU Hours 28 Dec 2024 · 0 repositories · arXiv:2412.20256
-
Transformer-Based Contrastive Meta-Learning For Low-Resource Generalizable Activity Recognition 28 Dec 2024 · 0 repositories · arXiv:2412.20290
-
Transforming CCTV cameras into NO₂ sensors at city scale for adaptive policymaking 28 Dec 2024 · 0 repositories · arXiv:2501.00056
-
VELoRA: A Low-Rank Adaptation Approach for Efficient RGB-Event based Recognition 28 Dec 2024 · 1 repository · arXiv:2412.20064
-
VisTabNet: Adapting Vision Transformers for Tabular Data 28 Dec 2024 · 1 repository · arXiv:2501.00057
-
YAD: Leveraging T5 for Improved Automatic Diacritization of Yorùbá Text 28 Dec 2024 · 1 repository · arXiv:2412.20218
-
A Survey on Large Language Model Acceleration based on KV Cache Management 27 Dec 2024 · 1 repository · arXiv:2412.19442
-
An Integrated Optimization and Deep Learning Pipeline for Predicting Live Birth Success in IVF Using Feature Optimization and Transformer-Based Models 27 Dec 2024 · 0 repositories · arXiv:2412.19696
-
Assessing Text Classification Methods for Cyberbullying Detection on Social Media Platforms 27 Dec 2024 · 0 repositories · arXiv:2412.19928
-
Can AI Help with Your Personal Finances? 27 Dec 2024 · 0 repositories · arXiv:2412.19784
-
Chimera: A Block-Based Neural Architecture Search Framework for Event-Based Object Detection 27 Dec 2024 · 0 repositories · arXiv:2412.19646
-
DeepSeek-V3 Technical Report 27 Dec 2024 · 4 repositories · arXiv:2412.19437Syntology official (archive's flag): 2 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT 27 Dec 2024 · 1 repository · arXiv:2412.19505Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Enhancing Cognitive Diagnosis by Modeling Learner Cognitive Structure State 27 Dec 2024 · 0 repositories · arXiv:2412.19759
-
Enhancing Fine-grained Image Classification through Attentive Batch Training 27 Dec 2024 · 0 repositories · arXiv:2412.19606
-
Feature Alignment-Based Knowledge Distillation for Efficient Compression of Large Language Models 27 Dec 2024 · 0 repositories · arXiv:2412.19449
-
Focusing Image Generation to Mitigate Spurious Correlations 27 Dec 2024 · 0 repositories · arXiv:2412.19457
-
Generalized Uncertainty-Based Evidential Fusion with Hybrid Multi-Head Attention for Weak-Supervised Temporal Action Localization 27 Dec 2024 · 1 repository · arXiv:2412.19418
-
Generative Pretrained Embedding and Hierarchical Irregular Time Series Representation for Daily Living Activity Recognition 27 Dec 2024 · 1 repository · arXiv:2412.19732
-
Gradient Weight-normalized Low-rank Projection for Efficient LLM Training 27 Dec 2024 · 1 repository · arXiv:2412.19616
-
Graph-attention-based Casual Discovery with Trust Region-navigated Clipping Policy Optimization 27 Dec 2024 · 0 repositories · arXiv:2412.19578
-
Hear the Scene: Audio-Enhanced Text Spotting 27 Dec 2024 · 0 repositories · arXiv:2412.19504
-
Hidformer: Transformer-Style Neural Network in Stock Price Forecasting 27 Dec 2024 · 0 repositories · arXiv:2412.19932
-
Long Context vs. RAG for LLMs: An Evaluation and Revisits 27 Dec 2024 · 1 repository · arXiv:2501.01880
-
Meta-Learning-Based Delayless Subband Adaptive Filter using Complex Self-Attention for Active Noise Control 27 Dec 2024 · 0 repositories · arXiv:2412.19471
-
MNet-SAt: A Multiscale Network with Spatial-enhanced Attention for Segmentation of Polyps in Colonoscopy 27 Dec 2024 · 0 repositories · arXiv:2412.19464
-
Multi-Condition Fault Diagnosis of Dynamic Systems: A Survey, Insights, and Prospects 27 Dec 2024 · 1 repository · arXiv:2412.19497
-
Multi-scale Latent Point Consistency Models for 3D Shape Generation 27 Dec 2024 · 0 repositories · arXiv:2412.19413
-
Nonconvex Stochastic Optimization under Heavy-Tailed Noises: Optimal Convergence without Gradient Clipping 27 Dec 2024 · 0 repositories · arXiv:2412.19529
-
Optimizing Local-Global Dependencies for Accurate 3D Human Pose Estimation 27 Dec 2024 · 1 repository · arXiv:2412.19676
-
P3S-Diffusion:A Selective Subject-driven Generation Framework via Point Supervision 27 Dec 2024 · 0 repositories · arXiv:2412.19533
-
Pre-training, Fine-tuning and Re-ranking: A Three-Stage Framework for Legal Question Answering 27 Dec 2024 · 0 repositories · arXiv:2412.19482
-
RAIN: Real-time Animation of Infinite Video Stream 27 Dec 2024 · 0 repositories · arXiv:2412.19489
-
Revisiting PCA for time series reduction in temporal dimension 27 Dec 2024 · 1 repository · arXiv:2412.19423
-
Scalable Hierarchical Reinforcement Learning for Hyper Scale Multi-Robot Task Planning 27 Dec 2024 · 0 repositories · arXiv:2412.19538
-
Seq2Seq Model-Based Chatbot with LSTM and Attention Mechanism for Enhanced User Interaction 27 Dec 2024 · 0 repositories · arXiv:2501.00049
-
Structural Similarity in Deep Features: Image Quality Assessment Robust to Geometrically Disparate Reference 27 Dec 2024 · 0 repositories · arXiv:2412.19553
-
StyleRWKV: High-Quality and High-Efficiency Style Transfer with RWKV-like Architecture 27 Dec 2024 · 0 repositories · arXiv:2412.19535
-
Text2Insight: Transform natural language text into insights seamlessly using multi-model architecture 27 Dec 2024 · 0 repositories · arXiv:2412.19718
-
Toward Adaptive Reasoning in Large Language Models with Thought Rollback 27 Dec 2024 · 1 repository · arXiv:2412.19707Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
VideoMaker: Zero-shot Customized Video Generation with the Inherent Force of Video Diffusion Models 27 Dec 2024 · 1 repository · arXiv:2412.19645Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
A novel framework for MCDM based on Z numbers and soft likelihood function 26 Dec 2024 · 0 repositories · arXiv:2412.19321
-
A Review of Resilience Enhancement Measures for Hydrogen-penetrated Multi-energy Systems 26 Dec 2024 · 0 repositories · arXiv:2412.19374
-
Advanced Knowledge Transfer: Refined Feature Distillation for Zero-Shot Quantization in Edge Computing 26 Dec 2024 · 1 repository · arXiv:2412.19125
-
Context-Aware Deep Learning for Multi Modal Depression Detection 26 Dec 2024 · 1 repository · arXiv:2412.19209
-
DAPoinTr: Domain Adaptive Point Transformer for Point Cloud Completion 26 Dec 2024 · 1 repository · arXiv:2412.19062
-
Dual Channel Multi-Attention in ViT for Biometric Authentication using Forehead Subcutaneous Vein Pattern and Periocular Pattern 26 Dec 2024 · 0 repositories · arXiv:2412.19160
-
GAIS: A Novel Approach to Instance Selection with Graph Attention Networks 26 Dec 2024 · 0 repositories · arXiv:2412.19201
-
Graph-Enhanced Dual-Stream Feature Fusion with Pre-Trained Model for Acoustic Traffic Monitoring 26 Dec 2024 · 0 repositories · arXiv:2412.19078
-
Indonesian-English Code-Switching Speech Synthesizer Utilizing Multilingual STEN-TTS and Bert LID 26 Dec 2024 · 0 repositories · arXiv:2412.19043
-
Learning Cross-Domain Representations for Transferable Drug Perturbations on Single-Cell Transcriptional Responses 26 Dec 2024 · 1 repository · arXiv:2412.19228
-
MEDEC: A Benchmark for Medical Error Detection and Correction in Clinical Notes 26 Dec 2024 · 1 repository · arXiv:2412.19260
-
Multi-matrix Factorization Attention 26 Dec 2024 · 0 repositories · arXiv:2412.19255
-
On the Expressiveness and Length Generalization of Selective State-Space Models on Regular Languages 26 Dec 2024 · 1 repository · arXiv:2412.19350
-
Personalized Dynamic Music Emotion Recognition with Dual-Scale Attention-Based Meta-Learning 26 Dec 2024 · 1 repository · arXiv:2412.19200
-
RAG with Differential Privacy 26 Dec 2024 · 1 repository · arXiv:2412.19291
-
Rethinking Masked Representation Learning for 3D Point Cloud Understanding 26 Dec 2024 · 1 repository