Methods › General › Attention Mechanisms › Attention › Papers, page 53
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 53 of 316: papers 5,201 to 5,300 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Audio-Visual Deepfake Detection With Local Temporal Inconsistencies 14 Jan 2025 · 0 repositories · arXiv:2501.08137
-
BMIP: Bi-directional Modality Interaction Prompt Learning for VLM 14 Jan 2025 · 0 repositories · arXiv:2501.07769
-
Change Captioning in Remote Sensing: Evolution to SAT-Cap -- A Single-Stage Transformer Approach 14 Jan 2025 · 0 repositories · arXiv:2501.08114
-
Cloud Removal With PolSAR-Optical Data Fusion Using A Two-Flow Residual Network 14 Jan 2025 · 0 repositories · arXiv:2501.07901
-
Comparative Analysis of Efficient Adapter-Based Fine-Tuning of State-of-the-Art Transformer Models 14 Jan 2025 · 0 repositories · arXiv:2501.08271
-
Comprehensive Metapath-based Heterogeneous Graph Transformer for Gene-Disease Association Prediction 14 Jan 2025 · 0 repositories · arXiv:2501.07970
-
Decision Transformers for RIS-Assisted Systems with Diffusion Model-Based Channel Acquisition 14 Jan 2025 · 0 repositories · arXiv:2501.08007
-
Decoding Interpretable Logic Rules from Neural Networks 14 Jan 2025 · 0 repositories · arXiv:2501.08281
-
Dynamic Multimodal Sentiment Analysis: Leveraging Cross-Modal Attention for Enabled Classification 14 Jan 2025 · 0 repositories · arXiv:2501.08085
-
EEG-ReMinD: Enhancing Neurodegenerative EEG Decoding through Self-Supervised State Reconstruction-Primed Riemannian Dynamics 14 Jan 2025 · 0 repositories · arXiv:2501.08139
-
Efficient Deep Learning-based Forward Solvers for Brain Tumor Growth Models 14 Jan 2025 · 1 repository · arXiv:2501.08226
-
Eliciting In-context Retrieval and Reasoning for Long-context Large Language Models 14 Jan 2025 · 0 repositories · arXiv:2501.08248
-
EmoNeXt: an Adapted ConvNeXt for Facial Emotion Recognition 14 Jan 2025 · 1 repository · arXiv:2501.08199
-
Exploring Narrative Clustering in Large Language Models: A Layerwise Analysis of BERT 14 Jan 2025 · 0 repositories · arXiv:2501.08053
-
Exploring Robustness of Multilingual LLMs on Real-World Noisy Data 14 Jan 2025 · 1 repository · arXiv:2501.08322
-
FramePainter: Endowing Interactive Image Editing with Video Diffusion Priors 14 Jan 2025 · 1 repository · arXiv:2501.08225Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
GAC-Net_Geometric and attention-based Network for Depth Completion 14 Jan 2025 · 0 repositories · arXiv:2501.07988
-
Hybrid Action Based Reinforcement Learning for Multi-Objective Compatible Autonomous Driving 14 Jan 2025 · 0 repositories · arXiv:2501.08096
-
Investigating Energy Efficiency and Performance Trade-offs in LLM Inference Across Tasks and DVFS Settings 14 Jan 2025 · 0 repositories · arXiv:2501.08219
-
Large Language Models for Knowledge Graph Embedding Techniques, Methods, and Challenges: A Survey 14 Jan 2025 · 0 repositories · arXiv:2501.07766
-
Large Language Models For Text Classification: Case Study And Comprehensive Review 14 Jan 2025 · 0 repositories · arXiv:2501.08457
-
Leveraging 2D Masked Reconstruction for Domain Adaptation of 3D Pose Estimation 14 Jan 2025 · 0 repositories · arXiv:2501.08408
-
Leveraging Metamemory Mechanisms for Enhanced Data-Free Code Generation in LLMs 14 Jan 2025 · 0 repositories · arXiv:2501.07892
-
LLaVA-ST: A Multimodal Large Language Model for Fine-Grained Spatial-Temporal Understanding 14 Jan 2025 · 1 repository · arXiv:2501.08282Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples)
-
Logarithmic Memory Networks (LMNs): Efficient Long-Range Sequence Modeling for Resource-Constrained Environments 14 Jan 2025 · 1 repository · arXiv:2501.07905
-
MD-Syn: Synergistic drug combination prediction based on the multidimensional feature fusion method and attention mechanisms 14 Jan 2025 · 0 repositories · arXiv:2501.07884
-
MiniMax-01: Scaling Foundation Models with Lightning Attention 14 Jan 2025 · 1 repository · arXiv:2501.08313Syntology 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Optimizing Language Models for Grammatical Acceptability: A Comparative Study of Fine-Tuning Techniques 14 Jan 2025 · 0 repositories · arXiv:2501.07853
-
Physics-Informed Machine Learning for Microscale Drying of Plant-Based Foods: A Systematic Review of Computational Models and Experimental Insights 14 Jan 2025 · 0 repositories · arXiv:2501.09034
-
PokerBench: Training Large Language Models to become Professional Poker Players 14 Jan 2025 · 1 repository · arXiv:2501.08328
-
Predicting 4D Hand Trajectory from Monocular Videos 14 Jan 2025 · 0 repositories · arXiv:2501.08329
-
PSReg: Prior-guided Sparse Mixture of Experts for Point Cloud Registration 14 Jan 2025 · 0 repositories · arXiv:2501.07762
-
READ: Reinforcement-based Adversarial Learning for Text Classification with Limited Labeled Data 14 Jan 2025 · 0 repositories · arXiv:2501.08035
-
ReARTeR: Retrieval-Augmented Reasoning with Trustworthy Process Rewarding 14 Jan 2025 · 1 repository · arXiv:2501.07861
-
Robust Hyperspectral Image Panshapring via Sparse Spatial-Spectral Representation 14 Jan 2025 · 0 repositories · arXiv:2501.07953
-
Selective Attention Merging for low resource tasks: A case study of Child ASR 14 Jan 2025 · 1 repository · arXiv:2501.08468
-
Spiking Neural Network Accelerator Architecture for Differential-Time Representation using Learned Encoding 14 Jan 2025 · 0 repositories · arXiv:2501.07952
-
Threshold Attention Network for Semantic Segmentation of Remote Sensing Images 14 Jan 2025 · 0 repositories · arXiv:2501.07984
-
Towards Lightweight Time Series Forecasting: a Patch-wise Transformer with Weak Data Enriching 14 Jan 2025 · 0 repositories · arXiv:2501.10448
-
Transforming Indoor Localization: Advanced Transformer Architecture for NLOS Dominated Wireless Environments with Distributed Sensors 14 Jan 2025 · 0 repositories · arXiv:2501.07774
-
UFGraphFR: An attempt at a federated recommendation system based on user text characteristics 14 Jan 2025 · 1 repository · arXiv:2501.08044
-
A Multi-Modal Deep Learning Framework for Pan-Cancer Prognosis 13 Jan 2025 · 1 repository · arXiv:2501.07016
-
AdaCS: Adaptive Normalization for Enhanced Code-Switching ASR 13 Jan 2025 · 1 repository · arXiv:2501.07102
-
An Adaptive Collocation Point Strategy For Physics Informed Neural Networks via the QR Discrete Empirical Interpolation Method 13 Jan 2025 · 1 repository · arXiv:2501.07700
-
An Investigation into Seasonal Variations in Energy Forecasting for Student Residences 13 Jan 2025 · 0 repositories · arXiv:2501.07423
-
Attention when you need 13 Jan 2025 · 0 repositories · arXiv:2501.07440
-
Bigger Isn't Always Better: Towards a General Prior for Medical Image Reconstruction 13 Jan 2025 · 1 repository · arXiv:2501.07376
-
BlobGEN-Vid: Compositional Text-to-Video Generation with Blob Video Representations 13 Jan 2025 · 0 repositories · arXiv:2501.07647
-
Code and Pixels: Multi-Modal Contrastive Pre-training for Enhanced Tabular Data Analysis 13 Jan 2025 · 0 repositories · arXiv:2501.07304
-
Comparative analysis of optical character recognition methods for Sámi texts from the National Library of Norway 13 Jan 2025 · 2 repositories · arXiv:2501.07300
-
D3MES: Diffusion Transformer with multihead equivariant self-attention for 3D molecule generation 13 Jan 2025 · 1 repository · arXiv:2501.07077
-
Dual Scale-aware Adaptive Masked Knowledge Distillation for Object Detection 13 Jan 2025 · 0 repositories · arXiv:2501.07101
-
EdgeTAM: On-Device Track Anything Model 13 Jan 2025 · 1 repository · arXiv:2501.07256Syntology official: no sample here; runs from other or unrecorded repositories · 7 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Efficient Event-based Delay Learning in Spiking Neural Networks 13 Jan 2025 · 1 repository · arXiv:2501.07331
-
Emergent effects of scaling on the functional hierarchies within large language models 13 Jan 2025 · 0 repositories · arXiv:2501.07359
-
Enhancing Image Generation Fidelity via Progressive Prompts 13 Jan 2025 · 1 repository · arXiv:2501.07070Syntology official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Enhancing Retrieval-Augmented Generation: A Study of Best Practices 13 Jan 2025 · 1 repository · arXiv:2501.07391Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Enhancing Talent Employment Insights Through Feature Extraction with LLM Finetuning 13 Jan 2025 · 0 repositories · arXiv:2501.07663
-
Estimating Musical Surprisal in Audio 13 Jan 2025 · 1 repository · arXiv:2501.07474
-
FinerWeb-10BT: Refining Web Data with LLM-Based Line-Level Filtering 13 Jan 2025 · 1 repository · arXiv:2501.07314
-
Future-Conditioned Recommendations with Multi-Objective Controllable Decision Transformer 13 Jan 2025 · 0 repositories · arXiv:2501.07212
-
GPT as a Monte Carlo Language Tree: A Probabilistic Perspective 13 Jan 2025 · 0 repositories · arXiv:2501.07641
-
How GPT learns layer by layer 13 Jan 2025 · 1 repository · arXiv:2501.07108
-
Intent-Interest Disentanglement and Item-Aware Intent Contrastive Learning for Sequential Recommendation 13 Jan 2025 · 0 repositories · arXiv:2501.07096
-
ListConRanker: A Contrastive Text Reranker with Listwise Encoding 13 Jan 2025 · 0 repositories · arXiv:2501.07111
-
MathReader : Text-to-Speech for Mathematical Documents 13 Jan 2025 · 1 repository · arXiv:2501.07088
-
MSV-Mamba: A Multiscale Vision Mamba Network for Echocardiography Segmentation 13 Jan 2025 · 0 repositories · arXiv:2501.07120
-
VDOR: A Video-based Dataset for Object Removal via Sequence Consistency 13 Jan 2025 · 0 repositories · arXiv:2501.07397
-
Parallel Key-Value Cache Fusion for Position Invariant RAG 13 Jan 2025 · 0 repositories · arXiv:2501.07523
-
PRKAN: Parameter-Reduced Kolmogorov-Arnold Networks 13 Jan 2025 · 1 repository · arXiv:2501.07032
-
Protego: Detecting Adversarial Examples for Vision Transformers via Intrinsic Capabilities 13 Jan 2025 · 0 repositories · arXiv:2501.07044
-
Scaling Up ESM2 Architectures for Long Protein Sequences Analysis: Long and Quantized Approaches 13 Jan 2025 · 0 repositories · arXiv:2501.07747
-
Social and Genetic Ties Drive Skewed Cross-Border Media Coverage of Disasters 13 Jan 2025 · 0 repositories · arXiv:2501.07615
-
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing 13 Jan 2025 · 1 repository · arXiv:2501.07554
-
Subject Representation Learning from EEG using Graph Convolutional Variational Autoencoders 13 Jan 2025 · 0 repositories · arXiv:2501.16626
-
The Quest for Visual Understanding: A Journey Through the Evolution of Visual Question Answering 13 Jan 2025 · 0 repositories · arXiv:2501.07109
-
UNetVL: Enhancing 3D Medical Image Segmentation with Chebyshev KAN Powered Vision-LSTM 13 Jan 2025 · 1 repository · arXiv:2501.07017
-
VAGeo: View-specific Attention for Cross-View Object Geo-Localization 13 Jan 2025 · 0 repositories · arXiv:2501.07194
-
WebWalker: Benchmarking LLMs in Web Traversal 13 Jan 2025 · 2 repositories · arXiv:2501.07572
-
Better Prompt Compression Without Multi-Layer Perceptrons 12 Jan 2025 · 0 repositories · arXiv:2501.06730
-
DRDT3: Diffusion-Refined Decision Test-Time Training Model 12 Jan 2025 · 0 repositories · arXiv:2501.06718
-
Eliza: A Web3 friendly AI Agent Operating System 12 Jan 2025 · 2 repositories · arXiv:2501.06781
-
Generative Artificial Intelligence-Supported Pentesting: A Comparison between Claude Opus, GPT-4, and Copilot 12 Jan 2025 · 0 repositories · arXiv:2501.06963
-
Local Foreground Selection aware Attentive Feature Reconstruction for few-shot fine-grained plant species classification 12 Jan 2025 · 0 repositories · arXiv:2501.06909
-
MFConvTr: Multi-Frequency Convolutional Transformer for Fetal Arrhythmia Detection in Non-Invasive fECG 12 Jan 2025 · 0 repositories · arXiv:2501.16649
-
MiniRAG: Towards Extremely Simple Retrieval-Augmented Generation 12 Jan 2025 · 1 repository · arXiv:2501.06713Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
ODPG: Outfitting Diffusion with Pose Guided Condition 12 Jan 2025 · 0 repositories · arXiv:2501.06769
-
On the Complexity of Global Necessary Reasons to Explain Classification 12 Jan 2025 · 0 repositories · arXiv:2501.06766
-
Procedural Fairness and Its Relationship with Distributive Fairness in Machine Learning 12 Jan 2025 · 1 repository · arXiv:2501.06753
-
TAPO: Task-Referenced Adaptation for Prompt Optimization 12 Jan 2025 · 1 repository · arXiv:2501.06689
-
Transforming Vision Transformer: Towards Efficient Multi-Task Asynchronous Learning 12 Jan 2025 · 1 repository · arXiv:2501.06884Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
What Is a Counterfactual Cause in Action Theories? 12 Jan 2025 · 0 repositories · arXiv:2501.06857
-
ZNO-Eval: Benchmarking reasoning capabilities of large language models in Ukrainian 12 Jan 2025 · 1 repository · arXiv:2501.06715
-
A Comparative Performance Analysis of Classification and Segmentation Models on Bangladeshi Pothole Dataset 11 Jan 2025 · 0 repositories · arXiv:2501.06602
-
Assessing instructor-AI cooperation for grading essay-type questions in an introductory sociology course 11 Jan 2025 · 1 repository · arXiv:2501.06461
-
CeViT: Copula-Enhanced Vision Transformer in multi-task learning and bi-group image covariates with an application to myopia screening 11 Jan 2025 · 1 repository · arXiv:2501.06540
-
CPDR: Towards Highly-Efficient Salient Object Detection via Crossed Post-decoder Refinement 11 Jan 2025 · 0 repositories · arXiv:2501.06441
-
Dual-Modality Representation Learning for Molecular Property Prediction 11 Jan 2025 · 0 repositories · arXiv:2501.06608
-
First Token Probability Guided RAG for Telecom Question Answering 11 Jan 2025 · 0 repositories · arXiv:2501.06468
-
Flash Window Attention: speedup the attention computation for Swin Transformer 11 Jan 2025 · 2 repositories · arXiv:2501.06480