Methods › General › Attention Mechanisms › Attention › Papers, page 96
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 96 of 316: papers 9,501 to 9,600 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
EDSNet: Efficient-DSNet for Video Summarization 23 Sep 2024 · 0 repositories · arXiv:2409.14724
-
PAPILLON: Efficient and Stealthy Fuzz Testing-Powered Jailbreaks for LLMs 23 Sep 2024 · 1 repository · arXiv:2409.14866Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Efficiently Dispatching Flash Attention For Partially Filled Attention Masks 23 Sep 2024 · 0 repositories · arXiv:2409.15097
-
Enhancing Scientific Reproducibility Through Automated BioCompute Object Creation Using Retrieval-Augmented Generation from Publications 23 Sep 2024 · 0 repositories · arXiv:2409.15076
-
GATher: Graph Attention Based Predictions of Gene-Disease Links 23 Sep 2024 · 0 repositories · arXiv:2409.16327
-
GEM-RAG: Graphical Eigen Memories For Retrieval Augmented Generation 23 Sep 2024 · 0 repositories · arXiv:2409.15566
-
Generalizing monocular colonoscopy image depth estimation by uncertainty-based global and local fusion network 23 Sep 2024 · 0 repositories · arXiv:2409.15006
-
Generative AI Is Not Ready for Clinical Use in Patient Education for Lower Back Pain Patients, Even With Retrieval-Augmented Generation 23 Sep 2024 · 0 repositories · arXiv:2409.15260
-
Generative LLM Powered Conversational AI Application for Personalized Risk Assessment: A Case Study in COVID-19 23 Sep 2024 · 0 repositories · arXiv:2409.15027
-
Goal-based Neural Physics Vehicle Trajectory Prediction Model 23 Sep 2024 · 0 repositories · arXiv:2409.15182
-
Graph Network Models To Detect Illicit Transactions In Block Chain 23 Sep 2024 · 0 repositories · arXiv:2410.07150
-
HydroVision: LiDAR-Guided Hydrometric Prediction with Vision Transformers and Hybrid Graph Learning 23 Sep 2024 · 0 repositories · arXiv:2409.15213
-
Improving Academic Skills Assessment with NLP and Ensemble Learning 23 Sep 2024 · 0 repositories · arXiv:2409.19013
-
Inference-Friendly Models With MixAttention 23 Sep 2024 · 1 repository · arXiv:2409.15012
-
Investigating Robot Dogs for Construction Monitoring: A Comparative Analysis of Specifications and On-site Requirements 23 Sep 2024 · 1 repository · arXiv:2409.15253
-
Kriformer: A Novel Spatiotemporal Kriging Approach Based on Graph Transformers 23 Sep 2024 · 0 repositories · arXiv:2409.14906
-
Learning When to Retrieve, What to Rewrite, and How to Respond in Conversational QA 23 Sep 2024 · 0 repositories · arXiv:2409.15515
-
Less yet robust: crucial region selection for scene recognition 23 Sep 2024 · 0 repositories · arXiv:2409.14741
-
Lessons Learned on Information Retrieval in Electronic Health Records: A Comparison of Embedding Models and Pooling Strategies 23 Sep 2024 · 0 repositories · arXiv:2409.15163
-
Location is Key: Leveraging Large Language Model for Functional Bug Localization in Verilog 23 Sep 2024 · 0 repositories · arXiv:2409.15186
-
M2OST: Many-to-one Regression for Predicting Spatial Transcriptomics from Digital Pathology Images 23 Sep 2024 · 1 repository · arXiv:2409.15092Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)
-
MemeCLIP: Leveraging CLIP Representations for Multimodal Meme Classification 23 Sep 2024 · 1 repository · arXiv:2409.14703
-
Methods for Convex (L₀,L₁)-Smooth Optimization: Clipping, Acceleration, and Adaptivity 23 Sep 2024 · 0 repositories · arXiv:2409.14989
-
Micrometer: Micromechanics Transformer for Predicting Mechanical Responses of Heterogeneous Materials 23 Sep 2024 · 0 repositories · arXiv:2410.05281
-
Multi-Modal Generative AI: Multi-modal LLM, Diffusion and Beyond 23 Sep 2024 · 0 repositories · arXiv:2409.14993
-
Optimizing News Text Classification with Bi-LSTM and Attention Mechanism for Efficient Data Processing 23 Sep 2024 · 0 repositories · arXiv:2409.15576
-
PALLM: Evaluating and Enhancing PALLiative Care Conversations with Large Language Models 23 Sep 2024 · 1 repository · arXiv:2409.15188
-
Privacy Policy Analysis through Prompt Engineering for LLMs 23 Sep 2024 · 0 repositories · arXiv:2409.14879
-
Probabilistically Aligned View-unaligned Clustering with Adaptive Template Selection 23 Sep 2024 · 0 repositories · arXiv:2409.14882
-
RACER: Rich Language-Guided Failure Recovery Policies for Imitation Learning 23 Sep 2024 · 0 repositories · arXiv:2409.14674
-
Retrieval Augmented Generation (RAG) and Beyond: A Comprehensive Survey on How to Make your LLMs use External Data More Wisely 23 Sep 2024 · 0 repositories · arXiv:2409.14924
-
Robust and Flexible Omnidirectional Depth Estimation with Multiple 360° Cameras 23 Sep 2024 · 0 repositories · arXiv:2409.14766
-
RoWSFormer: A Robust Watermarking Framework with Swin Transformer for Enhanced Geometric Attack Resilience 23 Sep 2024 · 0 repositories · arXiv:2409.14829
-
Safe Guard: an LLM-agent for Real-time Voice-based Hate Speech Detection in Social Virtual Reality 23 Sep 2024 · 0 repositories · arXiv:2409.15623
-
Scaling Laws of Decoder-Only Models on the Multilingual Machine Translation Task 23 Sep 2024 · 0 repositories · arXiv:2409.15051
-
SDBA: A Stealthy and Long-Lasting Durable Backdoor Attack in Federated Learning 23 Sep 2024 · 1 repository · arXiv:2409.14805
-
SOFI: Multi-Scale Deformable Transformer for Camera Calibration with Enhanced Line Queries 23 Sep 2024 · 1 repository · arXiv:2409.15553
-
TransUKAN:Computing-Efficient Hybrid KAN-Transformer for Enhanced Medical Image Segmentation 23 Sep 2024 · 0 repositories · arXiv:2409.14676
-
Beyond Words: Evaluating Large Language Models in Transportation Planning 22 Sep 2024 · 0 repositories · arXiv:2409.14516
-
Can pre-trained language models generate titles for research papers? 22 Sep 2024 · 1 repository · arXiv:2409.14602
-
EchoAtt: Attend, Copy, then Adjust for More Efficient Large Language Models 22 Sep 2024 · 0 repositories · arXiv:2409.14595
-
EM-DARTS: Hierarchical Differentiable Architecture Search for Eye Movement Recognition 22 Sep 2024 · 0 repositories · arXiv:2409.14432
-
Enhancing LLM-based Autonomous Driving Agents to Mitigate Perception Attacks 22 Sep 2024 · 0 repositories · arXiv:2409.14488
-
EQ-CBM: A Probabilistic Concept Bottleneck with Energy-based Models and Quantized Vectors 22 Sep 2024 · 0 repositories · arXiv:2409.14630
-
Evaluating the Quality of Code Comments Generated by Large Language Models for Novice Programmers 22 Sep 2024 · 0 repositories · arXiv:2409.14368
-
GroupDiff: Diffusion-based Group Portrait Editing 22 Sep 2024 · 1 repository · arXiv:2409.14379
-
Investigating Layer Importance in Large Language Models 22 Sep 2024 · 0 repositories · arXiv:2409.14381
-
J2N -- Nominal Adjective Identification and its Application 22 Sep 2024 · 1 repository · arXiv:2409.14374
-
Large Model Based Agents: State-of-the-Art, Cooperation Paradigms, Security and Privacy, and Future Trends 22 Sep 2024 · 0 repositories · arXiv:2409.14457
-
LLMs are One-Shot URL Classifiers and Explainers 22 Sep 2024 · 0 repositories · arXiv:2409.14306
-
More Effective LLM Compressed Tokens with Uniformly Spread Position Identifiers and Compression Loss 22 Sep 2024 · 0 repositories · arXiv:2409.14364
-
OStr-DARTS: Differentiable Neural Architecture Search based on Operation Strength 22 Sep 2024 · 1 repository · arXiv:2409.14433
-
Patch Ranking: Efficient CLIP by Learning to Rank Local Patches 22 Sep 2024 · 1 repository · arXiv:2409.14607
-
Prior Knowledge Distillation Network for Face Super-Resolution 22 Sep 2024 · 0 repositories · arXiv:2409.14385
-
Proof Automation with Large Language Models 22 Sep 2024 · 0 repositories · arXiv:2409.14274
-
Sparse Low-Ranked Self-Attention Transformer for Remaining Useful Lifetime Prediction of Optical Fiber Amplifiers 22 Sep 2024 · 0 repositories · arXiv:2409.14378
-
Thinking in Granularity: Dynamic Quantization for Image Super-Resolution by Intriguing Multi-Granularity Clues 22 Sep 2024 · 1 repository · arXiv:2409.14330
-
TrackNetV4: Enhancing Fast Sports Object Tracking with Motion Attention Maps 22 Sep 2024 · 0 repositories · arXiv:2409.14543
-
UU-Mamba: Uncertainty-aware U-Mamba for Cardiovascular Segmentation 22 Sep 2024 · 1 repository · arXiv:2409.14305
-
ChemEval: A Comprehensive Multi-Level Chemical Evaluation for Large Language Models 21 Sep 2024 · 1 repository · arXiv:2409.13989Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 0 violated, 17 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
SMART-RAG: Selection using Determinantal Matrices for Augmented Retrieval 21 Sep 2024 · 0 repositories · arXiv:2409.13992
-
Graph Neural Network Framework for Sentiment Analysis Using Syntactic Feature 21 Sep 2024 · 0 repositories · arXiv:2409.14000
-
Can LLMs replace Neil deGrasse Tyson? Evaluating the Reliability of LLMs as Science Communicators 21 Sep 2024 · 1 repository · arXiv:2409.14037
-
MultiMed: Multilingual Medical Speech Recognition via Attention Encoder Decoder 21 Sep 2024 · 1 repository · arXiv:2409.14074
-
Probing Context Localization of Polysemous Words in Pre-trained Language Model Sub-Layers 21 Sep 2024 · 0 repositories · arXiv:2409.14097
-
Interpreting Arithmetic Mechanism in Large Language Models through Comparative Neuron Analysis 21 Sep 2024 · 2 repositories · arXiv:2409.14144Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Towards Building Efficient Sentence BERT Models using Layer Pruning 21 Sep 2024 · 0 repositories · arXiv:2409.14168
-
QMOS: Enhancing LLMs for Telecommunication with Question Masked loss and Option Shuffling 21 Sep 2024 · 1 repository · arXiv:2409.14175
-
A Sinkhorn Regularized Adversarial Network for Image Guided DEM Super-resolution using Frequency Selective Hybrid Graph Transformer 21 Sep 2024 · 0 repositories · arXiv:2409.14198
-
AI Assistants for Spaceflight Procedures: Combining Generative Pre-Trained Transformer and Retrieval-Augmented Generation on Knowledge Graphs With Augmented Reality Cues 21 Sep 2024 · 0 repositories · arXiv:2409.14206
-
Are Music Foundation Models Better at Singing Voice Deepfake Detection? Far-Better Fuse them with Speech Foundation Models 21 Sep 2024 · 0 repositories · arXiv:2409.14131
-
Boolean Product Graph Neural Networks 21 Sep 2024 · 0 repositories · arXiv:2409.14001
-
Detecting Inpainted Video with Frequency Domain Insights 21 Sep 2024 · 0 repositories · arXiv:2409.13976
-
Developing a Thailand solar irradiance map using Himawari-8 satellite imageries and deep learning models 21 Sep 2024 · 1 repository · arXiv:2409.16320
-
Drift to Remember 21 Sep 2024 · 0 repositories · arXiv:2409.13997
-
Multilateral Cascading Network for Semantic Segmentation of Large-Scale Outdoor Point Clouds 21 Sep 2024 · 0 repositories · arXiv:2409.13983
-
FAMOUS: Flexible Accelerator for the Attention Mechanism of Transformer on UltraScale+ FPGAs 21 Sep 2024 · 0 repositories · arXiv:2409.14023
-
Generalizable Non-Line-of-Sight Imaging with Learnable Physical Priors 21 Sep 2024 · 0 repositories · arXiv:2409.14011
-
Knowledge in Triples for LLMs: Enhancing Table QA Accuracy with Semantic Extraction 21 Sep 2024 · 0 repositories · arXiv:2409.14192
-
Loop Neural Networks for Parameter Sharing 21 Sep 2024 · 0 repositories · arXiv:2409.14199
-
Monocular Event-Inertial Odometry with Adaptive decay-based Time Surface and Polarity-aware Tracking 21 Sep 2024 · 0 repositories · arXiv:2409.13971
-
MSDet: Receptive Field Enhanced Multiscale Detection for Tiny Pulmonary Nodule 21 Sep 2024 · 1 repository · arXiv:2409.14028
-
Multiple-Exit Tuning: Towards Inference-Efficient Adaptation for Vision Transformer 21 Sep 2024 · 0 repositories · arXiv:2409.13999
-
On Broad-Beam Reflection for Dual-Polarized RIS-Assisted MIMO Systems 21 Sep 2024 · 0 repositories · arXiv:2410.07134
-
ProTEA: Programmable Transformer Encoder Acceleration on FPGA 21 Sep 2024 · 0 repositories · arXiv:2409.13975
-
Semi-intrusive audio evaluation: Casting non-intrusive assessment as a multi-modal text prediction task 21 Sep 2024 · 0 repositories · arXiv:2409.14069
-
What is a Digital Twin Anyway? Deriving the Definition for the Built Environment from over 15,000 Scientific Publications 21 Sep 2024 · 0 repositories · arXiv:2409.19005
-
Window-based Channel Attention for Wavelet-enhanced Learned Image Compression 21 Sep 2024 · 0 repositories · arXiv:2409.14090
-
Contextual Compression in Retrieval-Augmented Generation for Large Language Models: A Survey 20 Sep 2024 · 1 repository · arXiv:2409.13385
-
3D-GSW: 3D Gaussian Splatting for Robust Watermarking 20 Sep 2024 · 0 repositories · arXiv:2409.13222
-
A Comparison between Financial and Gambling Markets 20 Sep 2024 · 0 repositories · arXiv:2409.13528
-
A Personalised 3D+t Mesh Generative Model for Unveiling Normal Heart Dynamics 20 Sep 2024 · 1 repository · arXiv:2409.13825
-
A Survey of 5G-Based Positioning for Industry 4.0: State of the Art and Enhanced Techniques 20 Sep 2024 · 0 repositories · arXiv:2409.13308
-
Aligning Language Models Using Follow-up Likelihood as Reward Signal 20 Sep 2024 · 1 repository · arXiv:2409.13948
-
Analysis of Gene Regulatory Networks from Gene Expression Using Graph Neural Networks 20 Sep 2024 · 1 repository · arXiv:2409.13664
-
Applying Pre-trained Multilingual BERT in Embeddings for Improved Malicious Prompt Injection Attacks Detection 20 Sep 2024 · 0 repositories · arXiv:2409.13331
-
AVG-LLaVA: A Large Multimodal Model with Adaptive Visual Granularity 20 Sep 2024 · 1 repository · arXiv:2410.02745Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Beyond the binary: Limitations and possibilities of gender-related speech technology research 20 Sep 2024 · 1 repository · arXiv:2409.13335
-
Brain-Cognition Fingerprinting via Graph-GCCA with Contrastive Learning 20 Sep 2024 · 0 repositories · arXiv:2409.13887
-
Cooperative Resilience in Artificial Intelligence Multiagent Systems 20 Sep 2024 · 1 repository · arXiv:2409.13187