Methods › General › Attention Mechanisms › Attention › Papers, page 81
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 81 of 316: papers 8,001 to 8,100 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Is GPT-4 Less Politically Biased than GPT-3.5? A Renewed Investigation of ChatGPT's Political Biases 28 Oct 2024 · 0 repositories · arXiv:2410.21008
-
Joint Audio-Visual Idling Vehicle Detection with Streamlined Input Dependencies 28 Oct 2024 · 0 repositories · arXiv:2410.21170
-
KA²ER: Knowledge Adaptive Amalgamation of ExpeRts for Medical Images Segmentation 28 Oct 2024 · 0 repositories · arXiv:2410.21085
-
KaLDeX: Kalman Filter based Linear Deformable Cross Attention for Retina Vessel Segmentation 28 Oct 2024 · 1 repository · arXiv:2410.21160
-
KD-LoRA: A Hybrid Approach to Efficient Fine-Tuning with LoRA and Knowledge Distillation 28 Oct 2024 · 1 repository · arXiv:2410.20777
-
Learning Optimal Combination Patterns for Lightweight Stereo Image Super-Resolution 28 Oct 2024 · 0 repositories
-
LiGAR: LiDAR-Guided Hierarchical Transformer for Multi-Modal Group Activity Recognition 28 Oct 2024 · 0 repositories · arXiv:2410.21108
-
LinFormer: A Linear-based Lightweight Transformer Architecture For Time-Aware MIMO Channel Prediction 28 Oct 2024 · 0 repositories · arXiv:2410.21351
-
LLMs are Biased Evaluators But Not Biased for Retrieval Augmented Generation 28 Oct 2024 · 1 repository · arXiv:2410.20833
-
Long Sequence Modeling with Attention Tensorization: From Sequence to Tensor Learning 28 Oct 2024 · 0 repositories · arXiv:2410.20926
-
M2rc-Eval: Massively Multilingual Repository-level Code Completion Evaluation 28 Oct 2024 · 0 repositories · arXiv:2410.21157
-
Modeling and Replication of the Prepayment Option of Mortgages including Behavioral Uncertainty 28 Oct 2024 · 0 repositories · arXiv:2410.21110
-
Multi-modal AI for comprehensive breast cancer prognostication 28 Oct 2024 · 0 repositories · arXiv:2410.21256
-
MultiTok: Variable-Length Tokenization for Efficient LLMs Adapted from LZW Compression 28 Oct 2024 · 1 repository · arXiv:2410.21548
-
Near Optimal Pure Exploration in Logistic Bandits 28 Oct 2024 · 0 repositories · arXiv:2410.20640
-
ODGS: 3D Scene Reconstruction from Omnidirectional Images with 3D Gaussian Splattings 28 Oct 2024 · 1 repository · arXiv:2410.20686
-
On Inductive Biases That Enable Generalization of Diffusion Transformers 28 Oct 2024 · 1 repository · arXiv:2410.21273
-
Pay Attention to Attention for Sequential Recommendation 28 Oct 2024 · 0 repositories · arXiv:2410.21048
-
Plan×RAG: Planning-guided Retrieval Augmented Generation 28 Oct 2024 · 0 repositories · arXiv:2410.20753
-
Provisioning for Solar-Powered Base Stations Driven by Conditional LSTM Networks 28 Oct 2024 · 0 repositories · arXiv:2410.20755
-
Relaxed Recursive Transformers: Effective Parameter Sharing with Layer-wise LoRA 28 Oct 2024 · 0 repositories · arXiv:2410.20672
-
Reprogramming Pretrained Target-Specific Diffusion Models for Dual-Target Drug Design 28 Oct 2024 · 1 repository · arXiv:2410.20688Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 3 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 5 pointer-only (licence)
-
Retrieval-Retro: Retrieval-based Inorganic Retrosynthesis with Expert Knowledge 28 Oct 2024 · 1 repository · arXiv:2410.21341Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
SandboxAQ's submission to MRL 2024 Shared Task on Multi-lingual Multi-task Information Retrieval 28 Oct 2024 · 0 repositories · arXiv:2410.21501
-
Semantic Editing Increment Benefits Zero-Shot Composed Image Retrieval 28 Oct 2024 · 2 repositories
-
Semantic Search Evaluation 28 Oct 2024 · 0 repositories · arXiv:2410.21549
-
SepMamba: State-space models for speaker separation using Mamba 28 Oct 2024 · 1 repository · arXiv:2410.20997
-
ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference 28 Oct 2024 · 1 repository · arXiv:2410.21465Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Simple Is Effective: The Roles of Graphs and Large Language Models in Knowledge-Graph-Based Retrieval-Augmented Generation 28 Oct 2024 · 1 repository · arXiv:2410.20724Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 15 harvested samples)
-
Stealthy Jailbreak Attacks on Large Language Models via Benign Data Mirroring 28 Oct 2024 · 0 repositories · arXiv:2410.21083
-
TGCA-PVT: Topic-Guided Context-Aware Pyramid Vision Transformer for Sticker Emotion Recognition 28 Oct 2024 · 1 repository
-
Unveiling Context-Aware Criteria in Self-Assessing LLMs 28 Oct 2024 · 0 repositories · arXiv:2410.21545
-
uOttawa at LegalLens-2024: Transformer-based Classification Experiments 28 Oct 2024 · 1 repository · arXiv:2410.21139
-
Visualizing attention zones in machine reading comprehension models 28 Oct 2024 · 0 repositories · arXiv:2410.20652
-
A Framework for Real-Time Volcano-Seismic Event Recognition Based on Multi-Station Seismograms and Semantic Segmentation Models 27 Oct 2024 · 1 repository · arXiv:2410.20595
-
Accelerating Augmentation Invariance Pretraining 27 Oct 2024 · 0 repositories · arXiv:2410.22364
-
Accelerating Direct Preference Optimization with Prefix Sharing 27 Oct 2024 · 1 repository · arXiv:2410.20305
-
Automatic Estimation of Singing Voice Musical Dynamics 27 Oct 2024 · 1 repository · arXiv:2410.20540
-
Deep Learning Based Dense Retrieval: A Comparative Study 27 Oct 2024 · 0 repositories · arXiv:2410.20315
-
Deep Learning-Driven Microstructure Characterization and Vickers Hardness Prediction of Mg-Gd Alloys 27 Oct 2024 · 0 repositories · arXiv:2410.20402
-
Depth Attention for Robust RGB Tracking 27 Oct 2024 · 1 repository · arXiv:2410.20395
-
Extracting Alpha from Financial Analyst Networks 27 Oct 2024 · 0 repositories · arXiv:2410.20597
-
GrounDiT: Grounding Diffusion Transformers via Noisy Patch Transplantation 27 Oct 2024 · 1 repository · arXiv:2410.20474Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Leveraging Auxiliary Task Relevance for Enhanced Bearing Fault Diagnosis through Curriculum Meta-learning 27 Oct 2024 · 0 repositories · arXiv:2410.20351
-
LLM Robustness Against Misinformation in Biomedical Question Answering 27 Oct 2024 · 1 repository · arXiv:2410.21330
-
Malinowski in the Age of AI: Can large language models create a text game based on an anthropological classic? 27 Oct 2024 · 0 repositories · arXiv:2410.20536
-
Network scaling and scale-driven loss balancing for intelligent poroelastography 27 Oct 2024 · 0 repositories · arXiv:2411.08886
-
Open-Vocabulary Object Detection via Language Hierarchy 27 Oct 2024 · 0 repositories · arXiv:2410.20371
-
ProtSCAPE: Mapping the landscape of protein conformations in molecular dynamics 27 Oct 2024 · 1 repository · arXiv:2410.20317Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
PViT: Prior-augmented Vision Transformer for Out-of-distribution Detection 27 Oct 2024 · 1 repository · arXiv:2410.20631
-
R^3AG: First Workshop on Refined and Reliable Retrieval Augmented Generation 27 Oct 2024 · 0 repositories · arXiv:2410.20598
-
RopeTP: Global Human Motion Recovery via Integrating Robust Pose Estimation with Diffusion Trajectory Prior 27 Oct 2024 · 0 repositories · arXiv:2410.20358
-
Sebica: Lightweight Spatial and Efficient Bidirectional Channel Attention Super Resolution Network 27 Oct 2024 · 1 repository · arXiv:2410.20546
-
Sequential Large Language Model-Based Hyper-parameter Optimization 27 Oct 2024 · 1 repository · arXiv:2410.20302
-
SympCam: Remote Optical Measurement of Sympathetic Arousal 27 Oct 2024 · 0 repositories · arXiv:2410.20552
-
TEAFormers: TEnsor-Augmented Transformers for Multi-Dimensional Time Series Forecasting 27 Oct 2024 · 0 repositories · arXiv:2410.20439
-
ThunderKittens: Simple, Fast, and Adorable AI Kernels 27 Oct 2024 · 1 repository · arXiv:2410.20399Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 21 harvested samples)
-
UTSRMorph: A Unified Transformer and Superresolution Network for Unsupervised Medical Image Registration 27 Oct 2024 · 1 repository · arXiv:2410.20348
-
Wavelet-based Mamba with Fourier Adjustment for Low-light Image Enhancement 27 Oct 2024 · 1 repository · arXiv:2410.20314
-
A Stack-Propagation Framework for Low-Resource Personalized Dialogue Generation 26 Oct 2024 · 0 repositories · arXiv:2410.20174
-
Attention to Patterns is all you need for Insider threat detection 26 Oct 2024 · 0 repositories
-
Beyond Simple Sum of Delayed Rewards: Non-Markovian Reward Modeling for Reinforcement Learning 26 Oct 2024 · 0 repositories · arXiv:2410.20176
-
CAVE-Net: Classifying Abnormalities in Video Capsule Endoscopy 26 Oct 2024 · 0 repositories · arXiv:2410.20231
-
Enhancing Lie Detection Accuracy: A Comparative Study of Classic ML, CNN, and GCN Models using Audio-Visual Features 26 Oct 2024 · 0 repositories · arXiv:2411.08885
-
GATES: Graph Attention Network with Global Expression Fusion for Deciphering Spatial Transcriptome Architectures 26 Oct 2024 · 1 repository · arXiv:2410.20159
-
Generative Adversarial Patches for Physical Attacks on Cross-Modal Pedestrian Re-Identification 26 Oct 2024 · 0 repositories · arXiv:2410.20097
-
Hybrid Deep Learning for Legal Text Analysis: Predicting Punishment Durations in Indonesian Court Rulings 26 Oct 2024 · 0 repositories · arXiv:2410.20104
-
LLMs Can Evolve Continually on Modality for X-Modal Reasoning 26 Oct 2024 · 1 repository · arXiv:2410.20178
-
Looking Beyond The Top-1: Transformers Determine Top Tokens In Order 26 Oct 2024 · 1 repository · arXiv:2410.20210Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
MarDini: Masked Autoregressive Diffusion for Video Generation at Scale 26 Oct 2024 · 0 repositories · arXiv:2410.20280
-
Mask-based Membership Inference Attacks for Retrieval-Augmented Generation 26 Oct 2024 · 0 repositories · arXiv:2410.20142
-
On the Gaussian process limit of Bayesian Additive Regression Trees 26 Oct 2024 · 0 repositories · arXiv:2410.20289
-
Think Carefully and Check Again! Meta-Generation Unlocking LLMs for Low-Resource Cross-Lingual Summarization 26 Oct 2024 · 0 repositories · arXiv:2410.20021
-
Transforming Precision: A Comparative Analysis of Vision Transformers, CNNs, and Traditional ML for Knee Osteoarthritis Severity Diagnosis 26 Oct 2024 · 0 repositories · arXiv:2410.20062
-
UniVST: A Unified Framework for Training-free Localized Video Style Transfer 26 Oct 2024 · 1 repository · arXiv:2410.20084Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 12 harvested samples)
-
Vulnerability of LLMs to Vertically Aligned Text Manipulations 26 Oct 2024 · 0 repositories · arXiv:2410.20016
-
A Multimodal Approach For Endoscopic VCE Image Classification Using BiomedCLIP-PubMedBERT 25 Oct 2024 · 1 repository · arXiv:2410.19944
-
A Tutorial on Teaching Data Analytics with Generative AI 25 Oct 2024 · 0 repositories · arXiv:2411.07244
-
Capsule Endoscopy Multi-classification via Gated Attention and Wavelet Transformations 25 Oct 2024 · 1 repository · arXiv:2410.19363
-
ChunkRAG: Novel LLM-Chunk Filtering Method for RAG Systems 25 Oct 2024 · 0 repositories · arXiv:2410.19572
-
Computational Bottlenecks of Training Small-scale Large Language Models 25 Oct 2024 · 0 repositories · arXiv:2410.19456
-
Counting Ability of Large Language Models and Impact of Tokenization 25 Oct 2024 · 1 repository · arXiv:2410.19730
-
Deep Learning for Classification of Inflammatory Bowel Disease Activity in Whole Slide Images of Colonic Histopathology 25 Oct 2024 · 0 repositories · arXiv:2410.19690
-
Enhancing Battery Storage Energy Arbitrage with Deep Reinforcement Learning and Time-Series Forecasting 25 Oct 2024 · 1 repository · arXiv:2410.20005
-
Exploring Self-Supervised Learning with U-Net Masked Autoencoders and EfficientNet B7 for Improved Classification 25 Oct 2024 · 1 repository · arXiv:2410.19899
-
FairMT-Bench: Benchmarking Fairness for Multi-turn Dialogue in Conversational LLMs 25 Oct 2024 · 0 repositories · arXiv:2410.19317
-
FISHNET: Financial Intelligence from Sub-querying, Harmonizing, Neural-Conditioning, Expert Swarms, and Task Planning 25 Oct 2024 · 0 repositories · arXiv:2410.19727
-
Frozen-DETR: Enhancing DETR with Image Understanding from Frozen Foundation Models 25 Oct 2024 · 0 repositories · arXiv:2410.19635
-
GeoLLaVA: Efficient Fine-Tuned Vision-Language Models for Temporal Change Detection in Remote Sensing 25 Oct 2024 · 1 repository · arXiv:2410.19552Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
GPT-4o System Card 25 Oct 2024 · 0 repositories · arXiv:2410.21276
-
Hybrid Memetic Search for Electric Vehicle Routing with Time Windows, Simultaneous Pickup-Delivery, and Partial Recharges 25 Oct 2024 · 0 repositories · arXiv:2410.19580
-
Integrating Large Language Models with Internet of Things Applications 25 Oct 2024 · 0 repositories · arXiv:2410.19223
-
Investigating the Role of Prompting and External Tools in Hallucination Rates of Large Language Models 25 Oct 2024 · 0 repositories · arXiv:2410.19385
-
KAHANI: Culturally-Nuanced Visual Storytelling Pipeline for Non-Western Cultures 25 Oct 2024 · 0 repositories · arXiv:2410.19419
-
Multi-Agent Reinforcement Learning with Selective State-Space Models 25 Oct 2024 · 0 repositories · arXiv:2410.19382
-
Multi-Class Abnormality Classification Task in Video Capsule Endoscopy 25 Oct 2024 · 1 repository · arXiv:2410.19973
-
Not All Heads Matter: A Head-Level KV Cache Compression Method with Integrated Retrieval and Reasoning 25 Oct 2024 · 1 repository · arXiv:2410.19258Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Provable optimal transport with transformers: The essence of depth and prompt engineering 25 Oct 2024 · 1 repository · arXiv:2410.19931
-
Robot Behavior Personalization from Sparse User Feedback 25 Oct 2024 · 0 repositories · arXiv:2410.19219
-
RobustKV: Defending Large Language Models against Jailbreak Attacks via KV Eviction 25 Oct 2024 · 0 repositories · arXiv:2410.19937