Methods › General › Attention Mechanisms › Attention › Papers, page 83
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 83 of 316: papers 8,201 to 8,300 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
TabDPT: Scaling Tabular Foundation Models 23 Oct 2024 · 1 repository · arXiv:2410.18164Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
TAGE: Trustworthy Attribute Group Editing for Stable Few-shot Image Generation 23 Oct 2024 · 0 repositories · arXiv:2410.17855
-
TP-Eval: Tap Multimodal LLMs' Potential in Evaluation by Customizing Prompts 23 Oct 2024 · 0 repositories · arXiv:2410.18071
-
TranSPORTmer: A Holistic Approach to Trajectory Understanding in Multi-Agent Sports 23 Oct 2024 · 0 repositories · arXiv:2410.17785
-
Value Residual Learning For Alleviating Attention Concentration In Transformers 23 Oct 2024 · 1 repository · arXiv:2410.17897Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
Which Client is Reliable?: A Reliable and Personalized Prompt-based Federated Learning for Medical Image Question Answering 23 Oct 2024 · 0 repositories · arXiv:2410.17484
-
A Bayesian Perspective on the Maximum Score Problem 22 Oct 2024 · 0 repositories · arXiv:2410.17153
-
A Statistical Analysis of LLMs' Self-Evaluation Using Proverbs 22 Oct 2024 · 0 repositories · arXiv:2410.16640
-
An Eye for an AI: Evaluating GPT-4o's Visual Perception Skills and Geometric Reasoning Skills Using Computer Graphics Questions 22 Oct 2024 · 0 repositories · arXiv:2410.16991
-
Assessment of Transformer-Based Encoder-Decoder Model for Human-Like Summarization 22 Oct 2024 · 0 repositories · arXiv:2410.16842
-
Audio-to-Score Conversion Model Based on Whisper methodology 22 Oct 2024 · 0 repositories · arXiv:2410.17209
-
Automated Spinal MRI Labelling from Reports Using a Large Language Model 22 Oct 2024 · 1 repository · arXiv:2410.17235
-
Captions Speak Louder than Images (CASLIE): Generalizing Foundation Models for E-commerce from High-quality Multimodal Instruction Data 22 Oct 2024 · 0 repositories · arXiv:2410.17337
-
Dhoroni: Exploring Bengali Climate Change and Environmental Views with a Multi-Perspective News Dataset and Natural Language Processing 22 Oct 2024 · 1 repository · arXiv:2410.17225
-
DI-MaskDINO: A Joint Object Detection and Instance Segmentation Model 22 Oct 2024 · 1 repository · arXiv:2410.16707
-
Distill-SynthKG: Distilling Knowledge Graph Synthesis Workflow for Improved Coverage and Efficiency 22 Oct 2024 · 0 repositories · arXiv:2410.16597
-
DNAHLM -- DNA sequence and Human Language mixed large language Model 22 Oct 2024 · 1 repository · arXiv:2410.16917
-
Do Vision-Language Models Represent Space and How? Evaluating Spatial Frame of Reference Under Ambiguities 22 Oct 2024 · 1 repository · arXiv:2410.17385Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Dynamic graph neural networks for enhanced volatility prediction in financial markets 22 Oct 2024 · 0 repositories · arXiv:2410.16858
-
Efficient Feature Extraction Using Light-Weight CNN Attention-Based Deep Learning Architectures for Ultrasound Fetal Plane Classification 22 Oct 2024 · 0 repositories · arXiv:2410.17396
-
End-to-End Optimization and Learning of Fair Court Schedules 22 Oct 2024 · 0 repositories · arXiv:2410.17415
-
EnvBridge: Bridging Diverse Environments with Cross-Environment Knowledge Transfer for Embodied AI 22 Oct 2024 · 0 repositories · arXiv:2410.16919
-
Exploring Possibilities of AI-Powered Legal Assistance in Bangladesh through Large Language Modeling 22 Oct 2024 · 1 repository · arXiv:2410.17210
-
FairLoRA: Unpacking Bias Mitigation in Vision Models with Fairness-Driven Low-Rank Adaptation 22 Oct 2024 · 0 repositories · arXiv:2410.17358
-
FastAttention: Extend FlashAttention2 to NPUs and Low-resource GPUs 22 Oct 2024 · 0 repositories · arXiv:2410.16663
-
From Attention to Activation: Unravelling the Enigmas of Large Language Models 22 Oct 2024 · 0 repositories · arXiv:2410.17174
-
Graph Transformers Dream of Electric Flow 22 Oct 2024 · 0 repositories · arXiv:2410.16699
-
High-Order Associative Learning Based on Memristive Circuits for Efficient Learning 22 Oct 2024 · 1 repository · arXiv:2410.16734
-
In Context Learning and Reasoning for Symbolic Regression with Large Language Models 22 Oct 2024 · 1 repository · arXiv:2410.17448
-
Interchangeable Token Embeddings for Extendable Vocabulary and Alpha-Equivalence 22 Oct 2024 · 0 repositories · arXiv:2410.17161
-
LiNo: Advancing Recursive Residual Decomposition of Linear and Nonlinear Patterns for Robust Time Series Forecasting 22 Oct 2024 · 1 repository · arXiv:2410.17159
-
Methods of improving LLM training stability 22 Oct 2024 · 0 repositories · arXiv:2410.16682
-
Optimal Design for Reward Modeling in RLHF 22 Oct 2024 · 0 repositories · arXiv:2410.17055
-
Order Matters: Exploring Order Sensitivity in Multimodal Large Language Models 22 Oct 2024 · 0 repositories · arXiv:2410.16983
-
PLDR-LLM: Large Language Model from Power Law Decoder Representations 22 Oct 2024 · 2 repositories · arXiv:2410.16703
-
Representation Shattering in Transformers: A Synthetic Study with Knowledge Editing 22 Oct 2024 · 0 repositories · arXiv:2410.17194Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Scattered Forest Search: Smarter Code Space Exploration with LLMs 22 Oct 2024 · 0 repositories · arXiv:2411.05010
-
SmartRAG: Jointly Learn RAG-Related Tasks From the Environment Feedback 22 Oct 2024 · 0 repositories · arXiv:2410.18141
-
SpikMamba: When SNN meets Mamba in Event-based Human Action Recognition 22 Oct 2024 · 1 repository · arXiv:2410.16746
-
Tracing the Development of the Virtual Particle Concept Using Semantic Change Detection 22 Oct 2024 · 1 repository · arXiv:2410.16855
-
A Dual Process VLA: Efficient Robotic Manipulation Leveraging VLM 21 Oct 2024 · 0 repositories · arXiv:2410.15549
-
A Fusion-Driven Approach of Attention-Based CNN-BiLSTM for Protein Family Classification -- ProFamNet 21 Oct 2024 · 0 repositories · arXiv:2410.17293
-
A Troublemaker with Contagious Jailbreak Makes Chaos in Honest Towns 21 Oct 2024 · 0 repositories · arXiv:2410.16155
-
AlignVSR: Audio-Visual Cross-Modal Alignment for Visual Speech Recognition 21 Oct 2024 · 1 repository · arXiv:2410.16438
-
All You Need is an Improving Column: Enhancing Column Generation for Parallel Machine Scheduling via Transformers 21 Oct 2024 · 0 repositories · arXiv:2410.15601
-
An Efficient System for Automatic Map Storytelling -- A Case Study on Historical Maps 21 Oct 2024 · 1 repository · arXiv:2410.15780
-
An Explainable Contrastive-based Dilated Convolutional Network with Transformer for Pediatric Pneumonia Detection 21 Oct 2024 · 0 repositories · arXiv:2410.16143
-
Arithmetic Transformers Can Length-Generalize in Both Operand Length and Count 21 Oct 2024 · 1 repository · arXiv:2410.15787
-
Beyond 2:4: exploring V:N:M sparsity for efficient transformer inference on GPUs 21 Oct 2024 · 0 repositories · arXiv:2410.16135
-
Boosting Jailbreak Transferability for Large Language Models 21 Oct 2024 · 1 repository · arXiv:2410.15645
-
Building A Coding Assistant via the Retrieval-Augmented Language Model 21 Oct 2024 · 1 repository · arXiv:2410.16229
-
CamI2V: Camera-Controlled Image-to-Video Diffusion Model 21 Oct 2024 · 1 repository · arXiv:2410.15957Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
CartesianMoE: Boosting Knowledge Sharing among Experts via Cartesian Product Routing in Mixture-of-Experts 21 Oct 2024 · 1 repository · arXiv:2410.16077
-
CausalGraph2LLM: Evaluating LLMs for Causal Queries 21 Oct 2024 · 1 repository · arXiv:2410.15939
-
CompassJudger-1: All-in-one Judge Model Helps Model Evaluation and Evolution 21 Oct 2024 · 1 repository · arXiv:2410.16256
-
Deep Graph Attention Networks 21 Oct 2024 · 1 repository · arXiv:2410.15640
-
Deep Learning and Data Augmentation for Detecting Self-Admitted Technical Debt 21 Oct 2024 · 1 repository · arXiv:2410.15804
-
Developing Retrieval Augmented Generation (RAG) based LLM Systems from PDFs: An Experience Report 21 Oct 2024 · 1 repository · arXiv:2410.15944
-
Diffusion Transformer Policy 21 Oct 2024 · 1 repository · arXiv:2410.15959
-
Disambiguating Monocular Reconstruction of 3D Clothed Human with Spatial-Temporal Transformer 21 Oct 2024 · 0 repositories · arXiv:2410.16337
-
Catastrophic Failure of LLM Unlearning via Quantization 21 Oct 2024 · 1 repository · arXiv:2410.16454Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Domain-Adaptive Pre-training of Self-Supervised Foundation Models for Medical Image Classification in Gastrointestinal Endoscopy 21 Oct 2024 · 1 repository · arXiv:2410.21302
-
Enabling Energy-Efficient Deployment of Large Language Models on Memristor Crossbar: A Synergy of Large and Small 21 Oct 2024 · 0 repositories · arXiv:2410.15977
-
Enhancing SNN-based Spatio-Temporal Learning: A Benchmark Dataset and Cross-Modality Attention Model 21 Oct 2024 · 1 repository · arXiv:2410.15689
-
ExDBN: Exact learning of Dynamic Bayesian Networks 21 Oct 2024 · 0 repositories · arXiv:2410.16100
-
Exploring Pretraining via Active Forgetting for Improving Cross Lingual Transfer for Decoder Language Models 21 Oct 2024 · 0 repositories · arXiv:2410.16168
-
Focus Where It Matters: Graph Selective State Focused Attention Networks 21 Oct 2024 · 0 repositories · arXiv:2410.15849
-
FusionLungNet: Multi-scale Fusion Convolution with Refinement Network for Lung CT Image Segmentation 21 Oct 2024 · 2 repositories · arXiv:2410.15812
-
Generalized Probabilistic Attention Mechanism in Transformers 21 Oct 2024 · 0 repositories · arXiv:2410.15578
-
Generalizing Motion Planners with Mixture of Experts for Autonomous Driving 21 Oct 2024 · 1 repository · arXiv:2410.15774Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
GReFEL: Geometry-Aware Reliable Facial Expression Learning under Bias and Imbalanced Data Distribution 21 Oct 2024 · 0 repositories · arXiv:2410.15927
-
Guardians of Discourse: Evaluating LLMs on Multilingual Offensive Language Detection 21 Oct 2024 · 0 repositories · arXiv:2410.15623
-
Improving Neuron-level Interpretability with White-box Language Models 21 Oct 2024 · 0 repositories · arXiv:2410.16443
-
Large Body Language Models 21 Oct 2024 · 0 repositories · arXiv:2410.16533
-
Large Language Models in Computer Science Education: A Systematic Literature Review 21 Oct 2024 · 1 repository · arXiv:2410.16349
-
Leveraging Retrieval-Augmented Generation for Culturally Inclusive Hakka Chatbots: Design Insights and User Perceptions 21 Oct 2024 · 0 repositories · arXiv:2410.15572
-
LightFusionRec: Lightweight Transformers-Based Cross-Domain Recommendation Model 21 Oct 2024 · 0 repositories · arXiv:2410.15656
-
LMHaze: Intensity-aware Image Dehazing with a Large-scale Multi-intensity Real Haze Dataset 21 Oct 2024 · 1 repository · arXiv:2410.16095
-
Long-distance Geomagnetic Navigation in GNSS-denied Environments with Deep Reinforcement Learning 21 Oct 2024 · 0 repositories · arXiv:2410.15837
-
MagicPIG: LSH Sampling for Efficient LLM Generation 21 Oct 2024 · 1 repository · arXiv:2410.16179Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Mesa-Extrapolation: A Weave Position Encoding Method for Enhanced Extrapolation in LLMs 21 Oct 2024 · 1 repository · arXiv:2410.15859Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Mitigating Object Hallucination via Concentric Causal Attention 21 Oct 2024 · 1 repository · arXiv:2410.15926Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Modelling Concurrent RTP Flows for End-to-end Predictions of QoS in Real Time Communications 21 Oct 2024 · 0 repositories · arXiv:2410.15846
-
MoRE: Multi-Modal Contrastive Pre-training with Transformers on X-Rays, ECGs, and Diagnostic Report 21 Oct 2024 · 1 repository · arXiv:2410.16239
-
Multi-head Sequence Tagging Model for Grammatical Error Correction 21 Oct 2024 · 2 repositories · arXiv:2410.16473
-
Natural GaLore: Accelerating GaLore for memory-efficient LLM Training and Fine-tuning 21 Oct 2024 · 1 repository · arXiv:2410.16029
-
Object-Centric Temporal Consistency via Conditional Autoregressive Inductive Biases 21 Oct 2024 · 0 repositories · arXiv:2410.15728
-
On Creating an English-Thai Code-switched Machine Translation in Medical Domain 21 Oct 2024 · 1 repository · arXiv:2410.16221
-
On the Design and Performance of Machine Learning Based Error Correcting Decoders 21 Oct 2024 · 0 repositories · arXiv:2410.15899
-
Promoting cross-modal representations to improve multimodal foundation models for physiological signals 21 Oct 2024 · 0 repositories · arXiv:2410.16424
-
RAG4ITOps: A Supervised Fine-Tunable and Comprehensive RAG Framework for IT Operations and Maintenance 21 Oct 2024 · 0 repositories · arXiv:2410.15805
-
Reflection-Bench: probing AI intelligence with reflection 21 Oct 2024 · 1 repository · arXiv:2410.16270
-
Revealing and Mitigating the Local Pattern Shortcuts of Mamba 21 Oct 2024 · 1 repository · arXiv:2410.15678
-
Rulebreakers Challenge: Revealing a Blind Spot in Large Language Models' Reasoning with Formal Logic 21 Oct 2024 · 0 repositories · arXiv:2410.16502
-
SeisLM: a Foundation Model for Seismic Waveforms 21 Oct 2024 · 1 repository · arXiv:2410.15765
-
Self-Explained Keywords Empower Large Language Models for Code Generation 21 Oct 2024 · 0 repositories · arXiv:2410.15966
-
SSMT: Few-Shot Traffic Forecasting with Single Source Meta-Transfer 21 Oct 2024 · 0 repositories · arXiv:2410.15589
-
Students Rather Than Experts: A New AI For Education Pipeline To Model More Human-Like And Personalised Early Adolescences 21 Oct 2024 · 0 repositories · arXiv:2410.15701
-
Surprising Patterns in Musical Influence Networks 21 Oct 2024 · 0 repositories · arXiv:2410.15996
-
TimeMixer++: A General Time Series Pattern Machine for Universal Predictive Analysis 21 Oct 2024 · 2 repositories · arXiv:2410.16032