Methods › General › Attention Mechanisms › Attention › Papers, page 95
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 95 of 316: papers 9,401 to 9,500 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
The Lou Dataset -- Exploring the Impact of Gender-Fair Language in German Text Classification 26 Sep 2024 · 0 repositories · arXiv:2409.17929
-
Towards the Mitigation of Confirmation Bias in Semi-supervised Learning: a Debiased Training Perspective 26 Sep 2024 · 0 repositories · arXiv:2409.18316
-
Trustworthy Text-to-Image Diffusion Models: A Timely and Focused Survey 26 Sep 2024 · 1 repository · arXiv:2409.18214
-
Unifying Dimensions: A Linear Adaptive Approach to Lightweight Image Super-Resolution 26 Sep 2024 · 1 repository · arXiv:2409.17597
-
Unsupervised Learning Based Multi-Scale Exposure Fusion 26 Sep 2024 · 0 repositories · arXiv:2409.17830
-
A Prompting-Based Representation Learning Method for Recommendation with Large Language Models 25 Sep 2024 · 0 repositories · arXiv:2409.16674
-
A Roadmap for Embodied and Social Grounding in LLMs 25 Sep 2024 · 0 repositories · arXiv:2409.16900
-
Accelerating Multi-Block Constrained Optimization Through Learning to Optimize 25 Sep 2024 · 0 repositories · arXiv:2409.17320
-
AgRegNet: A Deep Regression Network for Flower and Fruit Density Estimation, Localization, and Counting in Orchards 25 Sep 2024 · 0 repositories · arXiv:2409.17400
-
AlignedKV: Reducing Memory Access of KV-Cache with Precision-Aligned Quantization 25 Sep 2024 · 1 repository · arXiv:2409.16546
-
Targeted Neural Architectures in Multi-Objective Frameworks for Complete Glioma Characterization from Multimodal MRI 25 Sep 2024 · 0 repositories · arXiv:2409.17273
-
Attention Prompting on Image for Large Vision-Language Models 25 Sep 2024 · 1 repository · arXiv:2409.17143Syntology official (archive's flag): 6 ran · 8 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
Beyond Turing Test: Can GPT-4 Sway Experts' Decisions? 25 Sep 2024 · 0 repositories · arXiv:2409.16710
-
Block Expanded DINORET: Adapting Natural Domain Foundation Models for Retinal Imaging Without Catastrophic Forgetting 25 Sep 2024 · 0 repositories · arXiv:2409.17332
-
CodeInsight: A Curated Dataset of Practical Coding Solutions from Stack Overflow 25 Sep 2024 · 1 repository · arXiv:2409.16819
-
Deep Learning and Machine Learning, Advancing Big Data Analytics and Management: Handy Appetizer 25 Sep 2024 · 0 repositories · arXiv:2409.17120
-
Discovering the Gems in Early Layers: Accelerating Long-Context LLMs with 1000x Input Token Reduction 25 Sep 2024 · 1 repository · arXiv:2409.17422Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 1 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples) · 3 pointer-only (licence)
-
Enhancing Automatic Keyphrase Labelling with Text-to-Text Transfer Transformer (T5) Architecture: A Framework for Keyphrase Generation and Filtering 25 Sep 2024 · 0 repositories · arXiv:2409.16760
-
Enhancing Feature Selection and Interpretability in AI Regression Tasks Through Feature Attribution 25 Sep 2024 · 0 repositories · arXiv:2409.16787
-
Event-Triggered Non-Linear Control of Offshore MMC Grids for Asymmetrical AC Faults 25 Sep 2024 · 0 repositories · arXiv:2409.16743
-
Going Beyond U-Net: Assessing Vision Transformers for Semantic Segmentation in Microscopy Image Analysis 25 Sep 2024 · 0 repositories · arXiv:2409.16940
-
Gradient Boosting Decision Trees on Medical Diagnosis over Tabular Data 25 Sep 2024 · 1 repository · arXiv:2410.03705
-
Harnessing Diversity for Important Data Selection in Pretraining Large Language Models 25 Sep 2024 · 0 repositories · arXiv:2409.16986
-
HVT: A Comprehensive Vision Framework for Learning in Non-Euclidean Space 25 Sep 2024 · 1 repository · arXiv:2409.16897
-
INT-FlashAttention: Enabling Flash Attention for INT8 Quantization 25 Sep 2024 · 1 repository · arXiv:2409.16997
-
Investigating OCR-Sensitive Neurons to Improve Entity Recognition in Historical Documents 25 Sep 2024 · 1 repository · arXiv:2409.16934
-
LLaMa-SciQ: An Educational Chatbot for Answering Science MCQ 25 Sep 2024 · 0 repositories · arXiv:2409.16779
-
MCI-GRU: Stock Prediction Model Based on Multi-Head Cross-Attention and Improved GRU 25 Sep 2024 · 0 repositories · arXiv:2410.20679
-
Near-Field Multipath MIMO Channel Model for Imperfect Surface Reflection 25 Sep 2024 · 0 repositories · arXiv:2409.17041
-
Non-asymptotic Convergence of Training Transformers for Next-token Prediction 25 Sep 2024 · 0 repositories · arXiv:2409.17335
-
Non-stationary BERT: Exploring Augmented IMU Data For Robust Human Activity Recognition 25 Sep 2024 · 0 repositories · arXiv:2409.16730
-
Post-hoc Reward Calibration: A Case Study on Length Bias 25 Sep 2024 · 1 repository · arXiv:2409.17407Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Pre-trained Graphformer-based Ranking at Web-scale Search (Extended Abstract) 25 Sep 2024 · 0 repositories · arXiv:2409.16590
-
Probing Omissions and Distortions in Transformer-based RDF-to-Text Models 25 Sep 2024 · 0 repositories · arXiv:2409.16707
-
Quantum-Classical Sentiment Analysis 25 Sep 2024 · 0 repositories · arXiv:2409.16928
-
Severity Prediction in Mental Health: LLM-based Creation, Analysis, Evaluation of a Novel Multilingual Dataset 25 Sep 2024 · 0 repositories · arXiv:2409.17397
-
SWE2: SubWord Enriched and Significant Word Emphasized Framework for Hate Speech Detection 25 Sep 2024 · 0 repositories · arXiv:2409.16673
-
The Credibility Transformer 25 Sep 2024 · 0 repositories · arXiv:2409.16653
-
The Overfocusing Bias of Convolutional Neural Networks: A Saliency-Guided Regularization Approach 25 Sep 2024 · 0 repositories · arXiv:2409.17370
-
The Role of Language Models in Modern Healthcare: A Comprehensive Review 25 Sep 2024 · 0 repositories · arXiv:2409.16860
-
Trading through Earnings Seasons using Self-Supervised Contrastive Representation Learning 25 Sep 2024 · 0 repositories · arXiv:2409.17392
-
Using LLM for Real-Time Transcription and Summarization of Doctor-Patient Interactions into ePuskesmas in Indonesia 25 Sep 2024 · 0 repositories · arXiv:2409.17054
-
Zero-Shot Detection of LLM-Generated Text using Token Cohesiveness 25 Sep 2024 · 1 repository · arXiv:2409.16914Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
A Comprehensive Evaluation of Large Language Models on Mental Illnesses 24 Sep 2024 · 0 repositories · arXiv:2409.15687
-
AI Can Be Cognitively Biased: An Exploratory Study on Threshold Priming in LLM-Based Batch Relevance Assessment 24 Sep 2024 · 0 repositories · arXiv:2409.16022
-
Beyond Text-to-Text: An Overview of Multimodal and Generative Artificial Intelligence for Education Using Topic Modeling 24 Sep 2024 · 0 repositories · arXiv:2409.16376
-
Controlling Risk of Retrieval-augmented Generation: A Counterfactual Prompting Framework 24 Sep 2024 · 1 repository · arXiv:2409.16146Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Data Augmentation for Sparse Multidimensional Learning Performance Data Using Generative AI 24 Sep 2024 · 1 repository · arXiv:2409.15631
-
Disentangled Generation and Aggregation for Robust Radiance Fields 24 Sep 2024 · 0 repositories · arXiv:2409.15715
-
dnaGrinder: a lightweight and high-capacity genomic foundation model 24 Sep 2024 · 0 repositories · arXiv:2409.15697
-
Do the Right Thing, Just Debias! Multi-Category Bias Mitigation Using LLMs 24 Sep 2024 · 0 repositories · arXiv:2409.16371
-
Double-Path Adaptive-correlation Spatial-Temporal Inverted Transformer for Stock Time Series Forecasting 24 Sep 2024 · 0 repositories · arXiv:2409.15662
-
Ducho meets Elliot: Large-scale Benchmarks for Multimodal Recommendation 24 Sep 2024 · 1 repository · arXiv:2409.15857
-
Effectiveness of Cross-linguistic Extraction of Genetic Information using Generative Large Language Models 24 Sep 2024 · 1 repository
-
From Pixels to Words: Leveraging Explainability in Face Recognition through Interactive Natural Language Processing 24 Sep 2024 · 0 repositories · arXiv:2409.16089
-
GS-Net: Global Self-Attention Guided CNN for Multi-Stage Glaucoma Classification 24 Sep 2024 · 0 repositories · arXiv:2409.16082
-
IRSC: A Zero-shot Evaluation Benchmark for Information Retrieval through Semantic Comprehension in Retrieval-Augmented Generation Scenarios 24 Sep 2024 · 1 repository · arXiv:2409.15763
-
Language-based Audio Moment Retrieval 24 Sep 2024 · 1 repository · arXiv:2409.15672
-
Lessons and Insights from a Unifying Study of Parameter-Efficient Fine-Tuning (PEFT) in Visual Recognition 24 Sep 2024 · 2 repositories · arXiv:2409.16434
-
Lighter And Better: Towards Flexible Context Adaptation For Retrieval Augmented Generation 24 Sep 2024 · 0 repositories · arXiv:2409.15699
-
Looped Transformers for Length Generalization 24 Sep 2024 · 1 repository · arXiv:2409.15647Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Making Text Embedders Few-Shot Learners 24 Sep 2024 · 1 repository · arXiv:2409.15700
-
MaskBit: Embedding-free Image Generation via Bit Tokens 24 Sep 2024 · 1 repository · arXiv:2409.16211Syntology official (archive's flag): 8 ran · 8 ran (of which 1 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
MonoFormer: One Transformer for Both Diffusion and Autoregression 24 Sep 2024 · 1 repository · arXiv:2409.16280Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Neuromorphic Drone Detection: an Event-RGB Multimodal Approach 24 Sep 2024 · 1 repository · arXiv:2409.16099
-
Predicting Deterioration in Mild Cognitive Impairment with Survival Transformers, Extreme Gradient Boosting and Cox Proportional Hazard Modelling 24 Sep 2024 · 0 repositories · arXiv:2409.16231
-
Predicting Distance matrix with large language models 24 Sep 2024 · 0 repositories · arXiv:2409.16333
-
Selection of Prompt Engineering Techniques for Code Generation through Predicting Code Complexity 24 Sep 2024 · 0 repositories · arXiv:2409.16416
-
Self-attention as an attractor network: transient memories without backpropagation 24 Sep 2024 · 1 repository · arXiv:2409.16112
-
Sex Differences in Hierarchical and Modular Organization of Functional Brain Networks: Insights from Hierarchical Entropy and Modularity Analysis 24 Sep 2024 · 0 repositories · arXiv:2409.15833
-
Small Language Models: Survey, Measurements, and Insights 24 Sep 2024 · 1 repository · arXiv:2409.15790
-
Supervised Fine-Tuning Achieve Rapid Task Adaption Via Alternating Attention Head Activation Patterns 24 Sep 2024 · 0 repositories · arXiv:2409.15820
-
SurgIRL: Towards Life-Long Learning for Surgical Automation by Incremental Reinforcement Learning 24 Sep 2024 · 0 repositories · arXiv:2409.15651
-
SwiftDossier: Tailored Automatic Dossier for Drug Discovery with LLMs and Agents 24 Sep 2024 · 0 repositories · arXiv:2409.15817
-
Synatra: Turning Indirect Knowledge into Direct Demonstrations for Digital Agents at Scale 24 Sep 2024 · 0 repositories · arXiv:2409.15637
-
Task-oriented Prompt Enhancement via Script Generation 24 Sep 2024 · 0 repositories · arXiv:2409.16418
-
TiM4Rec: An Efficient Sequential Recommendation Model Based on Time-Aware Structured State Space Duality Model 24 Sep 2024 · 1 repository · arXiv:2409.16182
-
Towards Explainable Graph Neural Networks for Neurological Evaluation on EEG Signals 24 Sep 2024 · 0 repositories · arXiv:2410.07199
-
Transformer based time series prediction of the maximum power point for solar photovoltaic cells 24 Sep 2024 · 0 repositories · arXiv:2409.16342
-
Underground Mapping and Localization Based on Ground-Penetrating Radar 24 Sep 2024 · 0 repositories · arXiv:2409.16446
-
Unsupervised Attention Regularization Based Domain Adaptation for Oracle Character Recognition 24 Sep 2024 · 0 repositories · arXiv:2409.15893
-
VideoPatchCore: An Effective Method to Memorize Normality for Video Anomaly Detection 24 Sep 2024 · 1 repository · arXiv:2409.16225
-
WeSep: A Scalable and Flexible Toolkit Towards Generalizable Target Speaker Extraction 24 Sep 2024 · 1 repository · arXiv:2409.15799
-
XTRUST: On the Multilingual Trustworthiness of Large Language Models 24 Sep 2024 · 1 repository · arXiv:2409.15762
-
A Preliminary Study of o1 in Medicine: Are We Closer to an AI Doctor? 23 Sep 2024 · 0 repositories · arXiv:2409.15277
-
A-VL: Adaptive Attention for Large Vision-Language Models 23 Sep 2024 · 1 repository · arXiv:2409.14846
-
Advancing Depression Detection on Social Media Platforms Through Fine-Tuned Large Language Models 23 Sep 2024 · 0 repositories · arXiv:2409.14794
-
AEANet: Affinity Enhanced Attentional Networks for Arbitrary Style Transfer 23 Sep 2024 · 0 repositories · arXiv:2409.14652
-
CA-MHFA: A Context-Aware Multi-Head Factorized Attentive Pooling for SSL-Based Speaker Verification 23 Sep 2024 · 0 repositories · arXiv:2409.15234
-
Chattronics: using GPTs to assist in the design of data acquisition systems 23 Sep 2024 · 0 repositories · arXiv:2409.15183
-
Clinical-grade Multi-Organ Pathology Report Generation for Multi-scale Whole Slide Images via a Semantically Guided Medical Text Foundation Model 23 Sep 2024 · 1 repository · arXiv:2409.15574
-
Curb Your Attention: Causal Attention Gating for Robust Trajectory Prediction in Autonomous Driving 23 Sep 2024 · 0 repositories · arXiv:2410.07191
-
Deep Cost Ray Fusion for Sparse Depth Video Completion 23 Sep 2024 · 0 repositories · arXiv:2409.14935
-
Deep Reinforcement Learning-based Obstacle Avoidance for Robot Movement in Warehouse Environments 23 Sep 2024 · 0 repositories · arXiv:2409.14972
-
DepthART: Monocular Depth Estimation as Autoregressive Refinement Task 23 Sep 2024 · 0 repositories · arXiv:2409.15010
-
Designing Pre-training Datasets from Unlabeled Data for EEG Classification with Transformers 23 Sep 2024 · 0 repositories · arXiv:2410.07190
-
Diffusion-based RGB-D Semantic Segmentation with Deformable Attention Transformer 23 Sep 2024 · 0 repositories · arXiv:2409.15117
-
Dual Stream Graph Transformer Fusion Networks for Enhanced Brain Decoding 23 Sep 2024 · 0 repositories · arXiv:2410.07189
-
Dumpling GNN: Hybrid GNN Enables Better ADC Payload Activity Prediction Based on Chemical Structure 23 Sep 2024 · 0 repositories · arXiv:2410.05278
-
EDGE-Rec: Efficient and Data-Guided Edge Diffusion For Recommender Systems Graphs 23 Sep 2024 · 0 repositories · arXiv:2409.14689