Methods › General › Attention Mechanisms › Attention › Papers, page 51
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 51 of 316: papers 5,001 to 5,100 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Training Dialogue Systems by AI Feedback for Improving Overall Dialogue Impression 22 Jan 2025 · 0 repositories · arXiv:2501.12698
-
Unified CNNs and transformers underlying learning mechanism reveals multi-head attention modus vivendi 22 Jan 2025 · 0 repositories · arXiv:2501.12900
-
A Hybrid Attention Framework for Fake News Detection with Large Language Models 21 Jan 2025 · 0 repositories · arXiv:2501.11967
-
A Hybrid Supervised and Self-Supervised Graph Neural Network for Edge-Centric Applications 21 Jan 2025 · 0 repositories · arXiv:2501.12309
-
A Survey of Graph Retrieval-Augmented Generation for Customized Large Language Models 21 Jan 2025 · 1 repository · arXiv:2501.13958
-
Academic Case Reports Lack Diversity: Assessing the Presence and Diversity of Sociodemographic and Behavioral Factors related to Post COVID-19 Condition 21 Jan 2025 · 0 repositories · arXiv:2501.12538
-
Advancing the Understanding and Evaluation of AR-Generated Scenes: When Vision-Language Models Shine and Stumble 21 Jan 2025 · 1 repository · arXiv:2501.13964
-
ALoFTRAG: Automatic Local Fine Tuning for Retrieval Augmented Generation 21 Jan 2025 · 1 repository · arXiv:2501.11929
-
Assisting Mathematical Formalization with A Learning-based Premise Retriever 21 Jan 2025 · 1 repository · arXiv:2501.13959
-
Automatic Labelling with Open-source LLMs using Dynamic Label Schema Integration 21 Jan 2025 · 0 repositories · arXiv:2501.12332
-
Automatic selection of the best neural architecture for time series forecasting via multi-objective optimization and Pareto optimality conditions 21 Jan 2025 · 0 repositories · arXiv:2501.12215
-
Bridging the Communication Gap: Evaluating AI Labeling Practices for Trustworthy AI Development 21 Jan 2025 · 1 repository · arXiv:2501.11909
-
Coarse-to-Fine Lightweight Meta-Embedding for ID-Based Recommendation 21 Jan 2025 · 1 repository · arXiv:2501.11870
-
Comparative Approaches to Sentiment Analysis Using Datasets in Major European and Arabic Languages 21 Jan 2025 · 0 repositories · arXiv:2501.12540
-
Continuous 3D Perception Model with Persistent State 21 Jan 2025 · 0 repositories · arXiv:2501.12387
-
CroMe: Multimodal Fake News Detection using Cross-Modal Tri-Transformer and Metric Learning 21 Jan 2025 · 0 repositories · arXiv:2501.12422
-
Divide-Then-Aggregate: An Efficient Tool Learning Method via Parallel Tool Invocation 21 Jan 2025 · 0 repositories · arXiv:2501.12432
-
DLEN: Dual Branch of Transformer for Low-Light Image Enhancement in Dual Domains 21 Jan 2025 · 0 repositories · arXiv:2501.12235
-
Efficient Lung Ultrasound Severity Scoring Using Dedicated Feature Extractor 21 Jan 2025 · 1 repository · arXiv:2501.12524
-
Enhancing Retrosynthesis with Conformer: A Template-Free Method 21 Jan 2025 · 0 repositories · arXiv:2501.12434
-
Episodic Memories Generation and Evaluation Benchmark for Large Language Models 21 Jan 2025 · 1 repository · arXiv:2501.13121Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Fact-Preserved Personalized News Headline Generation 21 Jan 2025 · 1 repository · arXiv:2501.11828
-
PAINT: Paying Attention to INformed Tokens to Mitigate Hallucination in Large Vision-Language Model 21 Jan 2025 · 2 repositories · arXiv:2501.12206
-
FNIN: A Fourier Neural Operator-based Numerical Integration Network for Surface-form-gradients 21 Jan 2025 · 1 repository · arXiv:2501.11876
-
FOCUS: First Order Concentrated Updating Scheme 21 Jan 2025 · 0 repositories · arXiv:2501.12243
-
Generating Plausible Distractors for Multiple-Choice Questions via Student Choice Prediction 21 Jan 2025 · 0 repositories · arXiv:2501.13125
-
Harnessing Generative Pre-Trained Transformer for Datacenter Packet Trace Generation 21 Jan 2025 · 0 repositories · arXiv:2501.12033
-
Med-R²: Crafting Trustworthy LLM Physicians via Retrieval and Reasoning of Evidence-Based Medicine 21 Jan 2025 · 1 repository · arXiv:2501.11885
-
Multi-Modality Collaborative Learning for Sentiment Analysis 21 Jan 2025 · 1 repository · arXiv:2501.12424
-
Network-informed Prompt Engineering against Organized Astroturf Campaigns under Extreme Class Imbalance 21 Jan 2025 · 1 repository · arXiv:2501.11849
-
Noise-Resilient Point-wise Anomaly Detection in Time Series Using Weak Segment Labels 21 Jan 2025 · 1 repository · arXiv:2501.11959
-
Panoramic Interests: Stylistic-Content Aware Personalized Headline Generation 21 Jan 2025 · 1 repository · arXiv:2501.11900
-
Parallel Sequence Modeling via Generalized Spatial Propagation Network 21 Jan 2025 · 0 repositories · arXiv:2501.12381
-
Progressive Cross Attention Network for Flood Segmentation using Multispectral Satellite Imagery 21 Jan 2025 · 0 repositories · arXiv:2501.11923
-
Pushing the Limits of BFP on Narrow Precision LLM Inference 21 Jan 2025 · 0 repositories · arXiv:2502.00026
-
Rate-Aware Learned Speech Compression 21 Jan 2025 · 0 repositories · arXiv:2501.11999
-
Slot-BERT: Self-supervised Object Discovery in Surgical Video 21 Jan 2025 · 0 repositories · arXiv:2501.12477
-
SMamba: Sparse Mamba for Event-based Object Detection 21 Jan 2025 · 1 repository · arXiv:2501.11971Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 4 harvested samples) · 4 pointer-only (licence)
-
Speech Enhancement with Overlapped-Frame Information Fusion and Causal Self-Attention 21 Jan 2025 · 1 repository · arXiv:2501.12004
-
SVGS-DSGAT: An IoT-Enabled Innovation in Underwater Robotic Object Detection Technology 21 Jan 2025 · 0 repositories · arXiv:2501.12169
-
Test-time regression: a unifying framework for designing sequence models with associative memory 21 Jan 2025 · 0 repositories · arXiv:2501.12352
-
TokenVerse: Versatile Multi-concept Personalization in Token Modulation Space 21 Jan 2025 · 0 repositories · arXiv:2501.12224
-
Toward Scalable Graph Unlearning: A Node Influence Maximization based Approach 21 Jan 2025 · 0 repositories · arXiv:2501.11823
-
Towards Accurate Unified Anomaly Segmentation 21 Jan 2025 · 1 repository · arXiv:2501.12295
-
Vision-Language Models for Automated Chest X-ray Interpretation: Leveraging ViT and GPT-2 21 Jan 2025 · 0 repositories · arXiv:2501.12356
-
WaveNet-SF: A Hybrid Network for Retinal Disease Detection Based on Wavelet Transform in the Spatial-Frequency Domain 21 Jan 2025 · 0 repositories · arXiv:2501.11854
-
Adaptive parameters identification for nonlinear dynamics using deep permutation invariant networks 20 Jan 2025 · 0 repositories · arXiv:2501.11350
-
Advancing Oyster Phenotype Segmentation with Multi-Network Ensemble and Multi-Scale mechanism 20 Jan 2025 · 0 repositories · arXiv:2501.11203
-
Avoiding Shortcuts: Enhancing Channel-Robust Specific Emitter Identification via Single-Source Domain Generalization 20 Jan 2025 · 2 repositories
-
CatV2TON: Taming Diffusion Transformers for Vision-Based Virtual Try-On with Temporal Concatenation 20 Jan 2025 · 1 repository · arXiv:2501.11325
-
DLinear-based Prediction of Remaining Useful Life of Lithium-Ion Batteries: Feature Engineering through Explainable Artificial Intelligence 20 Jan 2025 · 0 repositories · arXiv:2501.11542
-
Early evidence of how LLMs outperform traditional systems on OCR/HTR tasks for historical records 20 Jan 2025 · 1 repository · arXiv:2501.11623
-
Episodic memory in AI agents poses risks that should be studied and mitigated 20 Jan 2025 · 0 repositories · arXiv:2501.11739
-
Explainable Lane Change Prediction for Near-Crash Scenarios Using Knowledge Graph Embeddings and Retrieval Augmented Generation 20 Jan 2025 · 0 repositories · arXiv:2501.11560
-
Generative AI-enabled Blockage Prediction for Robust Dual-Band mmWave Communication 20 Jan 2025 · 0 repositories · arXiv:2501.11763
-
Glinthawk: A Two-Tiered Architecture for Offline LLM Inference 20 Jan 2025 · 1 repository · arXiv:2501.11779
-
KEIR @ ECIR 2025: The Second Workshop on Knowledge-Enhanced Information Retrieval 20 Jan 2025 · 0 repositories · arXiv:2501.11499
-
Leveraging graph neural networks and mobility data for COVID-19 forecasting 20 Jan 2025 · 1 repository · arXiv:2501.11711
-
Mitigating Spatial Disparity in Urban Prediction Using Residual-Aware Spatiotemporal Graph Neural Networks: A Chicago Case Study 20 Jan 2025 · 0 repositories · arXiv:2501.11214
-
Multivariate Wireless Link Quality Prediction Based on Pre-trained Large Language Models 20 Jan 2025 · 0 repositories · arXiv:2501.11247
-
Neural Contextual Reinforcement Framework for Logical Structure Language Generation 20 Jan 2025 · 0 repositories · arXiv:2501.11417
-
PIKE-RAG: sPecIalized KnowledgE and Rationale Augmented Generation 20 Jan 2025 · 1 repository · arXiv:2501.11551Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Poison-RAG: Adversarial Data Poisoning Attacks on Retrieval-Augmented Generation in Recommender Systems 20 Jan 2025 · 1 repository · arXiv:2501.11759
-
Practical Modulo Sampling: Mitigating High-Frequency Components 20 Jan 2025 · 0 repositories · arXiv:2501.11330
-
PSO-based Sliding Mode Current Control of Grid-Forming Inverter in Rotating Frame 20 Jan 2025 · 0 repositories · arXiv:2501.11633
-
SEF-PNet: Speaker Encoder-Free Personalized Speech Enhancement with Local and Global Contexts Aggregation 20 Jan 2025 · 1 repository · arXiv:2501.11274
-
Sparse L0-norm based Kernel-free Quadratic Surface Support Vector Machines 20 Jan 2025 · 1 repository · arXiv:2501.11268
-
StyleSSP: Sampling StartPoint Enhancement for Training-free Diffusion-based Method for Style Transfer 20 Jan 2025 · 0 repositories · arXiv:2501.11319Syntology 7 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 7 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection 20 Jan 2025 · 0 repositories · arXiv:2501.11786
-
The Dual-use Dilemma in LLMs: Do Empowering Ethical Capacities Make a Degraded Utility? 20 Jan 2025 · 0 repositories · arXiv:2501.13952
-
Training-free Ultra Small Model for Universal Sparse Reconstruction in Compressed Sensing 20 Jan 2025 · 1 repository · arXiv:2501.11592
-
Trustformer: A Trusted Federated Transformer 20 Jan 2025 · 0 repositories · arXiv:2501.11706
-
TutorLLM: Customizing Learning Recommendations with Knowledge Tracing and Retrieval-Augmented Generation 20 Jan 2025 · 0 repositories · arXiv:2502.15709
-
Advancing General Multimodal Capability of Vision-language Models with Pyramid-descent Visual Position Encoding 19 Jan 2025 · 1 repository · arXiv:2501.10967
-
Chain-of-Reasoning: Towards Unified Mathematical Reasoning in Large Language Models via a Multi-Paradigm Perspective 19 Jan 2025 · 0 repositories · arXiv:2501.11110
-
Community Detection for Contextual-LSBM: Theoretical Limitations of Misclassification Rate and Efficient Algorithms 19 Jan 2025 · 0 repositories · arXiv:2501.11139
-
Dagger Behind Smile: Fool LLMs with a Happy Ending Story 19 Jan 2025 · 0 repositories · arXiv:2501.13115
-
DeepIFSAC: Deep Imputation of Missing Values Using Feature and Sample Attention within Contrastive Framework 19 Jan 2025 · 1 repository · arXiv:2501.10910
-
Design and Prototyping of Filtering Active STAR-RIS with Adjustable Power Splitting 19 Jan 2025 · 0 repositories · arXiv:2501.11062
-
Enhanced Suicidal Ideation Detection from Social Media Using a CNN-BiLSTM Hybrid Model 19 Jan 2025 · 0 repositories · arXiv:2501.11094
-
Enhancing Brain Tumor Segmentation Using Channel Attention and Transfer learning 19 Jan 2025 · 1 repository · arXiv:2501.11196
-
Enhancing Semantic Consistency of Large Language Models through Model Editing: An Interpretability-Oriented Approach 19 Jan 2025 · 0 repositories · arXiv:2501.11041
-
Few-shot Human Motion Recognition through Multi-Aspect mmWave FMCW Radar Data 19 Jan 2025 · 0 repositories · arXiv:2501.11028
-
From Arabic Text to Puzzles: LLM-Driven Development of Arabic Educational Crosswords 19 Jan 2025 · 0 repositories · arXiv:2501.11035
-
Generative Retrieval for Book search 19 Jan 2025 · 0 repositories · arXiv:2501.11034
-
HFGCN:Hypergraph Fusion Graph Convolutional Networks for Skeleton-Based Action Recognition 19 Jan 2025 · 0 repositories · arXiv:2501.11007
-
Leveraging counterfactual concepts for debugging and improving CNN model performance 19 Jan 2025 · 0 repositories · arXiv:2501.11087
-
LF-Steering: Latent Feature Activation Steering for Enhancing Semantic Consistency in Large Language Models 19 Jan 2025 · 0 repositories · arXiv:2501.11036
-
Modeling Attention during Dimensional Shifts with Counterfactual and Delayed Feedback 19 Jan 2025 · 1 repository · arXiv:2501.11161
-
Playing the Lottery With Concave Regularizers for Sparse Trainable Neural Networks 19 Jan 2025 · 0 repositories · arXiv:2501.11135
-
ProKeR: A Kernel Perspective on Few-Shot Adaptation of Large Vision-Language Models 19 Jan 2025 · 0 repositories · arXiv:2501.11175
-
A CNN-Transformer for Classification of Longitudinal 3D MRI Images -- A Case Study on Hepatocellular Carcinoma Prediction 18 Jan 2025 · 1 repository · arXiv:2501.10733
-
Building Short Value Chains for Animal Welfare-Friendly Products Adoption: Insights from a Restaurant-Based Study in Japan 18 Jan 2025 · 0 repositories · arXiv:2501.10680
-
CEReBrO: Compact Encoder for Representations of Brain Oscillations Using Efficient Alternating Attention 18 Jan 2025 · 0 repositories · arXiv:2501.10885
-
CS-Net:Contribution-based Sampling Network for Point Cloud Simplification 18 Jan 2025 · 0 repositories · arXiv:2501.10789
-
Dynamic Trend Fusion Module for Traffic Flow Prediction 18 Jan 2025 · 1 repository · arXiv:2501.10796
-
Efficient Auto-Labeling of Large-Scale Poultry Datasets (ALPD) Using Semi-Supervised Models, Active Learning, and Prompt-then-Detect Approach 18 Jan 2025 · 0 repositories · arXiv:2501.10809
-
FSMoE: A Flexible and Scalable Training System for Sparse Mixture-of-Experts Models 18 Jan 2025 · 0 repositories · arXiv:2501.10714
-
GEC-RAG: Improving Generative Error Correction via Retrieval-Augmented Generation for Automatic Speech Recognition Systems 18 Jan 2025 · 0 repositories · arXiv:2501.10734
-
HOPS: High-order Polynomials with Self-supervised Dimension Reduction for Load Forecasting 18 Jan 2025 · 0 repositories · arXiv:2501.10637