Methods › General › Attention Mechanisms › Attention › Papers, page 75
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 75 of 316: papers 7,401 to 7,500 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
LES-Talker: Fine-Grained Emotion Editing for Talking Head Generation in Linear Emotion Space 14 Nov 2024 · 0 repositories · arXiv:2411.09268
-
Local deployment of large-scale music AI models on commodity hardware 14 Nov 2024 · 0 repositories · arXiv:2411.09625
-
Local-Global Attention: An Adaptive Mechanism for Multi-Scale Feature Integration 14 Nov 2024 · 1 repository · arXiv:2411.09604
-
MM-Eval: A Hierarchical Benchmark for Modern Mongolian Evaluation in LLMs 14 Nov 2024 · 1 repository · arXiv:2411.09492
-
On the Surprising Effectiveness of Attention Transfer for Vision Transformers 14 Nov 2024 · 1 repository · arXiv:2411.09702Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
OpenGeMM: A High-Utilization GeMM Accelerator Generator with Lightweight RISC-V Control and Tight Memory Coupling 14 Nov 2024 · 1 repository · arXiv:2411.09543
-
Partial Multi-View Clustering via Meta-Learning and Contrastive Feature Alignment 14 Nov 2024 · 0 repositories · arXiv:2411.09758
-
Re-Parameterization of Lightweight Transformer for On-Device Speech Emotion Recognition 14 Nov 2024 · 0 repositories · arXiv:2411.09339
-
Reducing Reasoning Costs: The Path of Optimization for Chain of Thought via Sparse Attention Mechanism 14 Nov 2024 · 1 repository · arXiv:2411.09111
-
SAG-ViT: A Scale-Aware, High-Fidelity Patching Approach with Graph Attention for Vision Transformers 14 Nov 2024 · 1 repository · arXiv:2411.09420
-
Squeezed Attention: Accelerating Long Context Length LLM Inference 14 Nov 2024 · 1 repository · arXiv:2411.09688Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
Stability and Generalization for Distributed SGDA 14 Nov 2024 · 0 repositories · arXiv:2411.09365
-
The Future of Skill: What Is It to Be Skilled at Work? 14 Nov 2024 · 0 repositories · arXiv:2411.10488
-
Towards a Classification of Open-Source ML Models and Datasets for Software Engineering 14 Nov 2024 · 0 repositories · arXiv:2411.09683
-
A Large-Scale Study of Relevance Assessments with Large Language Models: An Initial Look 13 Nov 2024 · 1 repository · arXiv:2411.08275
-
A Transformer-Based Visual Piano Transcription Algorithm 13 Nov 2024 · 0 repositories · arXiv:2411.09037
-
AD-DINO: Attention-Dynamic DINO for Distance-Aware Embodied Reference Understanding 13 Nov 2024 · 0 repositories · arXiv:2411.08451
-
Advanced Nonlinear SCMA Codebook Design Based on Lattice Constellations 13 Nov 2024 · 0 repositories · arXiv:2411.08493
-
Analyst Reports and Stock Performance: Evidence from the Chinese Market 13 Nov 2024 · 0 repositories · arXiv:2411.08726
-
CamemBERT 2.0: A Smarter French Language Model Aged to Perfection 13 Nov 2024 · 0 repositories · arXiv:2411.08868
-
Continuous GNN-based Anomaly Detection on Edge using Efficient Adaptive Knowledge Graph Learning 13 Nov 2024 · 0 repositories · arXiv:2411.09072
-
FinRobot: AI Agent for Equity Research and Valuation with Large Language Models 13 Nov 2024 · 1 repository · arXiv:2411.08804Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples)
-
Flow reconstruction in time-varying geometries using graph neural networks 13 Nov 2024 · 0 repositories · arXiv:2411.08764
-
Fluoroformer: Scaling multiple instance learning to multiplexed images via attention-based channel fusion 13 Nov 2024 · 1 repository · arXiv:2411.08975
-
LLMStinger: Jailbreaking LLMs using RL fine-tuned LLMs 13 Nov 2024 · 0 repositories · arXiv:2411.08862
-
LogLLM: Log-based Anomaly Detection Using Large Language Models 13 Nov 2024 · 1 repository · arXiv:2411.08561
-
Oblique Bayesian additive regression trees 13 Nov 2024 · 0 repositories · arXiv:2411.08849
-
PerceiverS: A Multi-Scale Perceiver with Effective Segmentation for Long-Term Expressive Symbolic Music Generation 13 Nov 2024 · 0 repositories · arXiv:2411.08307
-
Quantity versus Diversity: Influence of Data on Detecting EEG Pathology with Advanced ML Models 13 Nov 2024 · 0 repositories · arXiv:2411.17709
-
ReMP: Reusable Motion Prior for Multi-domain 3D Human Pose Estimation and Motion Inbetweening 13 Nov 2024 · 0 repositories · arXiv:2411.09435
-
RESOLVE: Relational Reasoning with Symbolic and Object-Level Features Using Vector Symbolic Processing 13 Nov 2024 · 1 repository · arXiv:2411.08290
-
Responsible AI in Construction Safety: Systematic Evaluation of Large Language Models and Prompt Engineering 13 Nov 2024 · 0 repositories · arXiv:2411.08320
-
Retrieval Augmented Recipe Generation 13 Nov 2024 · 0 repositories · arXiv:2411.08715
-
SAD-TIME: a Spatiotemporal-fused network for depression detection with Automated multi-scale Depth-wise and TIME-interval-related common feature extractor 13 Nov 2024 · 0 repositories · arXiv:2411.08521
-
SAM-I2I: Unleash the Power of Segment Anything Model for Medical Image Translation 13 Nov 2024 · 0 repositories · arXiv:2411.12755
-
SASE: A Searching Architecture for Squeeze and Excitation Operations 13 Nov 2024 · 0 repositories · arXiv:2411.08333
-
Scale Contrastive Learning with Selective Attentions for Blind Image Quality Assessment 13 Nov 2024 · 0 repositories · arXiv:2411.09007
-
Towards Objective and Unbiased Decision Assessments with LLM-Enhanced Hierarchical Attention Networks 13 Nov 2024 · 1 repository · arXiv:2411.08504
-
Towards Optimizing a Retrieval Augmented Generation using Large Language Model on Academic Data 13 Nov 2024 · 0 repositories · arXiv:2411.08438
-
TRACE: Transformer-based Risk Assessment for Clinical Evaluation 13 Nov 2024 · 1 repository · arXiv:2411.08701
-
UIFormer: A Unified Transformer-based Framework for Incremental Few-Shot Object Detection and Instance Segmentation 13 Nov 2024 · 0 repositories · arXiv:2411.08569
-
VALTEST: Automated Validation of Language Model Generated Test Cases 13 Nov 2024 · 0 repositories · arXiv:2411.08254
-
A Preview of XiYan-SQL: A Multi-Generator Ensemble Framework for Text-to-SQL 13 Nov 2024 · 5 repositories · arXiv:2411.08599
-
Breaking the Low-Rank Dilemma of Linear Attention 12 Nov 2024 · 1 repository · arXiv:2411.07635Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
BudgetMLAgent: A Cost-Effective LLM Multi-Agent system for Automating Machine Learning Tasks 12 Nov 2024 · 0 repositories · arXiv:2411.07464
-
Can adversarial attacks by large language models be attributed? 12 Nov 2024 · 0 repositories · arXiv:2411.08003
-
Circuit Complexity Bounds for RoPE-based Transformer Architecture 12 Nov 2024 · 0 repositories · arXiv:2411.07602
-
Contrastive Language Prompting to Ease False Positives in Medical Anomaly Detection 12 Nov 2024 · 1 repository · arXiv:2411.07546
-
Controlled Evaluation of Syntactic Knowledge in Multilingual Language Models 12 Nov 2024 · 1 repository · arXiv:2411.07474
-
Deceiving Question-Answering Models: A Hybrid Word-Level Adversarial Approach 12 Nov 2024 · 1 repository · arXiv:2411.08248
-
DINO-LG: A Task-Specific DINO Model for Coronary Calcium Scoring 12 Nov 2024 · 0 repositories · arXiv:2411.07976
-
Efficient Federated Finetuning of Tiny Transformers with Resource-Constrained Devices 12 Nov 2024 · 0 repositories · arXiv:2411.07826
-
Emotion Classification of Children Expressions 12 Nov 2024 · 0 repositories · arXiv:2411.07708
-
Enhancing Link Prediction with Fuzzy Graph Attention Networks and Dynamic Negative Sampling 12 Nov 2024 · 0 repositories · arXiv:2411.07482
-
Evaluating ChatGPT-3.5 Efficiency in Solving Coding Problems of Different Complexity Levels: An Empirical Analysis 12 Nov 2024 · 1 repository · arXiv:2411.07529
-
Fair Summarization: Bridging Quality and Diversity in Extractive Summaries 12 Nov 2024 · 1 repository · arXiv:2411.07521Syntology official (archive's flag): 3 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Fast Disentangled Slim Tensor Learning for Multi-view Clustering 12 Nov 2024 · 1 repository · arXiv:2411.07685
-
FM-TS: Flow Matching for Time Series Generation 12 Nov 2024 · 1 repository · arXiv:2411.07506Syntology official (archive's flag): 20 ran · 20 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 2 honoured, 3 violated, 11 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 20 harvested samples) · 20 pointer-only (licence)
-
HMIL: Hierarchical Multi-Instance Learning for Fine-Grained Whole Slide Image Classification 12 Nov 2024 · 1 repository · arXiv:2411.07660
-
Improving Grapheme-to-Phoneme Conversion through In-Context Knowledge Retrieval with Large Language Models 12 Nov 2024 · 0 repositories · arXiv:2411.07563
-
Interaction Asymmetry: A General Principle for Learning Composable Abstractions 12 Nov 2024 · 1 repository · arXiv:2411.07784Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Joint multi-dimensional dynamic attention and transformer for general image restoration 12 Nov 2024 · 1 repository · arXiv:2411.07893
-
Large Language Models Can Self-Improve in Long-context Reasoning 12 Nov 2024 · 1 repository · arXiv:2411.08147Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Leveraging Multimodal Models for Enhanced Neuroimaging Diagnostics in Alzheimer's Disease 12 Nov 2024 · 0 repositories · arXiv:2411.07871
-
LLM App Squatting and Cloning 12 Nov 2024 · 0 repositories · arXiv:2411.07518
-
Multi-task Feature Enhancement Network for No-Reference Image Quality Assessment 12 Nov 2024 · 0 repositories · arXiv:2411.07556
-
Multimodal Clinical Reasoning through Knowledge-augmented Rationale Generation 12 Nov 2024 · 0 repositories · arXiv:2411.07611
-
New Emerged Security and Privacy of Pre-trained Model: a Survey and Outlook 12 Nov 2024 · 0 repositories · arXiv:2411.07691
-
Query Optimization for Parametric Knowledge Refinement in Retrieval-Augmented Large Language Models 12 Nov 2024 · 0 repositories · arXiv:2411.07820
-
Rendering-Oriented 3D Point Cloud Attribute Compression using Sparse Tensor-based Transformer 12 Nov 2024 · 0 repositories · arXiv:2411.07899
-
Retrieval Augmented Time Series Forecasting 12 Nov 2024 · 1 repository · arXiv:2411.08249Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Spatially Regularized Graph Attention Autoencoder Framework for Detecting Rainfall Extremes 12 Nov 2024 · 0 repositories · arXiv:2411.07753
-
Trustful LLMs: Customizing and Grounding Text Generation with Knowledge Bases and Dual Decoders 12 Nov 2024 · 0 repositories · arXiv:2411.07870
-
Two-Layer Attention Optimization for Bimanual Coordination 12 Nov 2024 · 0 repositories · arXiv:2411.07470
-
Unraveling the Gradient Descent Dynamics of Transformers 12 Nov 2024 · 0 repositories · arXiv:2411.07538
-
Verbosity ≠ Veracity: Demystify Verbosity Compensation Behavior of Large Language Models 12 Nov 2024 · 1 repository · arXiv:2411.07858
-
World Models: The Safety Perspective 12 Nov 2024 · 0 repositories · arXiv:2411.07690
-
A Unified Multi-Task Learning Architecture for Hate Detection Leveraging User-Based Information 11 Nov 2024 · 0 repositories · arXiv:2411.06855
-
Add-it: Training-Free Object Insertion in Images With Pretrained Diffusion Models 11 Nov 2024 · 1 repository · arXiv:2411.07232
-
AEROMamba: An efficient architecture for audio super-resolution using generative adversarial networks and state space models 11 Nov 2024 · 1 repository · arXiv:2411.07364
-
Ambient AI Scribing Support: Comparing the Performance of Specialized AI Agentic Architecture to Leading Foundational Models 11 Nov 2024 · 0 repositories · arXiv:2411.06713
-
An Efficient Memory Module for Graph Few-Shot Class-Incremental Learning 11 Nov 2024 · 1 repository · arXiv:2411.06659
-
AssistRAG: Boosting the Potential of Large Language Models with an Intelligent Information Assistant 11 Nov 2024 · 1 repository · arXiv:2411.06805Syntology official (archive's flag): 10 ran · 10 ran (of which 1 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Autonomous Droplet Microfluidic Design Framework with Large Language Models 11 Nov 2024 · 1 repository · arXiv:2411.06691
-
Can KAN Work? Exploring the Potential of Kolmogorov-Arnold Networks in Computer Vision 11 Nov 2024 · 0 repositories · arXiv:2411.06727
-
Cancer-Answer: Empowering Cancer Care with Advanced Large Language Models 11 Nov 2024 · 0 repositories · arXiv:2411.06946
-
ConvMixFormer- A Resource-efficient Convolution Mixer for Transformer-based Dynamic Hand Gesture Recognition 11 Nov 2024 · 1 repository · arXiv:2411.07118
-
Data-Driven Analysis of AI in Medical Device Software in China: Deep Learning and General AI Trends Based on Regulatory Data 11 Nov 2024 · 0 repositories · arXiv:2411.07378
-
Evaluating Large Language Models on Financial Report Summarization: An Empirical Study 11 Nov 2024 · 0 repositories · arXiv:2411.06852
-
Explore the Reasoning Capability of LLMs in the Chess Testbed 11 Nov 2024 · 0 repositories · arXiv:2411.06655
-
Fast and Robust Contextual Node Representation Learning over Dynamic Graphs 11 Nov 2024 · 0 repositories · arXiv:2411.07123
-
GTA-Net: An IoT-Integrated 3D Human Pose Estimation System for Real-Time Adolescent Sports Posture Correction 11 Nov 2024 · 0 repositories · arXiv:2411.06725
-
SynCL: A Synergistic Training Strategy with Instance-Aware Contrastive Learning for End-to-End Multi-Camera 3D Tracking 11 Nov 2024 · 0 repositories · arXiv:2411.06780
-
Invar-RAG: Invariant LLM-aligned Retrieval for Better Generation 11 Nov 2024 · 0 repositories · arXiv:2411.07021
-
Isochrony-Controlled Speech-to-Text Translation: A study on translating from Sino-Tibetan to Indo-European Languages 11 Nov 2024 · 0 repositories · arXiv:2411.07387
-
LA4SR: illuminating the dark proteome with generative AI 11 Nov 2024 · 0 repositories · arXiv:2411.06798
-
Layout Control and Semantic Guidance with Attention Loss Backward for T2I Diffusion Model 11 Nov 2024 · 0 repositories · arXiv:2411.06692
-
LongSafetyBench: Long-Context LLMs Struggle with Safety Issues 11 Nov 2024 · 1 repository · arXiv:2411.06899Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
MapSAM: Adapting Segment Anything Model for Automated Feature Detection in Historical Maps 11 Nov 2024 · 1 repository · arXiv:2411.06971
-
Modeling variable guide efficiency in pooled CRISPR screens with ContrastiveVI+ 11 Nov 2024 · 0 repositories · arXiv:2411.08072