Methods › General › Attention Mechanisms › Attention › Papers, page 104
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 104 of 316: papers 10,301 to 10,400 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
GSTAM: Efficient Graph Distillation with Structural Attention-Matching 29 Aug 2024 · 1 repository · arXiv:2408.16871
-
H-SGANet: Hybrid Sparse Graph Attention Network for Deformable Medical Image Registration 29 Aug 2024 · 0 repositories · arXiv:2408.16719
-
HyPA-RAG: A Hybrid Parameter Adaptive Retrieval-Augmented Generation System for AI Legal and Policy Applications 29 Aug 2024 · 0 repositories · arXiv:2409.09046
-
Integrating Features for Recognizing Human Activities through Optimized Parameters in Graph Convolutional Networks and Transformer Architectures 29 Aug 2024 · 0 repositories · arXiv:2408.16442
-
Jina-ColBERT-v2: A General-Purpose Multilingual Late Interaction Retriever 29 Aug 2024 · 0 repositories · arXiv:2408.16672
-
LLaVA-Chef: A Multi-modal Generative Model for Food Recipes 29 Aug 2024 · 1 repository · arXiv:2408.16889
-
LLaVA-SG: Leveraging Scene Graphs as Visual Semantic Expression in Vision-Language Models 29 Aug 2024 · 0 repositories · arXiv:2408.16224
-
Locally Grouped and Scale-Guided Attention for Dense Pest Counting 29 Aug 2024 · 0 repositories · arXiv:2408.16503
-
Machine learning models for daily rainfall forecasting in Northern Tropical Africa using tropical wave predictors 29 Aug 2024 · 1 repository · arXiv:2408.16349
-
On-device AI: Quantization-aware Training of Transformers in Time-Series 29 Aug 2024 · 0 repositories · arXiv:2408.16495
-
One-Shot Learning Meets Depth Diffusion in Multi-Object Videos 29 Aug 2024 · 0 repositories · arXiv:2408.16704
-
PartFormer: Awakening Latent Diverse Representation from Vision Transformer for Object Re-Identification 29 Aug 2024 · 0 repositories · arXiv:2408.16684
-
PolarBEVDet: Exploring Polar Representation for Multi-View 3D Object Detection in Bird's-Eye-View 29 Aug 2024 · 1 repository · arXiv:2408.16200
-
Prediction-Feedback DETR for Temporal Action Detection 29 Aug 2024 · 0 repositories · arXiv:2408.16729
-
PrivacyLens: Evaluating Privacy Norm Awareness of Language Models in Action 29 Aug 2024 · 1 repository · arXiv:2409.00138Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples)
-
SODAWideNet++: Combining Attention and Convolutions for Salient Object Detection 29 Aug 2024 · 1 repository · arXiv:2408.16645
-
Spiking Diffusion Models 29 Aug 2024 · 1 repository · arXiv:2408.16467
-
Super-Resolution works for coastal simulations 29 Aug 2024 · 0 repositories · arXiv:2408.16553
-
TempoKGAT: A Novel Graph Attention Network Approach for Temporal Graph Analysis 29 Aug 2024 · 1 repository · arXiv:2408.16391
-
Tiny-Toxic-Detector: A compact transformer-based model for toxic content detection 29 Aug 2024 · 0 repositories · arXiv:2409.02114
-
Toward Robust Early Detection of Alzheimer's Disease via an Integrated Multimodal Learning Approach 29 Aug 2024 · 1 repository · arXiv:2408.16343
-
Transformers Meet ACT-R: Repeat-Aware and Sequential Listening Session Recommendation 29 Aug 2024 · 1 repository · arXiv:2408.16578
-
UDD: Dataset Distillation via Mining Underutilized Regions 29 Aug 2024 · 0 repositories · arXiv:2408.16268
-
Uni-3DAD: GAN-Inversion Aided Universal 3D Anomaly Detection on Model-free Products 29 Aug 2024 · 0 repositories · arXiv:2408.16201
-
VideoLLM-MoD: Efficient Video-Language Streaming with Mixture-of-Depths Vision Computation 29 Aug 2024 · 0 repositories · arXiv:2408.16730
-
WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling 29 Aug 2024 · 1 repository · arXiv:2408.16532Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Weakly Supervised Object Detection for Automatic Tooth-marked Tongue Recognition 29 Aug 2024 · 1 repository · arXiv:2408.16451
-
Deep Learning-Based Automatic Multi-Level Airway Collapse Monitoring on Obstructive Sleep Apnea Patients 28 Aug 2024 · 0 repositories · arXiv:2408.16030
-
A Simple Baseline with Single-encoder for Referring Image Segmentation 28 Aug 2024 · 0 repositories · arXiv:2408.15521
-
AeroVerse: UAV-Agent Benchmark Suite for Simulating, Pre-training, Finetuning, and Evaluating Aerospace Embodied World Models 28 Aug 2024 · 0 repositories · arXiv:2408.15511
-
An Extremely Data-efficient and Generative LLM-based Reinforcement Learning Agent for Recommenders 28 Aug 2024 · 0 repositories · arXiv:2408.16032
-
ANVIL: Anomaly-based Vulnerability Identification without Labelled Training Data 28 Aug 2024 · 0 repositories · arXiv:2408.16028
-
CBF-LLM: Safe Control for LLM Alignment 28 Aug 2024 · 1 repository · arXiv:2408.15625
-
Conan-embedding: General Text Embedding with More and Better Negative Samples 28 Aug 2024 · 0 repositories · arXiv:2408.15710
-
Enhancing and Accelerating Large Language Models via Instruction-Aware Contextual Compression 28 Aug 2024 · 1 repository · arXiv:2408.15491
-
Evaluating Computational Representations of Character: An Austen Character Similarity Benchmark 28 Aug 2024 · 0 repositories · arXiv:2408.16131
-
Evaluating Named Entity Recognition Using Few-Shot Prompting with Large Language Models 28 Aug 2024 · 1 repository · arXiv:2408.15796
-
FRACTURED-SORRY-Bench: Framework for Revealing Attacks in Conversational Turns Undermining Refusal Efficacy and Defenses over SORRY-Bench (Automated Multi-shot Jailbreaks) 28 Aug 2024 · 0 repositories · arXiv:2408.16163
-
Grand canonical generative diffusion model for crystalline phases and grain boundaries 28 Aug 2024 · 0 repositories · arXiv:2408.15601
-
In-Context Imitation Learning via Next-Token Prediction 28 Aug 2024 · 1 repository · arXiv:2408.15980Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 2 pointer-only (licence)
-
Interactive Agents: Simulating Counselor-Client Psychological Counseling via Role-Playing LLM-to-LLM Interactions 28 Aug 2024 · 1 repository · arXiv:2408.15787Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Is Personality Prediction Possible Based on Reddit Comments? 28 Aug 2024 · 0 repositories · arXiv:2408.16089
-
Leveraging Large Language Models for Wireless Symbol Detection via In-Context Learning 28 Aug 2024 · 0 repositories · arXiv:2409.00124
-
Towards Logically Sound Natural Language Reasoning with Logic-Enhanced Language Model Agents 28 Aug 2024 · 1 repository · arXiv:2408.16081
-
LRP4RAG: Detecting Hallucinations in Retrieval-Augmented Generation via Layer-wise Relevance Propagation 28 Aug 2024 · 3 repositories · arXiv:2408.15533
-
MambaPlace:Text-to-Point-Cloud Cross-Modal Place Recognition with Attention Mamba Mechanisms 28 Aug 2024 · 1 repository · arXiv:2408.15740
-
Multi-modal Adversarial Training for Zero-Shot Voice Cloning 28 Aug 2024 · 0 repositories · arXiv:2408.15916
-
NAS-BNN: Neural Architecture Search for Binary Neural Networks 28 Aug 2024 · 1 repository · arXiv:2408.15484
-
Object Detection for Vehicle Dashcams using Transformers 28 Aug 2024 · 0 repositories · arXiv:2408.15809
-
PDSR: A Privacy-Preserving Diversified Service Recommendation Method on Distributed Data 28 Aug 2024 · 0 repositories · arXiv:2408.15688
-
RIDE: Boosting 3D Object Detection for LiDAR Point Clouds via Rotation-Invariant Analysis 28 Aug 2024 · 0 repositories · arXiv:2408.15643
-
Scaling Up Summarization: Leveraging Large Language Models for Long Text Extractive Summarization 28 Aug 2024 · 0 repositories · arXiv:2408.15801
-
SITransformer: Shared Information-Guided Transformer for Extreme Multimodal Summarization 28 Aug 2024 · 1 repository · arXiv:2408.15829
-
Temporal Attention for Cross-View Sequential Image Localization 28 Aug 2024 · 1 repository · arXiv:2408.15569
-
Toward Automated Simulation Research Workflow through LLM Prompt Engineering Design 28 Aug 2024 · 1 repository · arXiv:2408.15512Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Unleashing the Temporal-Spatial Reasoning Capacity of GPT for Training-Free Audio and Language Referenced Video Object Segmentation 28 Aug 2024 · 1 repository · arXiv:2408.15876
-
Visual Prompt Engineering for Medical Vision Language Models in Radiology 28 Aug 2024 · 0 repositories · arXiv:2408.15802
-
WebPilot: A Versatile and Autonomous Multi-Agent System for Web Task Execution with Strategic Exploration 28 Aug 2024 · 0 repositories · arXiv:2408.15978
-
A High Altitude Platform-Based 3D Geometrical Channel Model for Beamforming Characterization in Future 6G Flying Ad-Hoc Networks 27 Aug 2024 · 0 repositories · arXiv:2408.14986
-
A Survey of Large Language Models for European Languages 27 Aug 2024 · 0 repositories · arXiv:2408.15040
-
Alfie: Democratising RGBA Image Generation With No $$$ 27 Aug 2024 · 2 repositories · arXiv:2408.14826
-
Applying ViT in Generalized Few-shot Semantic Segmentation 27 Aug 2024 · 1 repository · arXiv:2408.14957
-
Automatic 8-tissue Segmentation for 6-month Infant Brains 27 Aug 2024 · 0 repositories · arXiv:2408.15198
-
CrossViewDiff: A Cross-View Diffusion Model for Satellite-to-Street View Synthesis 27 Aug 2024 · 0 repositories · arXiv:2408.14765
-
DualKanbaFormer: An Efficient Selective Sparse Framework for Multimodal Aspect-based Sentiment Analysis 27 Aug 2024 · 0 repositories · arXiv:2408.15379
-
Enhancing License Plate Super-Resolution: A Layout-Aware and Character-Driven Approach 27 Aug 2024 · 1 repository · arXiv:2408.15103
-
FastTextSpotter: A High-Efficiency Transformer for Multilingual Scene Text Spotting 27 Aug 2024 · 1 repository · arXiv:2408.14998
-
Graph Attention Inference of Network Topology in Multi-Agent Systems 27 Aug 2024 · 0 repositories · arXiv:2408.15449
-
GSIFN: A Graph-Structured and Interlaced-Masked Multimodal Transformer-based Fusion Network for Multimodal Sentiment Analysis 27 Aug 2024 · 1 repository · arXiv:2408.14809
-
Hierarchical Graph Interaction Transformer with Dynamic Token Clustering for Camouflaged Object Detection 27 Aug 2024 · 1 repository · arXiv:2408.15020
-
How transformers learn structured data: insights from hierarchical filtering 27 Aug 2024 · 1 repository · arXiv:2408.15138Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
HPT++: Hierarchically Prompting Vision-Language Models with Multi-Granularity Knowledge Generation and Improved Structure Modeling 27 Aug 2024 · 2 repositories · arXiv:2408.14812
-
Into the Unknown Unknowns: Engaged Human Learning through Participation in Language Model Agent Conversations 27 Aug 2024 · 1 repository · arXiv:2408.15232
-
Large Language Models for Disease Diagnosis: A Scoping Review 27 Aug 2024 · 0 repositories · arXiv:2409.00097
-
MeshUp: Multi-Target Mesh Deformation via Blended Score Distillation 27 Aug 2024 · 0 repositories · arXiv:2408.14899
-
MMASD+: A Novel Dataset for Privacy-Preserving Behavior Analysis of Children with Autism Spectrum Disorder 27 Aug 2024 · 1 repository · arXiv:2408.15077
-
MROVSeg: Breaking the Resolution Curse of Vision-Language Models in Open-Vocabulary Image Segmentation 27 Aug 2024 · 0 repositories · arXiv:2408.14776
-
Multi-Modal Instruction-Tuning Small-Scale Language-and-Vision Assistant for Semiconductor Electron Micrograph Analysis 27 Aug 2024 · 0 repositories · arXiv:2409.07463
-
Multitone PSK Modulation Design for Simultaneous Wireless Information and Power Transfer 27 Aug 2024 · 0 repositories · arXiv:2408.15083
-
Negation Blindness in Large Language Models: Unveiling the NO Syndrome in Image Generation 27 Aug 2024 · 0 repositories · arXiv:2409.00105
-
No Regrets: Investigating and Improving Regret Approximations for Curriculum Discovery 27 Aug 2024 · 1 repository · arXiv:2408.15099Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Parameter-Efficient Quantized Mixture-of-Experts Meets Vision-Language Instruction Tuning for Semiconductor Electron Micrograph Analysis 27 Aug 2024 · 0 repositories · arXiv:2408.15305
-
PAT: Pruning-Aware Tuning for Large Language Models 27 Aug 2024 · 2 repositories · arXiv:2408.14721
-
Personalized Video Summarization using Text-Based Queries and Conditional Modeling 27 Aug 2024 · 0 repositories · arXiv:2408.14743
-
RGDA-DDI: Residual graph attention network and dual-attention based framework for drug-drug interaction prediction 27 Aug 2024 · 1 repository · arXiv:2408.15310
-
S-MolSearch: 3D Semi-supervised Contrastive Learning for Bioactive Molecule Search 27 Aug 2024 · 0 repositories · arXiv:2409.07462
-
SpikingSSMs: Learning Long Sequences with Sparse and Parallel Spiking State Space Models 27 Aug 2024 · 1 repository · arXiv:2408.14909
-
Strategic Optimization and Challenges of Large Language Models in Object-Oriented Programming 27 Aug 2024 · 0 repositories · arXiv:2408.14834
-
TCNFormer: Temporal Convolutional Network Former for Short-Term Wind Speed Forecasting 27 Aug 2024 · 0 repositories · arXiv:2408.15737
-
The Benefits of Balance: From Information Projections to Variance Reduction 27 Aug 2024 · 0 repositories · arXiv:2408.15065
-
The Mamba in the Llama: Distilling and Accelerating Hybrid Models 27 Aug 2024 · 2 repositories · arXiv:2408.15237Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 14 harvested samples) · 5 pointer-only (licence)
-
The Uniqueness of LLaMA3-70B Series with Per-Channel Quantization 27 Aug 2024 · 0 repositories · arXiv:2408.15301
-
TourSynbio: A Multi-Modal Large Model and Agent Framework to Bridge Text and Protein Sequences for Protein Engineering 27 Aug 2024 · 1 repository · arXiv:2408.15299
-
A Permuted Autoregressive Approach to Word-Level Recognition for Urdu Digital Text 27 Aug 2024 · 0 repositories · arXiv:2408.15119
-
VHAKG: A Multi-modal Knowledge Graph Based on Synchronized Multi-view Videos of Daily Activities 27 Aug 2024 · 2 repositories · arXiv:2408.14895
-
Zero-Shot Visual Reasoning by Vision-Language Models: Benchmarking and Analysis 27 Aug 2024 · 0 repositories · arXiv:2409.00106
-
A Lightweight Insulator Defect Detection Model Based on Drone Images 26 Aug 2024 · 1 repository
-
A Multilateral Attention-enhanced Deep Neural Network for Disease Outbreak Forecasting: A Case Study on COVID-19 26 Aug 2024 · 0 repositories · arXiv:2408.14519
-
A novel k-generation propagation model for cyber risk and its application to cyber insurance 26 Aug 2024 · 0 repositories · arXiv:2408.14151
-
A Survey of Camouflaged Object Detection and Beyond 26 Aug 2024 · 1 repository · arXiv:2408.14562