Methods › General › Attention Mechanisms › Attention › Papers, page 124
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 124 of 316: papers 12,301 to 12,400 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
TheoremLlama: Transforming General-Purpose LLMs into Lean4 Experts 3 Jul 2024 · 1 repository · arXiv:2407.03203Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
Visual Grounding with Attention-Driven Constraint Balancing 3 Jul 2024 · 0 repositories · arXiv:2407.03243
-
When big data actually are low-rank, or entrywise approximation of certain function-generated matrices 3 Jul 2024 · 1 repository · arXiv:2407.03250
-
A Contrastive Learning Based Convolutional Neural Network for ERP Brain-Computer Interfaces 2 Jul 2024 · 0 repositories · arXiv:2407.04738
-
A Depression Detection Method Based on Multi-Modal Feature Fusion Using Cross-Attention 2 Jul 2024 · 0 repositories · arXiv:2407.12825
-
A Practical Review of Mechanistic Interpretability for Transformer-Based Language Models 2 Jul 2024 · 1 repository · arXiv:2407.02646
-
A Survey on Advancements in THz Technology for 6G: Systems, Circuits, Antennas, and Experiments 2 Jul 2024 · 0 repositories · arXiv:2407.01957
-
Adaptive Modality Balanced Online Knowledge Distillation for Brain-Eye-Computer based Dim Object Detection 2 Jul 2024 · 1 repository · arXiv:2407.01894
-
Assessing the Code Clone Detection Capability of Large Language Models 2 Jul 2024 · 0 repositories · arXiv:2407.02402
-
AXIAL: Attention-based eXplainability for Interpretable Alzheimer's Localized Diagnosis using 2D CNNs on 3D MRI brain scans 2 Jul 2024 · 1 repository · arXiv:2407.02418
-
BeNeRF: Neural Radiance Fields from a Single Blurry Image and Event Stream 2 Jul 2024 · 1 repository · arXiv:2407.02174Syntology official (archive's flag): 5 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Beyond Numeric Awards: In-Context Dueling Bandits with LLM Agents 2 Jul 2024 · 0 repositories · arXiv:2407.01887
-
CatMemo at the FinLLM Challenge Task: Fine-Tuning Large Language Models using Data Fusion in Financial Applications 2 Jul 2024 · 0 repositories · arXiv:2407.01953
-
Characterizing the Interpretability of Attention Maps in Digital Pathology 2 Jul 2024 · 0 repositories · arXiv:2407.02484
-
Classification of Power Quality Disturbances Using Resnet with Channel Attention Mechanism 2 Jul 2024 · 0 repositories · arXiv:2407.04739
-
CountFormer: Multi-View Crowd Counting Transformer 2 Jul 2024 · 1 repository · arXiv:2407.02047
-
Deep Learning Based Apparent Diffusion Coefficient Map Generation from Multi-parametric MR Images for Patients with Diffuse Gliomas 2 Jul 2024 · 0 repositories · arXiv:2407.02616
-
Diffusion Models for Tabular Data Imputation and Synthetic Data Generation 2 Jul 2024 · 0 repositories · arXiv:2407.02549
-
DM3D: Distortion-Minimized Weight Pruning for Lossless 3D Object Detection 2 Jul 2024 · 0 repositories · arXiv:2407.02098
-
Efficient Sparse Attention needs Adaptive Token Release 2 Jul 2024 · 0 repositories · arXiv:2407.02328
-
Ensemble of pre-trained language models and data augmentation for hate speech detection from Arabic tweets 2 Jul 2024 · 0 repositories · arXiv:2407.02448
-
Extracting and Encoding: Leveraging Large Language Models and Medical Knowledge to Enhance Radiological Text Representation 2 Jul 2024 · 1 repository · arXiv:2407.01948
-
Fake News Detection and Manipulation Reasoning via Large Vision-Language Models 2 Jul 2024 · 0 repositories · arXiv:2407.02042
-
FreeCG: Free the Design Space of Clebsch-Gordan Transform for Machine Learning Force Fields 2 Jul 2024 · 0 repositories · arXiv:2407.02263
-
GlyphDraw2: Automatic Generation of Complex Glyph Posters with Diffusion Models and Large Language Models 2 Jul 2024 · 1 repository · arXiv:2407.02252
-
GPTCast: a weather language model for precipitation nowcasting 2 Jul 2024 · 1 repository · arXiv:2407.02089
-
GRASP: A Grid-Based Benchmark for Evaluating Commonsense Spatial Reasoning 2 Jul 2024 · 0 repositories · arXiv:2407.01892
-
GVDIFF: Grounded Text-to-Video Generation with Diffusion Models 2 Jul 2024 · 0 repositories · arXiv:2407.01921
-
HC-GLAD: Dual Hyperbolic Contrastive Learning for Unsupervised Graph-Level Anomaly Detection 2 Jul 2024 · 1 repository · arXiv:2407.02057
-
HRSAM: Efficient Interactive Segmentation in High-Resolution Images 2 Jul 2024 · 1 repository · arXiv:2407.02109Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Improving Visual Storytelling with Multimodal Large Language Models 2 Jul 2024 · 0 repositories · arXiv:2407.02586
-
Integrate the Essence and Eliminate the Dross: Fine-Grained Self-Consistency for Free-Form Language Generation 2 Jul 2024 · 1 repository · arXiv:2407.02056Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Investigating Event-Based Cameras for Video Frame Interpolation in Sports 2 Jul 2024 · 0 repositories · arXiv:2407.02370
-
Joint-Dataset Learning and Cross-Consistent Regularization for Text-to-Motion Retrieval 2 Jul 2024 · 0 repositories · arXiv:2407.02104
-
Learning to Refine with Fine-Grained Natural Language Feedback 2 Jul 2024 · 1 repository · arXiv:2407.02397Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
LLM-Select: Feature Selection with Large Language Models 2 Jul 2024 · 0 repositories · arXiv:2407.02694
-
MeMemo: On-device Retrieval Augmentation for Private and Personalized Text Generation 2 Jul 2024 · 1 repository · arXiv:2407.01972
-
MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention 2 Jul 2024 · 2 repositories · arXiv:2407.02490Syntology official (archive's flag): 2 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples)
-
Language Model Alignment in Multilingual Trolley Problems 2 Jul 2024 · 2 repositories · arXiv:2407.02273Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Neurocache: Efficient Vector Retrieval for Long-range Language Modeling 2 Jul 2024 · 1 repository · arXiv:2407.02486
-
Occlusion-Aware Seamless Segmentation 2 Jul 2024 · 1 repository · arXiv:2407.02182Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
On the Anatomy of Attention 2 Jul 2024 · 0 repositories · arXiv:2407.02423
-
Open foundation models for Azerbaijani language 2 Jul 2024 · 0 repositories · arXiv:2407.02337
-
OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation 2 Jul 2024 · 0 repositories · arXiv:2407.02371
-
Pinyin Regularization in Error Correction for Chinese Speech Recognition with Large Language Models 2 Jul 2024 · 1 repository · arXiv:2407.01909
-
Predicting Visual Attention in Graphic Design Documents 2 Jul 2024 · 0 repositories · arXiv:2407.02439
-
RankRAG: Unifying Context Ranking with Retrieval-Augmented Generation in LLMs 2 Jul 2024 · 0 repositories · arXiv:2407.02485
-
Referring Atomic Video Action Recognition 2 Jul 2024 · 1 repository · arXiv:2407.01872
-
Research on target detection method of distracted driving behavior based on improved YOLOv8 2 Jul 2024 · 0 repositories · arXiv:2407.01864
-
SafaRi:Adaptive Sequence Transformer for Weakly Supervised Referring Expression Segmentation 2 Jul 2024 · 0 repositories · arXiv:2407.02389
-
SOAF: Scene Occlusion-aware Neural Acoustic Field 2 Jul 2024 · 0 repositories · arXiv:2407.02264
-
SeqAR: Jailbreak LLMs with Sequential Auto-Generated Characters 2 Jul 2024 · 1 repository · arXiv:2407.01902Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
The Art of Saying No: Contextual Noncompliance in Language Models 2 Jul 2024 · 1 repository · arXiv:2407.12043Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
The Solution for The PST-KDD-2024 OAG-Challenge 2 Jul 2024 · 0 repositories · arXiv:2407.12827
-
TrAME: Trajectory-Anchored Multi-View Editing for Text-Guided 3D Gaussian Splatting Manipulation 2 Jul 2024 · 0 repositories · arXiv:2407.02034
-
Unsupervised Face-Masked Speech Enhancement Using Generative Adversarial Networks With Human-in-the-Loop Assessment Metrics 2 Jul 2024 · 0 repositories · arXiv:2407.01939
-
What We Talk About When We Talk About LMs: Implicit Paradigm Shifts and the Ship of Language Models 2 Jul 2024 · 1 repository · arXiv:2407.01929
-
White-Box 3D-OMP-Transformer for ISAC 2 Jul 2024 · 0 repositories · arXiv:2407.02251
-
Why do LLaVA Vision-Language Models Reply to Images in English? 2 Jul 2024 · 0 repositories · arXiv:2407.02333
-
Zero-Shot Video Restoration and Enhancement Using Pre-Trained Image Diffusion Model 2 Jul 2024 · 0 repositories · arXiv:2407.01960
-
Embedded Prompt Tuning: Towards Enhanced Calibration of Pretrained Models for Medical Images 1 Jul 2024 · 1 repository · arXiv:2407.01003
-
A Global-Local Attention Mechanism for Relation Classification 1 Jul 2024 · 0 repositories · arXiv:2407.01424
-
AI Agents That Matter 1 Jul 2024 · 1 repository · arXiv:2407.01502Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Audio-Visual Approach For Multimodal Concurrent Speaker Detection 1 Jul 2024 · 0 repositories · arXiv:2407.01774
-
BERGEN: A Benchmarking Library for Retrieval-Augmented Generation 1 Jul 2024 · 1 repository · arXiv:2407.01102Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
CLEME2.0: Towards More Interpretable Evaluation by Disentangling Edits for Grammatical Error Correction 1 Jul 2024 · 1 repository · arXiv:2407.00934
-
CPT: Consistent Proxy Tuning for Black-box Optimization 1 Jul 2024 · 1 repository · arXiv:2407.01155
-
Cross-Modal Attention Alignment Network with Auxiliary Text Description for zero-shot sketch-based image retrieval 1 Jul 2024 · 0 repositories · arXiv:2407.00979
-
Cross-Slice Attention and Evidential Critical Loss for Uncertainty-Aware Prostate Cancer Detection 1 Jul 2024 · 1 repository · arXiv:2407.01146
-
CSFNet: A Cosine Similarity Fusion Network for Real-Time RGB-X Semantic Segmentation of Driving Scenes 1 Jul 2024 · 1 repository · arXiv:2407.01328
-
Deciphering the Factors Influencing the Efficacy of Chain-of-Thought: Probability, Memorization, and Noisy Reasoning 1 Jul 2024 · 1 repository · arXiv:2407.01687Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Understanding Multistationarity of Fully Open Reaction Networks 1 Jul 2024 · 0 repositories · arXiv:2407.01760
-
Domain Influence in MRI Medical Image Segmentation: spatial versus k-space inputs 1 Jul 2024 · 1 repository · arXiv:2407.01367
-
Eliminating Position Bias of Language Models: A Mechanistic Approach 1 Jul 2024 · 1 repository · arXiv:2407.01100Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Enhancing Speech-Driven 3D Facial Animation with Audio-Visual Guidance from Lip Reading Expert 1 Jul 2024 · 0 repositories · arXiv:2407.01034
-
Face4RAG: Factual Consistency Evaluation for Retrieval Augmented Generation in Chinese 1 Jul 2024 · 0 repositories · arXiv:2407.01080
-
FORA: Fast-Forward Caching in Diffusion Transformer Acceleration 1 Jul 2024 · 1 repository · arXiv:2407.01425Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 4 honoured, 0 violated, 1 with no contract checked; 6 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
GAT-Steiner: Rectilinear Steiner Minimal Tree Prediction Using GNNs 1 Jul 2024 · 0 repositories · arXiv:2407.01440
-
Ground Every Sentence: Improving Retrieval-Augmented LLMs with Interleaved Reference-Claim Generation 1 Jul 2024 · 0 repositories · arXiv:2407.01796
-
HGNET: A Hierarchical Feature Guided Network for Occupancy Flow Field Prediction 1 Jul 2024 · 0 repositories · arXiv:2407.01097
-
How Does Overparameterization Affect Features? 1 Jul 2024 · 0 repositories · arXiv:2407.00968
-
Hybrid RAG-empowered Multi-modal LLM for Secure Data Management in Internet of Medical Things: A Diffusion-based Contract Approach 1 Jul 2024 · 0 repositories · arXiv:2407.00978
-
Hypformer: Exploring Efficient Hyperbolic Transformer Fully in Hyperbolic Space 1 Jul 2024 · 1 repository · arXiv:2407.01290Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Image-to-Text Logic Jailbreak: Your Imagination can Help You Do Anything 1 Jul 2024 · 0 repositories · arXiv:2407.02534
-
Improving Trip Mode Choice Modeling Using Ensemble Synthesizer (ENSY) 1 Jul 2024 · 0 repositories · arXiv:2407.01769
-
Increasing Model Capacity for Free: A Simple Strategy for Parameter Efficient Fine-tuning 1 Jul 2024 · 1 repository · arXiv:2407.01320Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 7 pointer-only (licence)
-
Investigating the potential of Sparse Mixtures-of-Experts for multi-domain neural machine translation 1 Jul 2024 · 0 repositories · arXiv:2407.01126
-
Large Language Model Enhanced Knowledge Representation Learning: A Survey 1 Jul 2024 · 0 repositories · arXiv:2407.00936
-
A Lightweight UDF Learning Framework for 3D Reconstruction Based on Local Shape Functions 1 Jul 2024 · 1 repository · arXiv:2407.01330
-
Efficient Automated Circuit Discovery in Transformers using Contextual Decomposition 1 Jul 2024 · 0 repositories · arXiv:2407.00886
-
Multi-branch CNN and grouping cascade attention for medical image classification 1 Jul 2024 · 0 repositories
-
Multi-Modal Fusion-Based Multi-Task Semantic Communication System 1 Jul 2024 · 0 repositories · arXiv:2407.00964
-
Multi-State-Action Tokenisation in Decision Transformers for Multi-Discrete Action Spaces 1 Jul 2024 · 0 repositories · arXiv:2407.01310
-
Needle in the Haystack for Memory Based Large Language Models 1 Jul 2024 · 0 repositories · arXiv:2407.01437
-
Papez: Resource-Efficient Speech Separation with Auditory Working Memory 1 Jul 2024 · 1 repository · arXiv:2407.00888
-
Pictures Of MIDI: Controlled Music Generation via Graphical Prompts for Image-Based Diffusion Inpainting 1 Jul 2024 · 0 repositories · arXiv:2407.01499
-
Predicting DC-Link Capacitor Current Ripple in AC-DC Rectifier Circuits Using Fine-Tuned Large Language Models 1 Jul 2024 · 0 repositories · arXiv:2407.01724
-
Pron vs Prompt: Can Large Language Models already Challenge a World-Class Fiction Author at Creative Text Writing? 1 Jul 2024 · 0 repositories · arXiv:2407.01119
-
Race and Privacy in Broadcast Police Communications 1 Jul 2024 · 0 repositories · arXiv:2407.01817
-
Random Attention and Unobserved Reference Alternatives 1 Jul 2024 · 0 repositories · arXiv:2407.01528