Methods › General › Attention Mechanisms › Attention › Papers, page 125
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 125 of 316: papers 12,401 to 12,500 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Retrieval-augmented generation in multilingual settings 1 Jul 2024 · 1 repository · arXiv:2407.01463Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Roleplay-doh: Enabling Domain-Experts to Create LLM-simulated Patients via Eliciting and Adhering to Principles 1 Jul 2024 · 0 repositories · arXiv:2407.00870
-
Searching for Best Practices in Retrieval-Augmented Generation 1 Jul 2024 · 1 repository · arXiv:2407.01219Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Summary of a Haystack: A Challenge to Long-Context LLMs and RAG Systems 1 Jul 2024 · 1 repository · arXiv:2407.01370Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Terahertz Communication Multi-UAV-Assisted Mobile Edge Computing System 1 Jul 2024 · 0 repositories · arXiv:2407.01086
-
Memory³: Language Modeling with Explicit Memory 1 Jul 2024 · 0 repositories · arXiv:2407.01178
-
The Solution for Temporal Sound Localisation Task of ICCV 1st Perception Test Challenge 2023 1 Jul 2024 · 0 repositories · arXiv:2407.02318
-
Transferable-guided Attention Is All You Need for Video Domain Adaptation 1 Jul 2024 · 1 repository · arXiv:2407.01375
-
Uni-DVPS: Unified Model for Depth-Aware Video Panoptic Segmentation 1 Jul 2024 · 1 repository
-
Dynamic Universal Approximation Theory: The Basic Theory for Transformer-based Large Language Models 1 Jul 2024 · 0 repositories · arXiv:2407.00958
-
We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning? 1 Jul 2024 · 1 repository · arXiv:2407.01284Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
A Comparative Study of Quality Evaluation Methods for Text Summarization 30 Jun 2024 · 0 repositories · arXiv:2407.00747
-
Characterizing Stereotypical Bias from Privacy-preserving Pre-Training 30 Jun 2024 · 0 repositories · arXiv:2407.00764
-
Efficient Personalized Text-to-image Generation by Leveraging Textual Subspace 30 Jun 2024 · 1 repository · arXiv:2407.00608
-
Evaluation of Bias Towards Medical Professionals in Large Language Models 30 Jun 2024 · 0 repositories · arXiv:2407.12031
-
Exploring a Physics-Informed Decision Transformer for Distribution System Restoration: Methodology and Performance Analysis 30 Jun 2024 · 0 repositories · arXiv:2407.00808
-
GC-Bench: An Open and Unified Benchmark for Graph Condensation 30 Jun 2024 · 1 repository · arXiv:2407.00615Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Heterogeneous Graph Contrastive Learning with Spectral Augmentation 30 Jun 2024 · 0 repositories · arXiv:2407.00708
-
Instruct-IPT: All-in-One Image Processing Transformer via Weight Modulation 30 Jun 2024 · 1 repository · arXiv:2407.00676
-
LegalTurk Optimized BERT for Multi-Label Text Classification and NER 30 Jun 2024 · 0 repositories · arXiv:2407.00648
-
Multi-Agent Training for Pommerman: Curriculum Learning and Population-based Self-Play Approach 30 Jun 2024 · 0 repositories · arXiv:2407.00662
-
NAIST Simultaneous Speech Translation System for IWSLT 2024 30 Jun 2024 · 0 repositories · arXiv:2407.00826
-
Neurodevelopmental disorders modeling using isogeometric analysis, dynamic domain expansion and local refinement 30 Jun 2024 · 1 repository · arXiv:2407.00810
-
Parm: Efficient Training of Large Sparsely-Activated Models with Dedicated Schedules 30 Jun 2024 · 1 repository · arXiv:2407.00599
-
Prediction of Sentinel-2 multi-band imagery with attention BiLSTM for continuous earth surface monitoring 30 Jun 2024 · 0 repositories · arXiv:2407.00834
-
Safe Reinforcement Learning for Power System Control: A Review 30 Jun 2024 · 0 repositories · arXiv:2407.00681
-
A Two-stage Reinforcement Learning-based Approach for Multi-entity Task Allocation 29 Jun 2024 · 1 repository · arXiv:2407.00496
-
Answering real-world clinical questions using large language model based systems 29 Jun 2024 · 0 repositories · arXiv:2407.00541
-
Deciphering interventional dynamical causality from non-intervention systems 29 Jun 2024 · 0 repositories · arXiv:2407.01621
-
From RAG to RICHES: Retrieval Interlaced with Sequence Generation 29 Jun 2024 · 0 repositories · arXiv:2407.00361
-
Human-Agent Joint Learning for Efficient Robot Manipulation Skill Acquisition 29 Jun 2024 · 0 repositories · arXiv:2407.00299
-
Interpreting Pretrained Speech Models for Automatic Speech Assessment of Voice Disorders 29 Jun 2024 · 0 repositories · arXiv:2407.00531
-
Intrinsic PAPR for Point-level 3D Scene Albedo and Shading Editing 29 Jun 2024 · 0 repositories · arXiv:2407.00500
-
Learning Unsupervised Gaze Representation via Eye Mask Driven Information Bottleneck 29 Jun 2024 · 0 repositories · arXiv:2407.00315
-
LLM-Generated Natural Language Meets Scaling Laws: New Explorations and Data Augmentation Methods 29 Jun 2024 · 0 repositories · arXiv:2407.00322
-
Nonequilibrium dynamics and thermodynamics provide the underlying physical mechanism of the perceptual rivalry 29 Jun 2024 · 0 repositories · arXiv:2407.00350
-
Reconfigurable Intelligent Surface Identification in Mobile Networks: Opportunities and Challenges 29 Jun 2024 · 0 repositories · arXiv:2407.04731
-
Too Late to Train, Too Early To Use? A Study on Necessity and Viability of Low-Resource Bengali LLMs 29 Jun 2024 · 0 repositories · arXiv:2407.00416
-
Towards Universal Mesh Movement Networks 29 Jun 2024 · 1 repository · arXiv:2407.00382Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples)
-
Urban Visual Appeal According to ChatGPT: Contrasting AI and Human Insights 29 Jun 2024 · 0 repositories · arXiv:2407.14268
-
A Simple Attention-Based Mechanism for Bimodal Emotion Classification 28 Jun 2024 · 0 repositories · arXiv:2407.00134
-
AnomaLLMy -- Detecting anomalous tokens in black-box LLMs through low-confidence single-token predictions 28 Jun 2024 · 1 repository · arXiv:2406.19840
-
Applying RLAIF for Code Generation with API-usage in Lightweight LLMs 28 Jun 2024 · 0 repositories · arXiv:2406.20060
-
AstMatch: Adversarial Self-training Consistency Framework for Semi-Supervised Medical Image Segmentation 28 Jun 2024 · 1 repository · arXiv:2406.19649
-
Attention Meets UAVs: A Comprehensive Evaluation of DDoS Detection in Low-Cost UAVs 28 Jun 2024 · 0 repositories · arXiv:2406.19881
-
Beyond First-Order: A Multi-Scale Approach to Finger Knuckle Print Biometrics 28 Jun 2024 · 0 repositories · arXiv:2406.19672
-
BioMNER: A Dataset for Biomedical Method Entity Recognition 28 Jun 2024 · 0 repositories · arXiv:2406.20038
-
BMW Agents -- A Framework For Task Automation Through Multi-Agent Collaboration 28 Jun 2024 · 0 repositories · arXiv:2406.20041
-
Can GPT-4 Help Detect Quit Vaping Intentions? An Exploration of Automatic Data Annotation Approach 28 Jun 2024 · 0 repositories · arXiv:2407.00167
-
Composite Adaptive Disturbance Rejection in Robotics via Instrumental Variables based DREM 28 Jun 2024 · 0 repositories · arXiv:2406.19838
-
Covert Malicious Finetuning: Challenges in Safeguarding LLM Adaptation 28 Jun 2024 · 0 repositories · arXiv:2406.20053
-
Directly Training Temporal Spiking Neural Network with Sparse Surrogate Gradient 28 Jun 2024 · 0 repositories · arXiv:2406.19645
-
Evaluating Human Alignment and Model Faithfulness of LLM Rationale 28 Jun 2024 · 0 repositories · arXiv:2407.00219
-
EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model 28 Jun 2024 · 1 repository · arXiv:2406.20076Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Explainable Image Captioning using CNN- CNN architecture and Hierarchical Attention 28 Jun 2024 · 0 repositories · arXiv:2407.09556
-
Finite basis Kolmogorov-Arnold networks: domain decomposition for data-driven and physics-informed problems 28 Jun 2024 · 1 repository · arXiv:2406.19662
-
FootBots: A Transformer-based Architecture for Motion Prediction in Soccer 28 Jun 2024 · 0 repositories · arXiv:2406.19852
-
FRED: Flexible REduction-Distribution Interconnect and Communication Implementation for Wafer-Scale Distributed Training of DNN Models 28 Jun 2024 · 0 repositories · arXiv:2406.19580
-
Fuzzy Logic Guided Reward Function Variation: An Oracle for Testing Reinforcement Learning Programs 28 Jun 2024 · 1 repository · arXiv:2406.19812
-
Generative Iris Prior Embedded Transformer for Iris Restoration 28 Jun 2024 · 1 repository · arXiv:2407.00261
-
HouseCrafter: Lifting Floorplans to 3D Scenes with 2D Diffusion Model 28 Jun 2024 · 0 repositories · arXiv:2406.20077
-
InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache Management 28 Jun 2024 · 1 repository · arXiv:2406.19707Syntology 14 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
Joint Beamforming and Antenna Position Optimization for Movable Antenna-Assisted Spectrum Sharing 28 Jun 2024 · 0 repositories · arXiv:2406.19590
-
Less is More: Accurate Speech Recognition & Translation without Web-Scale Data 28 Jun 2024 · 0 repositories · arXiv:2406.19674
-
Machine Learning Predictors for Min-Entropy Estimation 28 Jun 2024 · 1 repository · arXiv:2406.19983
-
Mind the Gap: Analyzing Lacunae with Transformer-Based Transcription 28 Jun 2024 · 0 repositories · arXiv:2407.00250
-
Mixture of In-Context Experts Enhance LLMs' Long Context Awareness 28 Jun 2024 · 1 repository · arXiv:2406.19598Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Modeling the Real World with High-Density Visual Particle Dynamics 28 Jun 2024 · 0 repositories · arXiv:2406.19800
-
Multi-Satellite MIMO Systems for Direct User-Satellite Communications: A Survey 28 Jun 2024 · 0 repositories · arXiv:2407.00196
-
Multimodal Prototyping for cancer survival prediction 28 Jun 2024 · 1 repository · arXiv:2407.00224Syntology official (archive's flag): 7 ran · 7 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
PathGen-1.6M: 1.6 Million Pathology Image-text Pairs Generation through Multi-agent Collaboration 28 Jun 2024 · 1 repository · arXiv:2407.00203
-
PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators 28 Jun 2024 · 0 repositories · arXiv:2406.20083
-
PPTFormer: Pseudo Multi-Perspective Transformer for UAV Segmentation 28 Jun 2024 · 0 repositories · arXiv:2406.19632
-
ScaleBiO: Scalable Bilevel Optimization for LLM Data Reweighting 28 Jun 2024 · 0 repositories · arXiv:2406.19976
-
ShortcutsBench: A Large-Scale Real-world Benchmark for API-based Agents 28 Jun 2024 · 1 repository · arXiv:2407.00132Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
SK-VQA: Synthetic Knowledge Generation at Scale for Training Context-Augmented Multimodal LLMs 28 Jun 2024 · 0 repositories · arXiv:2406.19593
-
SMLT-MUGC: Small, Medium, and Large Texts -- Machine versus User-Generated Content Detection and Comparison 28 Jun 2024 · 0 repositories · arXiv:2407.12815
-
Solving Token Gradient Conflict in Mixture-of-Experts for Large Vision-Language Model 28 Jun 2024 · 1 repository · arXiv:2406.19905
-
The Computational Curse of Big Data for Bayesian Additive Regression Trees: A Hitting Time Analysis 28 Jun 2024 · 1 repository · arXiv:2406.19958
-
Uncertainty Quantification in Large Language Models Through Convex Hull Analysis 28 Jun 2024 · 0 repositories · arXiv:2406.19712
-
Vision Transformer with Key-select Routing Attention for Single Image Dehazing 28 Jun 2024 · 0 repositories · arXiv:2406.19703
-
A Sanity Check for AI-generated Image Detection 27 Jun 2024 · 2 repositories · arXiv:2406.19435Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 3 pointer-only (licence)
-
Aligning Teacher with Student Preferences for Tailored Training Data Generation 27 Jun 2024 · 0 repositories · arXiv:2406.19227
-
An Interpretable and Efficient Sleep Staging Algorithm: DetectsleepNet 27 Jun 2024 · 1 repository · arXiv:2406.19246
-
AutoPureData: Automated Filtering of Undesirable Web Data to Update LLM Knowledge 27 Jun 2024 · 1 repository · arXiv:2406.19271
-
AutoRAG-HP: Automatic Online Hyper-Parameter Tuning for Retrieval-Augmented Generation 27 Jun 2024 · 0 repositories · arXiv:2406.19251
-
Can Large Language Models Generate High-quality Patent Claims? 27 Jun 2024 · 1 repository · arXiv:2406.19465
-
Cost-efficient Active Illumination Camera For Hyper-spectral Reconstruction 27 Jun 2024 · 0 repositories · arXiv:2406.19560
-
Diminishing Stereotype Bias in Image Generation Model using Reinforcemenlent Learning Feedback 27 Jun 2024 · 0 repositories · arXiv:2407.09551
-
ELCoRec: Enhance Language Understanding with Co-Propagation of Numerical and Categorical Features for Recommendation 27 Jun 2024 · 0 repositories · arXiv:2406.18825
-
Enhancing Video-Language Representations with Structural Spatio-Temporal Alignment 27 Jun 2024 · 0 repositories · arXiv:2406.19255
-
Fairness and Bias in Multimodal AI: A Survey 27 Jun 2024 · 0 repositories · arXiv:2406.19097
-
Fibottention: Inceptive Visual Representation Learning with Diverse Attention Across Heads 27 Jun 2024 · 1 repository · arXiv:2406.19391
-
Fine-tuned network relies on generic representation to solve unseen cognitive task 27 Jun 2024 · 0 repositories · arXiv:2406.18926
-
From Artificial Needles to Real Haystacks: Improving Retrieval Capabilities in LLMs by Finetuning on Synthetic Data 27 Jun 2024 · 1 repository · arXiv:2406.19292
-
Granite-Function Calling Model: Introducing Function Calling Abilities via Multi-task Learning of Granular Tasks 27 Jun 2024 · 0 repositories · arXiv:2407.00121
-
Historia Magistra Vitae: Dynamic Topic Modeling of Roman Literature using Neural Embeddings 27 Jun 2024 · 0 repositories · arXiv:2406.18907
-
Human-Aware Vision-and-Language Navigation: Bridging Simulation to Reality with Dynamic Human Interactions 27 Jun 2024 · 1 repository · arXiv:2406.19236Syntology official (archive's flag): 7 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
IndoToxic2024: A Demographically-Enriched Dataset of Hate Speech and Toxicity Types for Indonesian Language 27 Jun 2024 · 0 repositories · arXiv:2406.19349
-
Learning Retrieval Augmentation for Personalized Dialogue Generation 27 Jun 2024 · 1 repository · arXiv:2406.18847Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 14 pointer-only (licence)