Methods › General › Attention Mechanisms › Attention › Papers, page 132
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 132 of 316: papers 13,101 to 13,200 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Tailoring Generative AI Chatbots for Multiethnic Communities in Disaster Preparedness Communication: Extending the CASA Paradigm 12 Jun 2024 · 1 repository · arXiv:2406.08411
-
Learning positional encodings in transformers depends on initialization 12 Jun 2024 · 0 repositories · arXiv:2406.08272
-
Transformer-based Model for ASR N-Best Rescoring and Rewriting 12 Jun 2024 · 0 repositories · arXiv:2406.08207
-
VeraCT Scan: Retrieval-Augmented Fake News Detection with Justifiable Reasoning 12 Jun 2024 · 0 repositories · arXiv:2406.10289
-
What If We Recaption Billions of Web Images with LLaMA-3? 12 Jun 2024 · 0 repositories · arXiv:2406.08478
-
Agent-SiMT: Agent-assisted Simultaneous Machine Translation with Large Language Models 11 Jun 2024 · 1 repository · arXiv:2406.06910
-
AI Sandbagging: Language Models can Strategically Underperform on Evaluations 11 Jun 2024 · 1 repository · arXiv:2406.07358Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Beyond Words: On Large Language Models Actionability in Mission-Critical Risk Analysis 11 Jun 2024 · 0 repositories · arXiv:2406.10273
-
Bilingual Sexism Classification: Fine-Tuned XLM-RoBERTa and GPT-3.5 Few-Shot Learning 11 Jun 2024 · 0 repositories · arXiv:2406.07287
-
COVID-19 Twitter Sentiment Classification Using Hybrid Deep Learning Model Based on Grid Search Methodology 11 Jun 2024 · 0 repositories · arXiv:2406.10266
-
DARA: Decomposition-Alignment-Reasoning Autonomous Language Agent for Question Answering over Knowledge Graphs 11 Jun 2024 · 1 repository · arXiv:2406.07080Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
DR-RAG: Applying Dynamic Document Relevance to Retrieval-Augmented Generation for Question-Answering 11 Jun 2024 · 0 repositories · arXiv:2406.07348
-
Effectively Compress KV Heads for LLM 11 Jun 2024 · 0 repositories · arXiv:2406.07056
-
Entropy-Reinforced Planning with Large Language Models for Drug Discovery 11 Jun 2024 · 1 repository · arXiv:2406.07025
-
Evolving Subnetwork Training for Large Language Models 11 Jun 2024 · 0 repositories · arXiv:2406.06962
-
FastAST: Accelerating Audio Spectrogram Transformer via Token Merging and Cross-Model Knowledge Distillation 11 Jun 2024 · 1 repository · arXiv:2406.07676
-
Flextron: Many-in-One Flexible Large Language Model 11 Jun 2024 · 0 repositories · arXiv:2406.10260
-
Grapevine Disease Prediction Using Climate Variables from Multi-Sensor Remote Sensing Imagery via a Transformer Model 11 Jun 2024 · 0 repositories · arXiv:2406.07094
-
GridPE: Unifying Positional Encoding in Transformers with a Grid Cell-Inspired Framework 11 Jun 2024 · 0 repositories · arXiv:2406.07049
-
MLLMGuard: A Multi-dimensional Safety Evaluation Suite for Multimodal Large Language Models 11 Jun 2024 · 1 repository · arXiv:2406.07594Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Multi-objective Reinforcement learning from AI Feedback 11 Jun 2024 · 1 repository · arXiv:2406.07295
-
Multimodal Belief Prediction 11 Jun 2024 · 1 repository · arXiv:2406.07466
-
Noise-robust Speech Separation with Fast Generative Correction 11 Jun 2024 · 1 repository · arXiv:2406.07461
-
Question-Answering (QA) Model for a Personalized Learning Assistant for Arabic Language 11 Jun 2024 · 0 repositories · arXiv:2406.08519
-
Samba: Simple Hybrid State Space Models for Efficient Unlimited Context Language Modeling 11 Jun 2024 · 2 repositories · arXiv:2406.07522Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Towards Generalized Hydrological Forecasting using Transformer Models for 120-Hour Streamflow Prediction 11 Jun 2024 · 0 repositories · arXiv:2406.07484
-
Unused information in token probability distribution of generative LLM: improving LLM reading comprehension through calculation of expected values 11 Jun 2024 · 1 repository · arXiv:2406.10267
-
UVIS: Unsupervised Video Instance Segmentation 11 Jun 2024 · 0 repositories · arXiv:2406.06908
-
Validating LLM-Generated Programs with Metamorphic Prompt Testing 11 Jun 2024 · 0 repositories · arXiv:2406.06864
-
A Comparative Survey of Vision Transformers for Feature Extraction in Texture Analysis 10 Jun 2024 · 0 repositories · arXiv:2406.06136
-
AGB-DE: A Corpus for the Automated Legal Assessment of Clauses in German Consumer Contracts 10 Jun 2024 · 1 repository · arXiv:2406.06809
-
Annotation alignment: Comparing LLM and human annotations of conversational safety 10 Jun 2024 · 0 repositories · arXiv:2406.06369
-
Can Language Models Serve as Text-Based World Simulators? 10 Jun 2024 · 0 repositories · arXiv:2406.06485
-
Compute Better Spent: Replacing Dense Layers with Structured Matrices 10 Jun 2024 · 1 repository · arXiv:2406.06248Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Data-Efficient Learning with Neural Programs 10 Jun 2024 · 1 repository · arXiv:2406.06246Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Diving into Underwater: Segment Anything Model Guided Underwater Salient Instance Segmentation and A Large-scale Dataset 10 Jun 2024 · 1 repository · arXiv:2406.06039
-
Emotion-Aware Speech Self-Supervised Representation Learning with Intensity Knowledge 10 Jun 2024 · 0 repositories · arXiv:2406.06646
-
Husky: A Unified, Open-Source Language Agent for Multi-Step Reasoning 10 Jun 2024 · 1 repository · arXiv:2406.06469Syntology official (archive's flag): 5 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 2 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
In-Context Learning and Fine-Tuning GPT for Argument Mining 10 Jun 2024 · 1 repository · arXiv:2406.06699
-
Learning Physical Simulation with Message Passing Transformer 10 Jun 2024 · 0 repositories · arXiv:2406.06060
-
Leveraging Large Language Models for Knowledge-free Weak Supervision in Clinical Natural Language Processing 10 Jun 2024 · 0 repositories · arXiv:2406.06723
-
LLM-dCache: Improving Tool-Augmented LLMs with GPT-Driven Localized Data Caching 10 Jun 2024 · 0 repositories · arXiv:2406.06799
-
PointABM:Integrating Bidirectional State Space Model with Multi-Head Self-Attention for Point Cloud Analysis 10 Jun 2024 · 0 repositories · arXiv:2406.06069
-
SecureNet: A Comparative Study of DeBERTa and Large Language Models for Phishing Detection 10 Jun 2024 · 0 repositories · arXiv:2406.06663
-
Symmetric Dot-Product Attention for Efficient Training of BERT Language Models 10 Jun 2024 · 0 repositories · arXiv:2406.06366
-
The Impact of Quantization on Retrieval-Augmented Generation: An Analysis of Small LLMs 10 Jun 2024 · 0 repositories · arXiv:2406.10251
-
UMBRELA: UMbrela is the (Open-Source Reproduction of the) Bing RELevance Assessor 10 Jun 2024 · 1 repository · arXiv:2406.06519
-
A Knowledge-Component-Based Methodology for Evaluating AI Assistants 9 Jun 2024 · 0 repositories · arXiv:2406.05603
-
Are Large Language Models Actually Good at Text Style Transfer? 9 Jun 2024 · 1 repository · arXiv:2406.05885
-
Attention as a Hypernetwork 9 Jun 2024 · 1 repository · arXiv:2406.05816Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
CAMS: Convolution and Attention-Free Mamba-based Cardiac Image Segmentation 9 Jun 2024 · 1 repository · arXiv:2406.05786
-
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation 9 Jun 2024 · 2 repositories · arXiv:2406.05654Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Exploring the Efficacy of Large Language Models (GPT-4) in Binary Reverse Engineering 9 Jun 2024 · 0 repositories · arXiv:2406.06637
-
GCtx-UNet: Efficient Network for Medical Image Segmentation 9 Jun 2024 · 1 repository · arXiv:2406.05891
-
Hidden Holes: topological aspects of language models 9 Jun 2024 · 0 repositories · arXiv:2406.05798
-
Large Language Models Memorize Sensor Datasets! Implications on Human Activity Recognition Research 9 Jun 2024 · 0 repositories · arXiv:2406.05900
-
Machine Against the RAG: Jamming Retrieval-Augmented Generation with Blocker Documents 9 Jun 2024 · 0 repositories · arXiv:2406.05870
-
MedREQAL: Examining Medical Knowledge Recall of Large Language Models via Question Answering 9 Jun 2024 · 0 repositories · arXiv:2406.05845
-
OD-DETR: Online Distillation for Stabilizing Training of Detection Transformer 9 Jun 2024 · 0 repositories · arXiv:2406.05791
-
RE-RAG: Improving Open-Domain QA Performance and Interpretability with Relevance Estimator in Retrieval-Augmented Generation 9 Jun 2024 · 1 repository · arXiv:2406.05794
-
SinkLoRA: Enhanced Efficiency and Chat Capabilities for Long-Context Large Language Models 9 Jun 2024 · 1 repository · arXiv:2406.05678Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Smiles2Dock: an open large-scale multi-task dataset for ML-based molecular docking 9 Jun 2024 · 1 repository · arXiv:2406.05738
-
Text2VP: Generative AI for Visual Programming and Parametric Modeling 9 Jun 2024 · 0 repositories · arXiv:2407.07732
-
Vision Mamba: Cutting-Edge Classification of Alzheimer's Disease with 3D MRI Scans 9 Jun 2024 · 0 repositories · arXiv:2406.05757
-
1st Place Winner of the 2024 Pixel-level Video Understanding in the Wild (CVPR'24 PVUW) Challenge in Video Panoptic Segmentation and Best Long Video Consistency of Video Semantic Segmentation 8 Jun 2024 · 0 repositories · arXiv:2406.05352
-
A Fine-tuning Dataset and Benchmark for Large Language Models for Protein Understanding 8 Jun 2024 · 1 repository · arXiv:2406.05540Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Advancing Semantic Textual Similarity Modeling: A Regression Framework with Translated ReLU and Smooth K2 Loss 8 Jun 2024 · 2 repositories · arXiv:2406.05326
-
Automata Extraction from Transformers 8 Jun 2024 · 1 repository · arXiv:2406.05564
-
Benchmarking Neural Decoding Backbones towards Enhanced On-edge iBCI Applications 8 Jun 2024 · 0 repositories · arXiv:2406.06626
-
Concept Formation and Alignment in Language Models: Bridging Statistical Patterns in Latent Space to Concept Taxonomy 8 Jun 2024 · 0 repositories · arXiv:2406.05315
-
Critical Phase Transition in Large Language Models 8 Jun 2024 · 0 repositories · arXiv:2406.05335
-
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts 8 Jun 2024 · 0 repositories · arXiv:2406.05569
-
G-Transformer: Counterfactual Outcome Prediction under Dynamic and Time-varying Treatment Regimes 8 Jun 2024 · 0 repositories · arXiv:2406.05504
-
MaTableGPT: GPT-based Table Data Extractor from Materials Science Literature 8 Jun 2024 · 0 repositories · arXiv:2406.05431
-
SelfDefend: LLMs Can Defend Themselves against Jailbreaking in a Practical Manner 8 Jun 2024 · 0 repositories · arXiv:2406.05498
-
Teaching-Assistant-in-the-Loop: Improving Knowledge Distillation from Imperfect Teacher Models in Low-Budget Scenarios 8 Jun 2024 · 0 repositories · arXiv:2406.05322
-
Toward Reliable Ad-hoc Scientific Information Extraction: A Case Study on Two Materials Datasets 8 Jun 2024 · 1 repository · arXiv:2406.05348
-
Transformer Conformal Prediction for Time Series 8 Jun 2024 · 1 repository · arXiv:2406.05332Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
U-Net Ensemble for Enhanced Semantic Segmentation in Remote Sensing Imagery 8 Jun 2024 · 0 repositories
-
VP-LLM: Text-Driven 3D Volume Completion with Large Language Models through Patchification 8 Jun 2024 · 0 repositories · arXiv:2406.05543
-
Are Large Language Models More Empathetic than Humans? 7 Jun 2024 · 0 repositories · arXiv:2406.05063
-
BAMO at SemEval-2024 Task 9: BRAINTEASER: A Novel Task Defying Common Sense 7 Jun 2024 · 1 repository · arXiv:2406.04947
-
BERTs are Generative In-Context Learners 7 Jun 2024 · 1 repository · arXiv:2406.04823Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 1 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 13 unverified (of 26 harvested samples)
-
Conti-Fuse: A Novel Continuous Decomposition-based Fusion Framework for Infrared and Visible Images 7 Jun 2024 · 0 repositories · arXiv:2406.04689
-
Corpus Poisoning via Approximate Greedy Gradient Descent 7 Jun 2024 · 1 repository · arXiv:2406.05087
-
CRAG -- Comprehensive RAG Benchmark 7 Jun 2024 · 2 repositories · arXiv:2406.04744Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
CTSyn: A Foundational Model for Cross Tabular Data Generation 7 Jun 2024 · 0 repositories · arXiv:2406.04619
-
DiNeR: a Large Realistic Dataset for Evaluating Compositional Generalization 7 Jun 2024 · 1 repository · arXiv:2406.04669
-
Diving Deep into the Motion Representation of Video-Text Models 7 Jun 2024 · 1 repository · arXiv:2406.05075
-
GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents 7 Jun 2024 · 2 repositories · arXiv:2406.06613Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Hints-In-Browser: Benchmarking Language Models for Programming Feedback Generation 7 Jun 2024 · 0 repositories · arXiv:2406.05053
-
DALD: Improving Logits-based Detector without Logits from Black-box LLMs 7 Jun 2024 · 1 repository · arXiv:2406.05232
-
Large Generative Graph Models 7 Jun 2024 · 0 repositories · arXiv:2406.05109
-
LlavaGuard: An Open VLM-based Framework for Safeguarding Vision Datasets and Models 7 Jun 2024 · 1 repository · arXiv:2406.05113Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
LLMs Are Not Intelligent Thinkers: Introducing Mathematical Topic Tree Benchmark for Comprehensive Evaluation of LLMs 7 Jun 2024 · 1 repository · arXiv:2406.05194
-
Logic Synthesis with Generative Deep Neural Networks 7 Jun 2024 · 0 repositories · arXiv:2406.04699
-
Low-Resource Cross-Lingual Summarization through Few-Shot Learning with Large Language Models 7 Jun 2024 · 0 repositories · arXiv:2406.04630
-
Mixture-of-Agents Enhances Large Language Model Capabilities 7 Jun 2024 · 3 repositories · arXiv:2406.04692Syntology 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
Multi-Head RAG: Solving Multi-Aspect Problems with LLMs 7 Jun 2024 · 2 repositories · arXiv:2406.05085
-
Multiplane Prior Guided Few-Shot Aerial Scene Rendering 7 Jun 2024 · 0 repositories · arXiv:2406.04961