Methods › General › Attention Mechanisms › Attention › Papers, page 148
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 148 of 316: papers 14,701 to 14,800 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Contact-aware Human Motion Generation from Textual Descriptions 23 Mar 2024 · 0 repositories · arXiv:2403.15709
-
EAGLE: A Domain Generalization Framework for AI-generated Text Detection 23 Mar 2024 · 0 repositories · arXiv:2403.15690
-
Fine Tuning LLM for Enterprise: Practical Guidelines and Recommendations 23 Mar 2024 · 0 repositories · arXiv:2404.10779
-
Improving Retrieval for RAG based Question Answering Models on Financial Documents 23 Mar 2024 · 0 repositories · arXiv:2404.07221
-
LlamBERT: Large-scale low-cost data annotation in NLP 23 Mar 2024 · 1 repository · arXiv:2403.15938
-
Once for Both: Single Stage of Importance and Sparsity Search for Vision Transformer Compression 23 Mar 2024 · 1 repository · arXiv:2403.15835Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
Technical Report: Masked Skeleton Sequence Modeling for Learning Larval Zebrafish Behavior Latent Embeddings 23 Mar 2024 · 0 repositories · arXiv:2403.15693
-
Towards a RAG-based Summarization Agent for the Electron-Ion Collider 23 Mar 2024 · 1 repository · arXiv:2403.15729
-
Understanding Emergent Abilities of Language Models from the Loss Perspective 23 Mar 2024 · 0 repositories · arXiv:2403.15796
-
Using Large Language Models for OntoClean-based Ontology Refinement 23 Mar 2024 · 0 repositories · arXiv:2403.15864
-
SOEN-101: Code Generation by Emulating Software Process Models Using Large Language Model Agents 23 Mar 2024 · 0 repositories · arXiv:2403.15852
-
Adapprox: Adaptive Approximation in Adam Optimization via Randomized Low-Rank Matrices 22 Mar 2024 · 0 repositories · arXiv:2403.14958
-
Blended RAG: Improving RAG (Retriever-Augmented Generation) Accuracy with Semantic Search and Hybrid Query-Based Retrievers 22 Mar 2024 · 1 repository · arXiv:2404.07220Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
BSNet: Box-Supervised Simulation-assisted Mean Teacher for 3D Instance Segmentation 22 Mar 2024 · 1 repository · arXiv:2403.15019Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Can large language models explore in-context? 22 Mar 2024 · 0 repositories · arXiv:2403.15371
-
Comprehensive Evaluation and Insights into the Use of Large Language Models in the Automation of Behavior-Driven Development Acceptance Test Formulation 22 Mar 2024 · 1 repository · arXiv:2403.14965
-
Construction of a Japanese Financial Benchmark for Large Language Models 22 Mar 2024 · 1 repository · arXiv:2403.15062Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
ESG Classification by Implicit Rule Learning via GPT-4 22 Mar 2024 · 0 repositories · arXiv:2403.15040
-
GTC: GNN-Transformer Co-contrastive Learning for Self-supervised Heterogeneous Graph Representation 22 Mar 2024 · 1 repository · arXiv:2403.15520Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 3 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Hierarchical Information Enhancement Network for Cascade Prediction in Social Networks 22 Mar 2024 · 0 repositories · arXiv:2403.15257
-
LLaVA-PruMerge: Adaptive Token Reduction for Efficient Large Multimodal Models 22 Mar 2024 · 1 repository · arXiv:2403.15388Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
MasonTigers at SemEval-2024 Task 1: An Ensemble Approach for Semantic Textual Relatedness 22 Mar 2024 · 0 repositories · arXiv:2403.14990
-
Measuring Gender and Racial Biases in Large Language Models 22 Mar 2024 · 0 repositories · arXiv:2403.15281
-
Neural Plasticity-Inspired Multimodal Foundation Model for Earth Observation 22 Mar 2024 · 1 repository · arXiv:2403.15356Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
On Zero-Shot Counterspeech Generation by LLMs 22 Mar 2024 · 1 repository · arXiv:2403.14938
-
Optimal path for Biomedical Text Summarization Using Pointer GPT 22 Mar 2024 · 0 repositories · arXiv:2404.08654
-
ParFormer: A Vision Transformer with Parallel Mixer and Sparse Channel Attention Patch Embedding 22 Mar 2024 · 0 repositories · arXiv:2403.15004
-
Reasoning-Enhanced Object-Centric Learning for Videos 22 Mar 2024 · 0 repositories · arXiv:2403.15245
-
Selecting Query-bag as Pseudo Relevance Feedback for Information-seeking Conversations 22 Mar 2024 · 0 repositories · arXiv:2404.04272
-
Selectively Informative Description can Reduce Undesired Embedding Entanglements in Text-to-Image Personalization 22 Mar 2024 · 0 repositories · arXiv:2403.15330
-
SensoryT5: Infusing Sensorimotor Norms into T5 for Enhanced Fine-grained Emotion Classification 22 Mar 2024 · 0 repositories · arXiv:2403.15574
-
Text Clustering with Large Language Model Embeddings 22 Mar 2024 · 0 repositories · arXiv:2403.15112
-
Vehicle Detection Performance in Nordic Region 22 Mar 2024 · 0 repositories · arXiv:2403.15017
-
A Chain-of-Thought Prompting Approach with LLMs for Evaluating Students' Formative Assessment Responses in Science 21 Mar 2024 · 0 repositories · arXiv:2403.14565
-
A Large-Scale Network Construction and Lightweighting Method for Point Cloud Semantic Segmentation 21 Mar 2024 · 1 repository
-
AI and Memory Wall 21 Mar 2024 · 0 repositories · arXiv:2403.14123
-
Assessing the Utility of Large Language Models for Phenotype-Driven Gene Prioritization in Rare Genetic Disorder Diagnosis 21 Mar 2024 · 0 repositories · arXiv:2403.14801
-
Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference 21 Mar 2024 · 1 repository · arXiv:2403.14520Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Emergent World Models and Latent Variable Estimation in Chess-Playing Language Models 21 Mar 2024 · 1 repository · arXiv:2403.15498Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Exploring the Potential of Large Language Models in Graph Generation 21 Mar 2024 · 0 repositories · arXiv:2403.14358
-
Extracting Emotion Phrases from Tweets using BART 21 Mar 2024 · 0 repositories · arXiv:2403.14050
-
FIT-RAG: Black-Box RAG with Factual Information and Token Reduction 21 Mar 2024 · 0 repositories · arXiv:2403.14374
-
K-Act2Emo: Korean Commonsense Knowledge Graph for Indirect Emotional Expression 21 Mar 2024 · 1 repository · arXiv:2403.14253
-
LDTR: Transformer-based Lane Detection with Anchor-chain Representation 21 Mar 2024 · 0 repositories · arXiv:2403.14354
-
Learning with SASQuaTCh: a Novel Variational Quantum Transformer Architecture with Kernel-Based Self-Attention 21 Mar 2024 · 0 repositories · arXiv:2403.14753
-
LLM-based Extraction of Contradictions from Patents 21 Mar 2024 · 0 repositories · arXiv:2403.14258
-
OTSeg: Multi-prompt Sinkhorn Attention for Zero-Shot Semantic Segmentation 21 Mar 2024 · 1 repository · arXiv:2403.14183Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
PSALM: Pixelwise SegmentAtion with Large Multi-Modal Model 21 Mar 2024 · 1 repository · arXiv:2403.14598Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
ReAct Meets ActRe: When Language Agents Enjoy Training Data Autonomy 21 Mar 2024 · 0 repositories · arXiv:2403.14589
-
S2LIC: Learned Image Compression with the SwinV2 Block, Adaptive Channel-wise and Global-inter Attention Context 21 Mar 2024 · 1 repository · arXiv:2403.14471
-
Speech-Aware Neural Diarization with Encoder-Decoder Attractor Guided by Attention Constraints 21 Mar 2024 · 0 repositories · arXiv:2403.14268
-
SpikeGraphormer: A High-Performance Graph Transformer with Spiking Graph Attention 21 Mar 2024 · 1 repository · arXiv:2403.15480Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
SpikingResformer: Bridging ResNet and Vision Transformer in Spiking Neural Networks 21 Mar 2024 · 2 repositories · arXiv:2403.14302Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Token Transformation Matters: Towards Faithful Post-hoc Explanation for Vision Transformer 21 Mar 2024 · 0 repositories · arXiv:2403.14552
-
Toward Multi-class Anomaly Detection: Exploring Class-aware Unified Model against Inter-class Interference 21 Mar 2024 · 0 repositories · arXiv:2403.14213
-
Unsupervised Audio-Visual Segmentation with Modality Alignment 21 Mar 2024 · 0 repositories · arXiv:2403.14203
-
VURF: A General-purpose Reasoning and Self-refinement Framework for Video Understanding 21 Mar 2024 · 1 repository · arXiv:2403.14743
-
AMP: Autoregressive Motion Prediction Revisited with Next Token Prediction for Autonomous Driving 20 Mar 2024 · 0 repositories · arXiv:2403.13331
-
AUD-TGN: Advancing Action Unit Detection with Temporal Convolution and GPT-2 in Wild Audiovisual Contexts 20 Mar 2024 · 0 repositories · arXiv:2403.13678
-
Ax-to-Grind Urdu: Benchmark Dataset for Urdu Fake News Detection 20 Mar 2024 · 1 repository · arXiv:2403.14037
-
DiffImpute: Tabular Data Imputation With Denoising Diffusion Probabilistic Model 20 Mar 2024 · 0 repositories · arXiv:2403.13863
-
Efficient argument classification with compact language models and ChatGPT-4 refinements 20 Mar 2024 · 0 repositories · arXiv:2403.15473
-
Facilitating Pornographic Text Detection for Open-Domain Dialogue Systems via Knowledge Distillation of Large Language Models 20 Mar 2024 · 1 repository · arXiv:2403.13250
-
High-confidence pseudo-labels for domain adaptation in COVID-19 detection 20 Mar 2024 · 0 repositories · arXiv:2403.13509
-
Incentivizing News Consumption on Social Media Platforms Using Large Language Models and Realistic Bot Accounts 20 Mar 2024 · 1 repository · arXiv:2403.13362
-
Motion Generation from Fine-grained Textual Descriptions 20 Mar 2024 · 1 repository · arXiv:2403.13518
-
MTP: Advancing Remote Sensing Foundation Model via Multi-Task Pretraining 20 Mar 2024 · 2 repositories · arXiv:2403.13430Syntology official (archive's flag): 1 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Natural Language as Policies: Reasoning for Coordinate-Level Embodied Control with LLMs 20 Mar 2024 · 0 repositories · arXiv:2403.13801
-
Open Access NAO (OAN): a ROS2-based software framework for HRI applications with the NAO robot 20 Mar 2024 · 0 repositories · arXiv:2403.13960
-
PARAMANU-AYN: Pretrain from scratch or Continual Pretraining of LLMs for Legal Domain Adaptation? 20 Mar 2024 · 0 repositories · arXiv:2403.13681
-
Portrait4D-v2: Pseudo Multi-View Data Creates Better 4D Head Synthesizer 20 Mar 2024 · 0 repositories · arXiv:2403.13570
-
Retina Vision Transformer (RetinaViT): Introducing Scaled Patches into Vision Transformers 20 Mar 2024 · 1 repository · arXiv:2403.13677
-
Rotary Position Embedding for Vision Transformer 20 Mar 2024 · 2 repositories · arXiv:2403.13298Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
T-Pixel2Mesh: Combining Global and Local Transformer for 3D Mesh Generation from a Single Image 20 Mar 2024 · 0 repositories · arXiv:2403.13663
-
Vi-Mistral-X: Building a Vietnamese Language Model with Advanced Continual Pre-training 20 Mar 2024 · 0 repositories · arXiv:2403.15470
-
VL-Mamba: Exploring State Space Models for Multimodal Learning 20 Mar 2024 · 0 repositories · arXiv:2403.13600
-
A Comparison of Deep Learning Architectures for Spacecraft Anomaly Detection 19 Mar 2024 · 0 repositories · arXiv:2403.12864
-
Automated Data Curation for Robust Language Model Fine-Tuning 19 Mar 2024 · 0 repositories · arXiv:2403.12776
-
Automatic Information Extraction From Employment Tribunal Judgements Using Large Language Models 19 Mar 2024 · 0 repositories · arXiv:2403.12936
-
Automatic Summarization of Doctor-Patient Encounter Dialogues Using Large Language Model through Prompt Tuning 19 Mar 2024 · 0 repositories · arXiv:2403.13089
-
Can AI Outperform Human Experts in Creating Social Media Creatives? 19 Mar 2024 · 0 repositories · arXiv:2404.00018
-
DeblurDiNAT: A Compact Model with Exceptional Generalization and Visual Fidelity on Unseen Domains 19 Mar 2024 · 1 repository · arXiv:2403.13163
-
Diffusion-Driven Self-Supervised Learning for Shape Reconstruction and Pose Estimation 19 Mar 2024 · 1 repository · arXiv:2403.12728
-
Emotion Recognition Using Transformers with Masked Learning 19 Mar 2024 · 1 repository · arXiv:2403.13731
-
Efficient Encoder-Decoder Transformer Decoding for Decomposable Tasks 19 Mar 2024 · 1 repository · arXiv:2403.13112
-
Fine-Tuning Pre-trained Language Models to Detect In-Game Trash Talks 19 Mar 2024 · 0 repositories · arXiv:2403.15458
-
FlowerFormer: Empowering Neural Architecture Encoding using a Flow-aware Graph Transformer 19 Mar 2024 · 1 repository · arXiv:2403.12821Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
GraphERE: Jointly Multiple Event-Event Relation Extraction via Graph-Enhanced Event Embeddings 19 Mar 2024 · 0 repositories · arXiv:2403.12523
-
Improved EATFormer: A Vision Transformer for Medical Image Classification 19 Mar 2024 · 0 repositories · arXiv:2403.13167
-
End-to-End Neuro-Symbolic Reinforcement Learning with Textual Explanations 19 Mar 2024 · 1 repository · arXiv:2403.12451Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
Instructing Large Language Models to Identify and Ignore Irrelevant Conditions 19 Mar 2024 · 1 repository · arXiv:2403.12744
-
LHMKE: A Large-scale Holistic Multi-subject Knowledge Evaluation Benchmark for Chinese Large Language Models 19 Mar 2024 · 0 repositories · arXiv:2403.12601
-
LLMLingua-2: Data Distillation for Efficient and Faithful Task-Agnostic Prompt Compression 19 Mar 2024 · 1 repository · arXiv:2403.12968
-
Multimodal Fusion Method with Spatiotemporal Sequences and Relationship Learning for Valence-Arousal Estimation 19 Mar 2024 · 0 repositories · arXiv:2403.12425
-
Pipelined Biomedical Event Extraction Rivaling Joint Learning 19 Mar 2024 · 0 repositories · arXiv:2403.12386
-
Pragmatic Competence Evaluation of Large Language Models for the Korean Language 19 Mar 2024 · 1 repository · arXiv:2403.12675
-
RankPrompt: Step-by-Step Comparisons Make Language Models Better Reasoners 19 Mar 2024 · 0 repositories · arXiv:2403.12373
-
SEVEN: Pruning Transformer Model by Reserving Sentinels 19 Mar 2024 · 1 repository · arXiv:2403.12688
-
Simple Hack for Transformers against Heavy Long-Text Classification on a Time- and Memory-Limited GPU Service 19 Mar 2024 · 0 repositories · arXiv:2403.12563
-
Quantifying uncertainty in lung cancer segmentation with foundation models applied to mixed-domain datasets 19 Mar 2024 · 0 repositories · arXiv:2403.13113