Methods › General › Attention Mechanisms › Attention › Papers, page 162
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 162 of 316: papers 16,101 to 16,200 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
MMToM-QA: Multimodal Theory of Mind Question Answering 16 Jan 2024 · 1 repository · arXiv:2401.08743
-
Mobile Contactless Palmprint Recognition: Use of Multiscale, Multimodel Embeddings 16 Jan 2024 · 0 repositories · arXiv:2401.08111
-
RAG vs Fine-tuning: Pipelines, Tradeoffs, and a Case Study on Agriculture 16 Jan 2024 · 0 repositories · arXiv:2401.08406
-
RoTBench: A Multi-Level Benchmark for Evaluating the Robustness of Large Language Models in Tool Learning 16 Jan 2024 · 1 repository · arXiv:2401.08326Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 15 harvested samples)
-
Scalable Pre-training of Large Autoregressive Image Models 16 Jan 2024 · 2 repositories · arXiv:2401.08541
-
Small Object Detection by DETR via Information Augmentation and Adaptive Feature Fusion 16 Jan 2024 · 0 repositories · arXiv:2401.08017
-
Solving Continual Offline Reinforcement Learning with Decision Transformer 16 Jan 2024 · 0 repositories · arXiv:2401.08478
-
Statistical Test for Attention Map in Vision Transformer 16 Jan 2024 · 1 repository · arXiv:2401.08169
-
Transcending the Limit of Local Window: Advanced Super-Resolution Transformer with Adaptive Token Dictionary 16 Jan 2024 · 1 repository · arXiv:2401.08209Syntology official (archive's flag): 11 ran · 11 ran (of which 9 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 15 harvested samples)
-
Tuning Language Models by Proxy 16 Jan 2024 · 2 repositories · arXiv:2401.08565Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Video Quality Assessment Based on Swin TransformerV2 and Coarse to Fine Strategy 16 Jan 2024 · 0 repositories · arXiv:2401.08522
-
A character-based steganography using masked language modeling 15 Jan 2024 · 1 repository
-
A Deep Hierarchical Feature Sparse Framework for Occluded Person Re-Identification 15 Jan 2024 · 0 repositories · arXiv:2401.07469
-
A Novel Approach for Automatic Program Repair using Round-Trip Translation with Large Language Models 15 Jan 2024 · 1 repository · arXiv:2401.07994
-
Combining Image- and Geometric-based Deep Learning for Shape Regression: A Comparison to Pixel-level Methods for Segmentation in Chest X-Ray 15 Jan 2024 · 0 repositories · arXiv:2401.07542
-
Consolidating Trees of Robotic Plans Generated Using Large Language Models to Improve Reliability 15 Jan 2024 · 0 repositories · arXiv:2401.07868
-
Exploiting GPT-4 Vision for Zero-shot Point Cloud Understanding 15 Jan 2024 · 0 repositories · arXiv:2401.07572
-
Graph database while computationally efficient filters out quickly the ESG integrated equities in investment management 15 Jan 2024 · 0 repositories · arXiv:2401.07483
-
Graph Transformer GANs with Graph Masked Modeling for Architectural Layout Generation 15 Jan 2024 · 0 repositories · arXiv:2401.07721
-
Image Similarity using An Ensemble of Context-Sensitive Models 15 Jan 2024 · 1 repository · arXiv:2401.07951
-
Towards Efficient Methods in Medical Question Answering using Knowledge Graph Embeddings 15 Jan 2024 · 1 repository · arXiv:2401.07977Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Leveraging the power of transformers for guilt detection in text 15 Jan 2024 · 0 repositories · arXiv:2401.07414
-
Milestones in Bengali Sentiment Analysis leveraging Transformer-models: Fundamentals, Challenges and Future Directions 15 Jan 2024 · 0 repositories · arXiv:2401.07847
-
One for All: Toward Unified Foundation Models for Earth Vision 15 Jan 2024 · 0 repositories · arXiv:2401.07527
-
SemEval-2017 Task 4: Sentiment Analysis in Twitter using BERT 15 Jan 2024 · 1 repository · arXiv:2401.07944
-
The Chronicles of RAG: The Retriever, the Chunk and the Generator 15 Jan 2024 · 0 repositories · arXiv:2401.07883
-
Understanding YTHDF2-mediated mRNA Degradation By m6A-BERT-Deg 15 Jan 2024 · 1 repository · arXiv:2401.08004
-
3D Landmark Detection on Human Point Clouds: A Benchmark and A Dual Cascade Point Transformer Framework 14 Jan 2024 · 0 repositories · arXiv:2401.07251
-
Harnessing Large Language Models Over Transformer Models for Detecting Bengali Depressive Social Media Text: A Comprehensive Study 14 Jan 2024 · 1 repository · arXiv:2401.07310
-
Killer Apps: Low-Speed, Large-Scale AI Weapons 14 Jan 2024 · 0 repositories · arXiv:2402.01663
-
Learning to be Homo Economicus: Can an LLM Learn Preferences from Choice 14 Jan 2024 · 0 repositories · arXiv:2401.07345
-
MapGPT: Map-Guided Prompting with Adaptive Path Planning for Vision-and-Language Navigation 14 Jan 2024 · 0 repositories · arXiv:2401.07314
-
Promptformer: Prompted Conformer Transducer for ASR 14 Jan 2024 · 0 repositories · arXiv:2401.07360
-
Streamlining the Selection Phase of Systematic Literature Reviews (SLRs) Using AI-Enabled GPT-4 Assistant API 14 Jan 2024 · 0 repositories · arXiv:2402.18582
-
A Novel Multi-Stage Prompting Approach for Language Agnostic MCQ Generation using GPT 13 Jan 2024 · 1 repository · arXiv:2401.07098
-
Assessing Large Language Models in Mechanical Engineering Education: A Study on Mechanics-Focused Conceptual Understanding 13 Jan 2024 · 0 repositories · arXiv:2401.12983
-
Bridging the Preference Gap between Retrievers and LLMs 13 Jan 2024 · 0 repositories · arXiv:2401.06954
-
Combining Confidence Elicitation and Sample-based Methods for Uncertainty Quantification in Misinformation Mitigation 13 Jan 2024 · 0 repositories · arXiv:2401.08694
-
GoMatching: A Simple Baseline for Video Text Spotting via Long and Short Term Matching 13 Jan 2024 · 1 repository · arXiv:2401.07080
-
Knowledge Distillation of Black-Box Large Language Models 13 Jan 2024 · 0 repositories · arXiv:2401.07013
-
Transformer for Object Re-Identification: A Survey 13 Jan 2024 · 1 repository · arXiv:2401.06960Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
A Survey on the Applications of Frontier AI, Foundation Models, and Large Language Models to Intelligent Transportation Systems 12 Jan 2024 · 0 repositories · arXiv:2401.06831
-
Adapting Large Language Models for Document-Level Machine Translation 12 Jan 2024 · 0 repositories · arXiv:2401.06468
-
An investigation of structures responsible for gender bias in BERT and DistilBERT 12 Jan 2024 · 0 repositories · arXiv:2401.06495
-
Comparing GPT-4 and Open-Source Language Models in Misinformation Mitigation 12 Jan 2024 · 0 repositories · arXiv:2401.06920
-
Domain Adaptation for Time series Transformers using One-step fine-tuning 12 Jan 2024 · 0 repositories · arXiv:2401.06524
-
Few-Shot Detection of Machine-Generated Text using Style Representations 12 Jan 2024 · 1 repository · arXiv:2401.06712Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Fine-grained Hallucination Detection and Editing for Language Models 12 Jan 2024 · 0 repositories · arXiv:2401.06855
-
Human-AI Collaborative Essay Scoring: A Dual-Process Framework with LLMs 12 Jan 2024 · 1 repository · arXiv:2401.06431
-
Health-LLM: Large Language Models for Health Prediction via Wearable Sensor Data 12 Jan 2024 · 1 repository · arXiv:2401.06866Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs 12 Jan 2024 · 2 repositories · arXiv:2401.06373Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Hyper-STTN: Social Group-aware Spatial-Temporal Transformer Network for Human Trajectory Prediction with Hypergraph Reasoning 12 Jan 2024 · 0 repositories · arXiv:2401.06344
-
Improved Learned Sparse Retrieval with Corpus-Specific Vocabularies 12 Jan 2024 · 1 repository · arXiv:2401.06703
-
Intention Analysis Makes LLMs A Good Jailbreak Defender 12 Jan 2024 · 1 repository · arXiv:2401.06561Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Mapping Transformer Leveraged Embeddings for Cross-Lingual Document Representation 12 Jan 2024 · 1 repository · arXiv:2401.06583
-
Mission: Impossible Language Models 12 Jan 2024 · 1 repository · arXiv:2401.06416Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples)
-
PersianMind: A Cross-Lingual Persian-English Large Language Model 12 Jan 2024 · 0 repositories · arXiv:2401.06466
-
PizzaCommonSense: Learning to Model Commonsense Reasoning about Intermediate Steps in Cooking Recipes 12 Jan 2024 · 1 repository · arXiv:2401.06930
-
UPDP: A Unified Progressive Depth Pruner for CNN and Vision Transformer 12 Jan 2024 · 0 repositories · arXiv:2401.06426
-
Video Super-Resolution Transformer with Masked Inter&Intra-Frame Attention 12 Jan 2024 · 1 repository · arXiv:2401.06312
-
Analyzing Regional Impacts of Climate Change using Natural Language Processing Techniques 11 Jan 2024 · 0 repositories · arXiv:2401.06817
-
Autocompletion of Chief Complaints in the Electronic Health Records using Large Language Models 11 Jan 2024 · 0 repositories · arXiv:2401.06088
-
Brain Tumor Radiogenomic Classification 11 Jan 2024 · 0 repositories · arXiv:2401.09471
-
Learning Segmented 3D Gaussians via Efficient Feature Unprojection for Zero-shot Neural Scene Segmentation 11 Jan 2024 · 0 repositories · arXiv:2401.05925
-
Efficient Image Deblurring Networks based on Diffusion Models 11 Jan 2024 · 1 repository · arXiv:2401.05907Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 2 violated, 8 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Evidence to Generate (E2G): A Single-agent Two-step Prompting for Context Grounded and Retrieval Augmented Reasoning 11 Jan 2024 · 0 repositories · arXiv:2401.05787
-
Investigating Data Contamination for Pre-training Language Models 11 Jan 2024 · 0 repositories · arXiv:2401.06059
-
Masked Attribute Description Embedding for Cloth-Changing Person Re-identification 11 Jan 2024 · 1 repository · arXiv:2401.05646
-
Mutation-based Consistency Testing for Evaluating the Code Understanding Capability of LLMs 11 Jan 2024 · 0 repositories · arXiv:2401.05940
-
Parrot: Pareto-optimal Multi-Reward Reinforcement Learning Framework for Text-to-Image Generation 11 Jan 2024 · 0 repositories · arXiv:2401.05675
-
Prompt-based mental health screening from social media text 11 Jan 2024 · 0 repositories · arXiv:2401.05912
-
Surface Normal Estimation with Transformers 11 Jan 2024 · 0 repositories · arXiv:2401.05745
-
Surgical-DINO: Adapter Learning of Foundation Models for Depth Estimation in Endoscopic Surgery 11 Jan 2024 · 1 repository · arXiv:2401.06013Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
The Benefits of a Concise Chain of Thought on Problem-Solving in Large Language Models 11 Jan 2024 · 1 repository · arXiv:2401.05618Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Transforming Image Super-Resolution: A ConvFormer-based Efficient Approach 11 Jan 2024 · 1 repository · arXiv:2401.05633
-
Adaptive-avg-pooling based Attention Vision Transformer for Face Anti-spoofing 10 Jan 2024 · 0 repositories · arXiv:2401.04953
-
AdvMT: Adversarial Motion Transformer for Long-term Human Motion Prediction 10 Jan 2024 · 0 repositories · arXiv:2401.05018
-
Deep learning in motion deblurring: current status, benchmarks and future prospects 10 Jan 2024 · 1 repository · arXiv:2401.05055
-
AutoAct: Automatic Agent Learning from Scratch for QA via Self-Planning 10 Jan 2024 · 1 repository · arXiv:2401.05268Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
CADgpt: Harnessing Natural Language Processing for 3D Modelling to Enhance Computer-Aided Design Workflows 10 Jan 2024 · 0 repositories · arXiv:2401.05476
-
Can AI Write Classical Chinese Poetry like Humans? An Empirical Study Inspired by Turing Test 10 Jan 2024 · 0 repositories · arXiv:2401.04952
-
Derm-T2IM: Harnessing Synthetic Skin Lesion Data via Stable Diffusion Models for Enhanced Skin Disease Classification using ViT and CNN 10 Jan 2024 · 0 repositories · arXiv:2401.05159
-
Diffusion-based Pose Refinement and Muti-hypothesis Generation for 3D Human Pose Estimaiton 10 Jan 2024 · 1 repository · arXiv:2401.04921
-
Efficient Fine-Tuning with Domain Adaptation for Privacy-Preserving Vision Transformer 10 Jan 2024 · 0 repositories · arXiv:2401.05126
-
I am a Strange Dataset: Metalinguistic Tests for Language Models 10 Jan 2024 · 1 repository · arXiv:2401.05300
-
InfiAgent-DABench: Evaluating Agents on Data Analysis Tasks 10 Jan 2024 · 1 repository · arXiv:2401.05507Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Knowledge-aware Graph Transformer for Pedestrian Trajectory Prediction 10 Jan 2024 · 0 repositories · arXiv:2401.04872
-
Knowledge Sharing in Manufacturing using Large Language Models: User Evaluation and Model Benchmarking 10 Jan 2024 · 0 repositories · arXiv:2401.05200
-
Leveraging Print Debugging to Improve Code Generation in Large Language Models 10 Jan 2024 · 0 repositories · arXiv:2401.05319
-
Can Active Label Correction Improve LLM-based Modular AI Systems? 10 Jan 2024 · 0 repositories · arXiv:2401.05467
-
Monte Carlo Tree Search for Recipe Generation using GPT-2 10 Jan 2024 · 0 repositories · arXiv:2401.05199
-
Motion Guided Token Compression for Efficient Masked Video Modeling 10 Jan 2024 · 0 repositories · arXiv:2402.18577
-
Reinforcement Learning for Optimizing RAG for Domain Chatbots 10 Jan 2024 · 0 repositories · arXiv:2401.06800
-
SPT: Spectral Transformer for Red Giant Stars Age and Mass Estimation 10 Jan 2024 · 0 repositories · arXiv:2401.04900
-
An Assessment on Comprehending Mental Health through Large Language Models 9 Jan 2024 · 0 repositories · arXiv:2401.04592
-
Arabic Text Diacritization In The Age Of Transfer Learning: Token Classification Is All You Need 9 Jan 2024 · 0 repositories · arXiv:2401.04848
-
DebugBench: Evaluating Debugging Capability of Large Language Models 9 Jan 2024 · 1 repository · arXiv:2401.04621Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples)
-
DedustNet: A Frequency-dominated Swin Transformer-based Wavelet Network for Agricultural Dust Removal 9 Jan 2024 · 0 repositories · arXiv:2401.04750
-
DepressionEmo: A novel dataset for multilabel classification of depression emotions 9 Jan 2024 · 1 repository · arXiv:2401.04655
-
Fighting Fire with Fire: Adversarial Prompting to Generate a Misinformation Detection Dataset 9 Jan 2024 · 0 repositories · arXiv:2401.04481