Methods › General › Attention Mechanisms › Attention › Papers, page 163
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 163 of 316: papers 16,201 to 16,300 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Informed AI Regulation: Comparing the Ethical Frameworks of Leading LLM Chatbots Using an Ethics-Based Audit to Assess Moral Reasoning and Normative Values 9 Jan 2024 · 1 repository · arXiv:2402.01651
-
Iterative Feedback Network for Unsupervised Point Cloud Registration 9 Jan 2024 · 1 repository · arXiv:2401.04357
-
Language Detection for Transliterated Content 9 Jan 2024 · 0 repositories · arXiv:2401.04619
-
Phishing Website Detection through Multi-Model Analysis of HTML Content 9 Jan 2024 · 0 repositories · arXiv:2401.04820
-
Setting the Record Straight on Transformer Oversmoothing 9 Jan 2024 · 0 repositories · arXiv:2401.04301
-
Skin Cancer Segmentation and Classification Using Vision Transformer for Automatic Analysis in Dermatoscopy-based Non-invasive Digital System 9 Jan 2024 · 0 repositories · arXiv:2401.04746
-
T-PRIME: Transformer-based Protocol Identification for Machine-learning at the Edge 9 Jan 2024 · 1 repository · arXiv:2401.04837
-
WaveletFormerNet: A Transformer-based Wavelet Network for Real-world Non-homogeneous and Dense Fog Removal 9 Jan 2024 · 0 repositories · arXiv:2401.04550
-
A Philosophical Introduction to Language Models -- Part I: Continuity With Classic Debates 8 Jan 2024 · 0 repositories · arXiv:2401.03910
-
Advancing Spatial Reasoning in Large Language Models: An In-Depth Evaluation and Enhancement Using the StepGame Benchmark 8 Jan 2024 · 1 repository · arXiv:2401.03991
-
An Exploratory Study on Automatic Identification of Assumptions in the Development of Deep Learning Frameworks 8 Jan 2024 · 1 repository · arXiv:2401.03653
-
Anatomy of Neural Language Models 8 Jan 2024 · 1 repository · arXiv:2401.03797
-
Attention-Guided Erasing: A Novel Augmentation Method for Enhancing Downstream Breast Density Classification 8 Jan 2024 · 0 repositories · arXiv:2401.03912
-
Can Large Language Models Beat Wall Street? Unveiling the Potential of AI in Stock Selection 8 Jan 2024 · 0 repositories · arXiv:2401.03737
-
Distortions in Judged Spatial Relations in Large Language Models 8 Jan 2024 · 0 repositories · arXiv:2401.04218
-
Efficient Multiscale Multimodal Bottleneck Transformer for Audio-Video Classification 8 Jan 2024 · 0 repositories · arXiv:2401.04023
-
Efficient Selective Audio Masked Multimodal Bottleneck Transformer for Audio-Video Classification 8 Jan 2024 · 0 repositories · arXiv:2401.04154
-
GloTSFormer: Global Video Text Spotting Transformer 8 Jan 2024 · 1 repository · arXiv:2401.03694
-
Gramformer: Learning Crowd Counting via Graph-Modulated Transformer 8 Jan 2024 · 1 repository · arXiv:2401.03870Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Advancing bioinformatics with large language models: components, applications and perspectives 8 Jan 2024 · 0 repositories · arXiv:2401.04155
-
LF-ViT: Reducing Spatial Redundancy in Vision Transformer for Efficient Image Recognition 8 Jan 2024 · 1 repository · arXiv:2402.00033Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 7 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
LLM4PLC: Harnessing Large Language Models for Verifiable Programming of PLCs in Industrial Control Systems 8 Jan 2024 · 1 repository · arXiv:2401.05443
-
MARG: Multi-Agent Review Generation for Scientific Papers 8 Jan 2024 · 1 repository · arXiv:2401.04259Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
Mixtral of Experts 8 Jan 2024 · 6 repositories · arXiv:2401.04088Syntology 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 5 harvested samples)
-
MoE-Mamba: Efficient Selective State Space Models with Mixture of Experts 8 Jan 2024 · 1 repository · arXiv:2401.04081
-
MS-DETR: Efficient DETR Training with Mixed Supervision 8 Jan 2024 · 1 repository · arXiv:2401.03989Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 1 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 3 pointer-only (licence)
-
Why Solving Multi-agent Path Finding with Large Language Model has not Succeeded Yet 8 Jan 2024 · 0 repositories · arXiv:2401.03630
-
Can generative AI and ChatGPT outperform humans on cognitive-demanding problem-solving tasks in science? 7 Jan 2024 · 0 repositories · arXiv:2401.15081
-
EAT: Self-Supervised Pre-Training with Efficient Audio Transformer 7 Jan 2024 · 1 repository · arXiv:2401.03497Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 3 pointer-only (licence)
-
Escalation Risks from Language Models in Military and Diplomatic Decision-Making 7 Jan 2024 · 1 repository · arXiv:2401.03408Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
InFoBench: Evaluating Instruction Following Ability in Large Language Models 7 Jan 2024 · 1 repository · arXiv:2401.03601Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
On Leveraging Large Language Models for Enhancing Entity Resolution: A Cost-efficient Approach 7 Jan 2024 · 0 repositories · arXiv:2401.03426
-
PEneo: Unifying Line Extraction, Line Grouping, and Entity Linking for End-to-end Document Pair Extraction 7 Jan 2024 · 1 repository · arXiv:2401.03472
-
RoBERTurk: Adjusting RoBERTa for Turkish 7 Jan 2024 · 0 repositories · arXiv:2401.03515
-
See360: Novel Panoramic View Interpolation 7 Jan 2024 · 1 repository · arXiv:2401.03431
-
The NPU-ASLP-LiAuto System Description for Visual Speech Recognition in CNVSRC 2023 7 Jan 2024 · 2 repositories · arXiv:2401.06788
-
CharPoet: A Chinese Classical Poetry Generation System Based on Token-free LLM 7 Jan 2024 · 0 repositories · arXiv:2401.03512
-
Towards Effective Multiple-in-One Image Restoration: A Sequential and Prompt Learning Strategy 7 Jan 2024 · 2 repositories · arXiv:2401.03379Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Exploring Defeasibility in Causal Reasoning 6 Jan 2024 · 0 repositories · arXiv:2401.03183
-
Multimodal Informative ViT: Information Aggregation and Distribution for Hyperspectral and LiDAR Classification 6 Jan 2024 · 1 repository · arXiv:2401.03179
-
PIXAR: Auto-Regressive Language Modeling in Pixel Space 6 Jan 2024 · 0 repositories · arXiv:2401.03321
-
PosDiffNet: Positional Neural Diffusion for Point Cloud Registration in a Large Field of View with Perturbations 6 Jan 2024 · 0 repositories · arXiv:2401.03167
-
Realism in Action: Anomaly-Aware Diagnosis of Brain Tumors from Medical Images Using YOLOv8 and DeiT 6 Jan 2024 · 0 repositories · arXiv:2401.03302
-
SecureReg: Combining NLP and MLP for Enhanced Detection of Malicious Domain Name Registrations 6 Jan 2024 · 0 repositories · arXiv:2401.03196
-
Using Large Language Models to Assess Tutors' Performance in Reacting to Students Making Math Errors 6 Jan 2024 · 0 repositories · arXiv:2401.03238
-
Vision Transformers and Bi-LSTM for Alzheimer's Disease Diagnosis from 3D MRI 6 Jan 2024 · 0 repositories · arXiv:2401.03132
-
A Cost-Efficient FPGA Implementation of Tiny Transformer Model using Neural ODE 5 Jan 2024 · 0 repositories · arXiv:2401.02721
-
A Random Ensemble of Encrypted models for Enhancing Robustness against Adversarial Examples 5 Jan 2024 · 0 repositories · arXiv:2401.02633
-
AST-T5: Structure-Aware Pretraining for Code Generation and Understanding 5 Jan 2024 · 1 repository · arXiv:2401.03003Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Can Large Language Models Understand Molecules? 5 Jan 2024 · 2 repositories · arXiv:2402.00024
-
CRUXEval: A Benchmark for Code Reasoning, Understanding and Execution 5 Jan 2024 · 1 repository · arXiv:2401.03065
-
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism 5 Jan 2024 · 1 repository · arXiv:2401.02954
-
Denoising Vision Transformers 5 Jan 2024 · 1 repository · arXiv:2401.02957Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
From LLM to Conversational Agent: A Memory Enhanced Architecture with Fine-Tuning of Large Language Models 5 Jan 2024 · 0 repositories · arXiv:2401.02777
-
Natural Language Programming in Medicine: Administering Evidence Based Clinical Workflows with Autonomous Agents Powered by Generative Large Language Models 5 Jan 2024 · 0 repositories · arXiv:2401.02851
-
Geometric-Facilitated Denoising Diffusion Model for 3D Molecule Generation 5 Jan 2024 · 1 repository · arXiv:2401.02683Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 1 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples)
-
German Text Embedding Clustering Benchmark 5 Jan 2024 · 1 repository · arXiv:2401.02709
-
Latte: Latent Diffusion Transformer for Video Generation 5 Jan 2024 · 4 repositories · arXiv:2401.03048Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples)
-
Parameter-Efficient Sparsity Crafting from Dense to Mixture-of-Experts for Instruction Tuning on General Tasks 5 Jan 2024 · 2 repositories · arXiv:2401.02731
-
PeFoMed: Parameter Efficient Fine-tuning of Multimodal Large Language Models for Medical Imaging 5 Jan 2024 · 1 repository · arXiv:2401.02797
-
Prompt-driven Latent Domain Generalization for Medical Image Classification 5 Jan 2024 · 2 repositories · arXiv:2401.03002
-
SPFormer: Enhancing Vision Transformer with Superpixel Representation 5 Jan 2024 · 0 repositories · arXiv:2401.02931
-
A novel method to enhance pneumonia detection via a model-level ensembling of CNN and vision transformer 4 Jan 2024 · 0 repositories · arXiv:2401.02358
-
Are LLMs Robust for Spoken Dialogues? 4 Jan 2024 · 0 repositories · arXiv:2401.02297
-
Beyond Extraction: Contextualising Tabular Data for Efficient Summarisation by Language Models 4 Jan 2024 · 0 repositories · arXiv:2401.02333
-
Blar-SQL: Faster, Stronger, Smaller NL2SQL 4 Jan 2024 · 0 repositories · arXiv:2401.02997
-
ClassWise-SAM-Adapter: Parameter Efficient Fine-tuning Adapts Segment Anything to SAR Domain for Semantic Segmentation 4 Jan 2024 · 1 repository · arXiv:2401.02326
-
Exploring Boundary of GPT-4V on Marine Analysis: A Preliminary Case Study 4 Jan 2024 · 0 repositories · arXiv:2401.02147
-
GridFormer: Point-Grid Transformer for Surface Reconstruction 4 Jan 2024 · 1 repository · arXiv:2401.02292Syntology official (archive's flag): 9 ran · 9 ran (of which 6 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
L3Cube-IndicNews: News-based Short Text and Long Document Classification Datasets in Indic Languages 4 Jan 2024 · 1 repository · arXiv:2401.02254
-
LLaMA Pro: Progressive LLaMA with Block Expansion 4 Jan 2024 · 1 repository · arXiv:2401.02415
-
Re-evaluating the Memory-balanced Pipeline Parallelism: BPipe 4 Jan 2024 · 0 repositories · arXiv:2401.02088
-
Shayona@SMM4H23: COVID-19 Self diagnosis classification using BERT and LightGBM models 4 Jan 2024 · 0 repositories · arXiv:2401.02158
-
Spikformer V2: Join the High Accuracy Club on ImageNet with an SNN Ticket 4 Jan 2024 · 3 repositories · arXiv:2401.02020Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Text2MDT: Extracting Medical Decision Trees from Medical Texts 4 Jan 2024 · 1 repository · arXiv:2401.02034
-
TR-DETR: Task-Reciprocal Transformer for Joint Moment Retrieval and Highlight Detection 4 Jan 2024 · 1 repository · arXiv:2401.02309Syntology official (archive's flag): 13 ran · 13 ran (of which 5 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 1 violated, 10 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Studying and Recommending Information Highlighting in Stack Overflow Answers 3 Jan 2024 · 1 repository · arXiv:2401.01472
-
AIGCBench: Comprehensive Evaluation of Image-to-Video Content Generated by AI 3 Jan 2024 · 2 repositories · arXiv:2401.01651Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
AstroLLaMA-Chat: Scaling AstroLLaMA with Conversational and Diverse Datasets 3 Jan 2024 · 0 repositories · arXiv:2401.01916
-
Context-Guided Spatio-Temporal Video Grounding 3 Jan 2024 · 2 repositories · arXiv:2401.01578Syntology official (archive's flag): 15 ran · 22 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 1 violated, 16 with no contract checked; 5 where Syntology's instrument failed) · 12 unverified (of 34 harvested samples) · 34 pointer-only (licence)
-
CRA-PCN: Point Cloud Completion with Intra- and Inter-level Cross-Resolution Transformers 3 Jan 2024 · 1 repository · arXiv:2401.01552Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 5 where Syntology's instrument failed) · 16 unverified (of 33 harvested samples) · 33 pointer-only (licence)
-
Enhancing Multilingual Information Retrieval in Mixed Human Resources Environments: A RAG Model Implementation for Multicultural Enterprise 3 Jan 2024 · 0 repositories · arXiv:2401.01511
-
FullLoRA-AT: Efficiently Boosting the Robustness of Pretrained Vision Transformers 3 Jan 2024 · 0 repositories · arXiv:2401.01752
-
GPT-4V(ision) is a Generalist Web Agent, if Grounded 3 Jan 2024 · 1 repository · arXiv:2401.01614Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Incremental FastPitch: Chunk-based High Quality Text to Speech 3 Jan 2024 · 0 repositories · arXiv:2401.01755
-
The Internet of Things in the Era of Generative AI: Vision and Challenges 3 Jan 2024 · 0 repositories · arXiv:2401.01923
-
Iterative Mask Filling: An Effective Text Augmentation Method Using Masked Language Modeling 3 Jan 2024 · 0 repositories · arXiv:2401.01830
-
Kernel-U-Net: Multivariate Time Series Forecasting using Custom Kernels 3 Jan 2024 · 0 repositories · arXiv:2401.01479
-
Large Language Model Capabilities in Perioperative Risk Prediction and Prognostication 3 Jan 2024 · 1 repository · arXiv:2401.01620
-
MLPs Compass: What is learned when MLPs are combined with PLMs? 3 Jan 2024 · 0 repositories · arXiv:2401.01667
-
Natural Language Processing and Multimodal Stock Price Prediction 3 Jan 2024 · 0 repositories · arXiv:2401.01487
-
Revisiting Zero-Shot Abstractive Summarization in the Era of Large Language Models from the Perspective of Position Bias 3 Jan 2024 · 1 repository · arXiv:2401.01989
-
Sports-QA: A Large-Scale Video Question Answering Benchmark for Complex and Professional Sports 3 Jan 2024 · 1 repository · arXiv:2401.01505
-
Team IELAB at TREC Clinical Trial Track 2023: Enhancing Clinical Trial Retrieval with Neural Rankers and Large Language Models 3 Jan 2024 · 0 repositories · arXiv:2401.01566
-
TPC-ViT: Token Propagation Controller for Efficient Vision Transformer 3 Jan 2024 · 0 repositories · arXiv:2401.01470
-
Towards Robust Semantic Segmentation against Patch-based Attack via Attention Refinement 3 Jan 2024 · 0 repositories · arXiv:2401.01750
-
Transformer Neural Autoregressive Flows 3 Jan 2024 · 0 repositories · arXiv:2401.01855
-
Transformer RGBT Tracking with Spatio-Temporal Multimodal Tokens 3 Jan 2024 · 0 repositories · arXiv:2401.01674
-
A Novel Transformer-Based Self-Supervised Learning Method to Enhance Photoplethysmogram Signal Artifact Detection 2 Jan 2024 · 0 repositories · arXiv:2401.01013
-
CharacterEval: A Chinese Benchmark for Role-Playing Conversational Agent Evaluation 2 Jan 2024 · 1 repository · arXiv:2401.01275Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)