Methods › General › Attention Mechanisms › Attention › Papers, page 72
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 72 of 316: papers 7,101 to 7,200 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Unraveling Arithmetic in Large Language Models: The Role of Algebraic Structures 25 Nov 2024 · 0 repositories · arXiv:2411.16260
-
VICON: Vision In-Context Operator Networks for Multi-Physics Fluid Dynamics Prediction 25 Nov 2024 · 1 repository · arXiv:2411.16063
-
VIRES: Video Instance Repainting via Sketch and Text Guided Generation 25 Nov 2024 · 0 repositories · arXiv:2411.16199
-
VQ-SGen: A Vector Quantized Stroke Representation for Creative Sketch Generation 25 Nov 2024 · 0 repositories · arXiv:2411.16446
-
WTDUN: Wavelet Tree-Structured Sampling and Deep Unfolding Network for Image Compressed Sensing 25 Nov 2024 · 0 repositories · arXiv:2411.16336
-
A General Sensing-assisted Channel Estimation Framework in Distributed MIMO Network 24 Nov 2024 · 0 repositories · arXiv:2411.15995
-
A Method for Building Large Language Models with Predefined KV Cache Capacity 24 Nov 2024 · 0 repositories · arXiv:2411.15785
-
Beyond adaptive gradient: Fast-Controlled Minibatch Algorithm for large-scale optimization 24 Nov 2024 · 1 repository · arXiv:2411.15795
-
Development of Pre-Trained Transformer-based Models for the Nepali Language 24 Nov 2024 · 0 repositories · arXiv:2411.15734
-
FastTrackTr:Towards Fast Multi-Object Tracking with Transformers 24 Nov 2024 · 0 repositories · arXiv:2411.15811
-
Fixing the Perspective: A Critical Examination of Zero-1-to-3 24 Nov 2024 · 0 repositories · arXiv:2411.15706
-
Gradient Norm Regularization Second-Order Algorithms for Solving Nonconvex-Strongly Concave Minimax Problems 24 Nov 2024 · 0 repositories · arXiv:2411.15769
-
Investigating Factuality in Long-Form Text Generation: The Roles of Self-Known and Self-Unknown 24 Nov 2024 · 0 repositories · arXiv:2411.15993
-
LetsTalk: Latent Diffusion Transformer for Talking Video Synthesis 24 Nov 2024 · 0 repositories · arXiv:2411.16748
-
LLaMA-MoE v2: Exploring Sparsity of LLaMA from Perspective of Mixture-of-Experts with Post-Training 24 Nov 2024 · 1 repository · arXiv:2411.15708Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
LTCF-Net: A Transformer-Enhanced Dual-Channel Fourier Framework for Low-Light Image Restoration 24 Nov 2024 · 0 repositories · arXiv:2411.15740
-
Medical Slice Transformer: Improved Diagnosis and Explainability on 3D Medical Images with DINOv2 24 Nov 2024 · 1 repository · arXiv:2411.15802
-
Nimbus: Secure and Efficient Two-Party Inference for Transformers 24 Nov 2024 · 1 repository · arXiv:2411.15707Syntology official: harvested, nothing ran · 0 ran · 5 unverified (of 5 harvested samples)
-
PR-MIM: Delving Deeper into Partial Reconstruction in Masked Image Modeling 24 Nov 2024 · 0 repositories · arXiv:2411.15746
-
RAMIE: Retrieval-Augmented Multi-task Information Extraction with Large Language Models on Dietary Supplements 24 Nov 2024 · 0 repositories · arXiv:2411.15700
-
ResCLIP: Residual Attention for Training-free Dense Vision-language Inference 24 Nov 2024 · 1 repository · arXiv:2411.15851
-
Self-Calibrated CLIP for Training-Free Open-Vocabulary Segmentation 24 Nov 2024 · 1 repository · arXiv:2411.15869Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
Test-time Alignment-Enhanced Adapter for Vision-Language Models 24 Nov 2024 · 1 repository · arXiv:2411.15735
-
Text-Guided Coarse-to-Fine Fusion Network for Robust Remote Sensing Visual Question Answering 24 Nov 2024 · 0 repositories · arXiv:2411.15770
-
TransFair: Transferring Fairness from Ocular Disease Classification to Progression Prediction 24 Nov 2024 · 0 repositories · arXiv:2412.00051
-
A Comparative Analysis of Transformer and LSTM Models for Detecting Suicidal Ideation on Reddit 23 Nov 2024 · 1 repository · arXiv:2411.15404
-
"All that Glitters": Approaches to Evaluations with Unreliable Model and Human Annotations 23 Nov 2024 · 1 repository · arXiv:2411.15634
-
Best of Both Worlds: Advantages of Hybrid Graph Sequence Models 23 Nov 2024 · 0 repositories · arXiv:2411.15671
-
ChatBCI: A P300 Speller BCI Leveraging Large Language Models for Improved Sentence Composition in Realistic Scenarios 23 Nov 2024 · 0 repositories · arXiv:2411.15395
-
Circuit design in biology and machine learning. II. Anomaly detection 23 Nov 2024 · 0 repositories · arXiv:2411.15647
-
Devils in Middle Layers of Large Vision-Language Models: Interpreting, Detecting and Mitigating Object Hallucinations via Attention Lens 23 Nov 2024 · 1 repository · arXiv:2411.16724Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Enhancing Instruction-Following Capability of Visual-Language Models by Reducing Image Redundancy 23 Nov 2024 · 0 repositories · arXiv:2411.15453
-
Federated Learning in Chemical Engineering: A Tutorial on a Framework for Privacy-Preserving Collaboration Across Distributed Data Sources 23 Nov 2024 · 1 repository · arXiv:2411.16737
-
Federated PCA and Estimation for Spiked Covariance Matrices: Optimal Rates and Efficient Algorithm 23 Nov 2024 · 0 repositories · arXiv:2411.15660
-
FG-CXR: A Radiologist-Aligned Gaze Dataset for Enhancing Interpretability in Chest X-Ray Report Generation 23 Nov 2024 · 1 repository · arXiv:2411.15413Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples)
-
freePruner: A Training-free Approach for Large Multimodal Model Acceleration 23 Nov 2024 · 0 repositories · arXiv:2411.15446
-
GeoAI-Enhanced Community Detection on Spatial Networks with Graph Deep Learning 23 Nov 2024 · 1 repository · arXiv:2411.15428
-
Improving Next Tokens via Second-Last Predictions with Generate and Refine 23 Nov 2024 · 0 repositories · arXiv:2411.15661
-
Inducing Human-like Biases in Moral Reasoning Language Models 23 Nov 2024 · 0 repositories · arXiv:2411.15386
-
Large-Scale Text-to-Image Model with Inpainting is a Zero-Shot Subject-Driven Image Generator 23 Nov 2024 · 1 repository · arXiv:2411.15466
-
Semantic Shield: Defending Vision-Language Models Against Backdooring and Poisoning via Fine-grained Knowledge Alignment 23 Nov 2024 · 1 repository · arXiv:2411.15673
-
Scaling Structure Aware Virtual Screening to Billions of Molecules with SPRINT 23 Nov 2024 · 1 repository · arXiv:2411.15418
-
TANGNN: a Concise, Scalable and Effective Graph Neural Networks with Top-m Attention Mechanism for Graph Representation Learning 23 Nov 2024 · 1 repository · arXiv:2411.15458
-
Towards Satellite Image Road Graph Extraction: A Global-Scale Dataset and A Novel Method 23 Nov 2024 · 1 repository · arXiv:2411.16733
-
Traditional Chinese Medicine Case Analysis System for High-Level Semantic Abstraction: Optimized with Prompt and RAG 23 Nov 2024 · 0 repositories · arXiv:2411.15491
-
Training an Open-Vocabulary Monocular 3D Object Detection Model without 3D Data 23 Nov 2024 · 0 repositories · arXiv:2411.15657
-
A Real-Time DETR Approach to Bangladesh Road Object Detection for Autonomous Vehicles 22 Nov 2024 · 0 repositories · arXiv:2411.15110
-
An Attention-based Framework for Fair Contrastive Learning 22 Nov 2024 · 0 repositories · arXiv:2411.14765
-
Astro-HEP-BERT: A bidirectional language model for studying the meanings of concepts in astrophysics and high energy physics 22 Nov 2024 · 0 repositories · arXiv:2411.14877
-
Boundless Across Domains: A New Paradigm of Adaptive Feature and Cross-Attention for Domain Generalization in Medical Image Segmentation 22 Nov 2024 · 0 repositories · arXiv:2411.14883
-
Comparative Analysis of Pooling Mechanisms in LLMs: A Sentiment Analysis Perspective 22 Nov 2024 · 0 repositories · arXiv:2411.14654
-
Learning Modality-Aware Representations: Adaptive Group-wise Interaction Network for Multimodal MRI Synthesis 22 Nov 2024 · 1 repository · arXiv:2411.14684
-
Cross-Modal Pre-Aligned Method with Global and Local Information for Remote-Sensing Image and Text Retrieval 22 Nov 2024 · 0 repositories · arXiv:2411.14704
-
Defective Edge Detection Using Cascaded Ensemble Canny Operator 22 Nov 2024 · 0 repositories · arXiv:2411.14868
-
Detecting Visual Triggers in Cannabis Imagery: A CLIP-Based Multi-Labeling Framework with Local-Global Aggregation 22 Nov 2024 · 0 repositories · arXiv:2412.08648
-
Don't Mesh with Me: Generating Constructive Solid Geometry Instead of Meshes by Fine-Tuning a Code-Generation LLM 22 Nov 2024 · 0 repositories · arXiv:2411.15279
-
EfficientViM: Efficient Vision Mamba with Hidden State Mixer based State Space Duality 22 Nov 2024 · 2 repositories · arXiv:2411.15241
-
ElastiFormer: Learned Redundancy Reduction in Transformer via Self-Distillation 22 Nov 2024 · 0 repositories · arXiv:2411.15281
-
Evaluating Vision Transformer Models for Visual Quality Control in Industrial Manufacturing 22 Nov 2024 · 1 repository · arXiv:2411.14953
-
Exploring the Use of Machine Learning Weather Models in Data Assimilation 22 Nov 2024 · 0 repositories · arXiv:2411.14677
-
Fast High-Quality Enhanced Imaging Algorithm for Layered Dielectric Targets Based on MMW MIMO-SAR System 22 Nov 2024 · 0 repositories · arXiv:2411.14837
-
HeadRouter: A Training-free Image Editing Framework for MM-DiTs by Adaptively Routing Attention Heads 22 Nov 2024 · 0 repositories · arXiv:2411.15034
-
ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models 22 Nov 2024 · 0 repositories · arXiv:2411.15268
-
AI Foundation Models for Wearable Movement Data in Mental Health Research 22 Nov 2024 · 1 repository · arXiv:2411.15240
-
J-Invariant Volume Shuffle for Self-Supervised Cryo-Electron Tomogram Denoising on Single Noisy Volume 22 Nov 2024 · 0 repositories · arXiv:2411.15248
-
KBAlign: Efficient Self Adaptation on Specific Knowledge Bases 22 Nov 2024 · 1 repository · arXiv:2411.14790
-
MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs 22 Nov 2024 · 4 repositories · arXiv:2411.15296
-
Multi-granularity Interest Retrieval and Refinement Network for Long-Term User Behavior Modeling in CTR Prediction 22 Nov 2024 · 3 repositories · arXiv:2411.15005
-
Multiset Transformer: Advancing Representation Learning in Persistence Diagrams 22 Nov 2024 · 1 repository · arXiv:2411.14662
-
OminiControl: Minimal and Universal Control for Diffusion Transformer 22 Nov 2024 · 2 repositories · arXiv:2411.15098Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Point Cloud Understanding via Attention-Driven Contrastive Learning 22 Nov 2024 · 0 repositories · arXiv:2411.14744
-
Purrfessor: A Fine-tuned Multimodal LLaVA Diet Health Chatbot 22 Nov 2024 · 0 repositories · arXiv:2411.14925
-
Recursive Gaussian Process State Space Model 22 Nov 2024 · 2 repositories · arXiv:2411.14679
-
RED: Effective Trajectory Representation Learning with Comprehensive Information 22 Nov 2024 · 0 repositories · arXiv:2411.15096
-
Resolution-Agnostic Transformer-based Climate Downscaling 22 Nov 2024 · 0 repositories · arXiv:2411.14774
-
SafeLight: Enhancing Security in Optical Convolutional Neural Network Accelerators 22 Nov 2024 · 0 repositories · arXiv:2411.16712
-
ScribeAgent: Towards Specialized Web Agents Using Production-Scale Workflow Data 22 Nov 2024 · 1 repository · arXiv:2411.15004Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Simplifying CLIP: Unleashing the Power of Large-Scale Models on Consumer-level Computers 22 Nov 2024 · 0 repositories · arXiv:2411.14789
-
TEXGen: a Generative Diffusion Model for Mesh Textures 22 Nov 2024 · 1 repository · arXiv:2411.14740
-
Transforming NLU with Babylon: A Case Study in Development of Real-time, Edge-Efficient, Multi-Intent Translation System for Automated Drive-Thru Ordering 22 Nov 2024 · 0 repositories · arXiv:2411.15372
-
When Spatial meets Temporal in Action Recognition 22 Nov 2024 · 0 repositories · arXiv:2411.15284
-
An accuracy improving method for advertising click through rate prediction based on enhanced xDeepFM model 21 Nov 2024 · 0 repositories · arXiv:2411.15223
-
An Experimental Study on Data Augmentation Techniques for Named Entity Recognition on Low-Resource Domains 21 Nov 2024 · 0 repositories · arXiv:2411.14551
-
Assessment of LLM Responses to End-user Security Questions 21 Nov 2024 · 0 repositories · arXiv:2411.14571
-
Benchmarking GPT-4 against Human Translators: A Comprehensive Evaluation Across Languages, Domains, and Expertise Levels 21 Nov 2024 · 1 repository · arXiv:2411.13775Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
BERT-Based Approach for Automating Course Articulation Matrix Construction with Explainable AI 21 Nov 2024 · 1 repository · arXiv:2411.14254
-
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation 21 Nov 2024 · 1 repository · arXiv:2411.13836
-
Contrasting local and global modeling with machine learning and satellite data: A case study estimating tree canopy height in African savannas 21 Nov 2024 · 0 repositories · arXiv:2411.14354
-
DINO-X: A Unified Vision Model for Open-World Object Detection and Understanding 21 Nov 2024 · 1 repository · arXiv:2411.14347
-
Do I Know This Entity? Knowledge Awareness and Hallucinations in Language Models 21 Nov 2024 · 0 repositories · arXiv:2411.14257
-
Evaluating the Robustness of Analogical Reasoning in Large Language Models 21 Nov 2024 · 1 repository · arXiv:2411.14215Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Explaining GPT-4's Schema of Depression Using Machine Behavior Analysis 21 Nov 2024 · 0 repositories · arXiv:2411.13800
-
Exploring applications of topological data analysis in stock index movement prediction 21 Nov 2024 · 1 repository · arXiv:2411.13881
-
FastRAG: Retrieval Augmented Generation for Semi-structured Data 21 Nov 2024 · 0 repositories · arXiv:2411.13773
-
G-RAG: Knowledge Expansion in Material Science 21 Nov 2024 · 1 repository · arXiv:2411.14592
-
Generative Fuzzy System for Sequence Generation 21 Nov 2024 · 0 repositories · arXiv:2411.13867
-
Global and Local Attention-Based Transformer for Hyperspectral Image Change Detection 21 Nov 2024 · 1 repository · arXiv:2411.14109
-
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI 21 Nov 2024 · 1 repository · arXiv:2411.14522Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
Graph Domain Adaptation with Dual-branch Encoder and Two-level Alignment for Whole Slide Image-based Survival Prediction 21 Nov 2024 · 0 repositories · arXiv:2411.14001
-
Learning from "Silly" Questions Improves Large Language Models, But Only Slightly 21 Nov 2024 · 0 repositories · arXiv:2411.14121