Methods › General › Attention Mechanisms › Attention › Papers, page 84
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 84 of 316: papers 8,301 to 8,400 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Towards a Reliable Offline Personal AI Assistant for Long Duration Spaceflight 21 Oct 2024 · 0 repositories · arXiv:2410.16397
-
Using GPT Models for Qualitative and Quantitative News Analytics in the 2024 US Presidental Election Process 21 Oct 2024 · 0 repositories · arXiv:2410.15884
-
ViMoE: An Empirical Study of Designing Vision Mixture-of-Experts 21 Oct 2024 · 0 repositories · arXiv:2410.15732
-
Weighted Diversified Sampling for Efficient Data-Driven Single-Cell Gene-Gene Interaction Discovery 21 Oct 2024 · 0 repositories · arXiv:2410.15616
-
Who's Who: Large Language Models Meet Knowledge Conflicts in Practice 21 Oct 2024 · 1 repository · arXiv:2410.15737
-
YOLO11 and Vision Transformers based 3D Pose Estimation of Immature Green Fruits in Commercial Apple Orchards for Robotic Thinning 21 Oct 2024 · 0 repositories · arXiv:2410.19846
-
A case study of social media and its perceived effects to student in academic performance. 20 Oct 2024 · 0 repositories
-
A case study of social media and its perceived effects to student in academic performance 20 Oct 2024 · 0 repositories
-
A Heterogeneous Network-based Contrastive Learning Approach for Predicting Drug-Target Interaction 20 Oct 2024 · 1 repository · arXiv:2411.00801
-
Advancing Gasoline Consumption Forecasting: A Novel Hybrid Model Integrating Transformers, LSTM, and CNN 20 Oct 2024 · 0 repositories · arXiv:2410.16336
-
AttCDCNet: Attention-enhanced Chest Disease Classification using X-Ray Images 20 Oct 2024 · 0 repositories · arXiv:2410.15437
-
Back to School: Translation Using Grammar Books 20 Oct 2024 · 1 repository · arXiv:2410.15263Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
BRIEF: Bridging Retrieval and Inference for Multi-hop Reasoning via Compression 20 Oct 2024 · 1 repository · arXiv:2410.15277
-
Comparative Analysis of LSTM, GRU, and Transformer Models for Stock Price Prediction 20 Oct 2024 · 0 repositories · arXiv:2411.05790
-
Concept Complement Bottleneck Model for Interpretable Medical Image Diagnosis 20 Oct 2024 · 0 repositories · arXiv:2410.15446
-
ContextDet: Temporal Action Detection with Adaptive Context Aggregation 20 Oct 2024 · 0 repositories · arXiv:2410.15279
-
Contextual Augmented Multi-Model Programming (CAMP): A Hybrid Local-Cloud Copilot Framework 20 Oct 2024 · 1 repository · arXiv:2410.15285
-
ConTReGen: Context-driven Tree-structured Retrieval for Open-domain Long-form Text Generation 20 Oct 2024 · 0 repositories · arXiv:2410.15511
-
Data Augmentation via Diffusion Model to Enhance AI Fairness 20 Oct 2024 · 0 repositories · arXiv:2410.15470
-
Do RAG Systems Cover What Matters? Evaluating and Optimizing Responses with Sub-Question Coverage 20 Oct 2024 · 1 repository · arXiv:2410.15531
-
Does ChatGPT Have a Poetic Style? 20 Oct 2024 · 1 repository · arXiv:2410.15299
-
Evaluating Consistencies in LLM responses through a Semantic Clustering of Question Answering 20 Oct 2024 · 0 repositories · arXiv:2410.15440
-
Exploring Social Desirability Response Bias in Large Language Models: Evidence from GPT-4 Simulations 20 Oct 2024 · 0 repositories · arXiv:2410.15442
-
Fractional-order spike-timing-dependent gradient descent for multi-layer spiking neural networks 20 Oct 2024 · 0 repositories · arXiv:2410.15293
-
FrameBridge: Improving Image-to-Video Generation with Bridge Models 20 Oct 2024 · 0 repositories · arXiv:2410.15371
-
GSSF: Generalized Structural Sparse Function for Deep Cross-modal Metric Learning 20 Oct 2024 · 1 repository · arXiv:2410.15266
-
Heterogeneous Graph Reinforcement Learning for Dependency-aware Multi-task Allocation in Spatial Crowdsourcing 20 Oct 2024 · 0 repositories · arXiv:2410.15449
-
IKDP: Inverse Kinematics through Diffusion Process 20 Oct 2024 · 0 repositories · arXiv:2410.15341
-
Interweaving Insights: High-Order Feature Interaction for Fine-Grained Visual Recognition 20 Oct 2024 · 1 repository
-
LlamaLens: Specialized Multilingual LLM for Analyzing News and Social Media Content 20 Oct 2024 · 0 repositories · arXiv:2410.15308
-
Lossless KV Cache Compression to 2% 20 Oct 2024 · 0 repositories · arXiv:2410.15252
-
LTPNet Integration of Deep Learning and Environmental Decision Support Systems for Renewable Energy Demand Forecasting 20 Oct 2024 · 0 repositories · arXiv:2410.15286
-
MMDS: A Multimodal Medical Diagnosis System Integrating Image Analysis and Knowledge-based Departmental Consultation 20 Oct 2024 · 0 repositories · arXiv:2410.15403
-
Multi-Layer Feature Fusion with Cross-Channel Attention-Based U-Net for Kidney Tumor Segmentation 20 Oct 2024 · 0 repositories · arXiv:2410.15472
-
SDP4Bit: Toward 4-bit Communication Quantization in Sharded Data Parallelism for LLM Training 20 Oct 2024 · 0 repositories · arXiv:2410.15526
-
SEA: State-Exchange Attention for High-Fidelity Physics Based Transformers 20 Oct 2024 · 1 repository · arXiv:2410.15495Syntology official (archive's flag): 11 ran · 11 ran (of which 7 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
TAGExplainer: Narrating Graph Explanations for Text-Attributed Graph Learning Models 20 Oct 2024 · 0 repositories · arXiv:2410.15268
-
TrackMe:A Simple and Effective Multiple Object Tracking Annotation Tool 20 Oct 2024 · 1 repository · arXiv:2410.15518
-
Training Language Models to Critique With Multi-agent Feedback 20 Oct 2024 · 0 repositories · arXiv:2410.15287
-
Unveiling and Consulting Core Experts in Retrieval-Augmented MoE-based LLMs 20 Oct 2024 · 0 repositories · arXiv:2410.15438
-
When Machine Unlearning Meets Retrieval-Augmented Generation (RAG): Keep Secret or Forget Knowledge? 20 Oct 2024 · 0 repositories · arXiv:2410.15267
-
A comparative study of NeuralODE and Universal ODE approaches to solving Chandrasekhar White Dwarf equation 19 Oct 2024 · 0 repositories · arXiv:2410.14998
-
Accelerate Coastal Ocean Circulation Model with AI Surrogate 19 Oct 2024 · 0 repositories · arXiv:2410.14952
-
Bias Amplification: Language Models as Increasingly Biased Media 19 Oct 2024 · 0 repositories · arXiv:2410.15234
-
Crafting Tomorrow: The Influence of Design Choices on Fresh Content in Social Media Recommendation 19 Oct 2024 · 0 repositories · arXiv:2410.15174
-
EPT-1.5 Technical Report 19 Oct 2024 · 0 repositories · arXiv:2410.15076
-
Evaluation Of P300 Speller Performance Using Large Language Models Along With Cross-Subject Training 19 Oct 2024 · 1 repository · arXiv:2410.15161
-
EViT-Unet: U-Net Like Efficient Vision Transformer for Medical Image Segmentation on Mobile and Edge Devices 19 Oct 2024 · 1 repository · arXiv:2410.15036
-
Incorporating Group Prior into Variational Inference for Tail-User Behavior Modeling in CTR Prediction 19 Oct 2024 · 0 repositories · arXiv:2410.15098
-
Independent Feature Enhanced Crossmodal Fusion for Match-Mismatch Classification of Speech Stimulus and EEG Response 19 Oct 2024 · 0 repositories · arXiv:2410.15078
-
LLaVA-Ultra: Large Chinese Language and Vision Assistant for Ultrasound 19 Oct 2024 · 0 repositories · arXiv:2410.15074
-
LSS-SKAN: Efficient Kolmogorov-Arnold Networks based on Single-Parameterized Function 19 Oct 2024 · 2 repositories · arXiv:2410.14951
-
MCCoder: Streamlining Motion Control with LLM-Assisted Code Generation and Rigorous Verification 19 Oct 2024 · 1 repository · arXiv:2410.15154
-
Medical-GAT: Cancer Document Classification Leveraging Graph-Based Residual Network for Scenarios with Limited Data 19 Oct 2024 · 0 repositories · arXiv:2410.15198
-
Personalized Federated Learning with Adaptive Feature Aggregation and Knowledge Transfer 19 Oct 2024 · 0 repositories · arXiv:2410.15073
-
SemiHVision: Enhancing Medical Multimodal Models with a Semi-Human Annotated Dataset and Fine-Tuned Instruction Generation 19 Oct 2024 · 1 repository · arXiv:2410.14948
-
Spatial-Mamba: Effective Visual State Space Models via Structure-Aware State Fusion 19 Oct 2024 · 1 repository · arXiv:2410.15091Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 14 harvested samples) · 2 pointer-only (licence)
-
Toward Robust RALMs: Revealing the Impact of Imperfect Retrieval on Retrieval-Augmented Language Models 19 Oct 2024 · 1 repository · arXiv:2410.15107
-
Visual Navigation of Digital Libraries: Retrieval and Classification of Images in the National Library of Norway's Digitised Book Collection 19 Oct 2024 · 1 repository · arXiv:2410.14969
-
A novel approach towards the classification of Bone Fracture from Musculoskeletal Radiography images using Attention Based Transfer Learning 18 Oct 2024 · 0 repositories · arXiv:2410.14833
-
Aligning AI Agents via Information-Directed Sampling 18 Oct 2024 · 0 repositories · arXiv:2410.14807
-
Attention-Guided Residual U-Net with SE Connection and ASPP for Watershed-Based Cell Segmentation in Microscopy Images 18 Oct 2024 · 1 repository
-
Auto Detecting Cognitive Events Using Machine Learning on Pupillary Data 18 Oct 2024 · 0 repositories · arXiv:2410.14174
-
Automated Genre-Aware Article Scoring and Feedback Using Large Language Models 18 Oct 2024 · 0 repositories · arXiv:2410.14165
-
Backdoored Retrievers for Prompt Injection Attacks on Retrieval Augmented Generation of Large Language Models 18 Oct 2024 · 0 repositories · arXiv:2410.14479
-
CausalChat: Interactive Causal Model Development and Refinement Using Large Language Models 18 Oct 2024 · 0 repositories · arXiv:2410.14146
-
CELI: Controller-Embedded Language Model Interactions 18 Oct 2024 · 0 repositories · arXiv:2410.14627
-
Croc: Pretraining Large Multimodal Models with Cross-Modal Comprehension 18 Oct 2024 · 1 repository · arXiv:2410.14332
-
DFlow: Diverse Dialogue Flow Simulation with Large Language Models 18 Oct 2024 · 0 repositories · arXiv:2410.14853
-
Effects of Soft-Domain Transfer and Named Entity Information on Deception Detection 18 Oct 2024 · 0 repositories · arXiv:2410.14814
-
FashionR2R: Texture-preserving Rendered-to-Real Image Translation with Diffusion Models 18 Oct 2024 · 0 repositories · arXiv:2410.14429
-
Feint and Attack: Attention-Based Strategies for Jailbreaking and Protecting LLMs 18 Oct 2024 · 0 repositories · arXiv:2410.16327
-
Flame quality monitoring of flare stack based on deep visual features 18 Oct 2024 · 0 repositories · arXiv:2410.19823
-
GESH-Net: Graph-Enhanced Spherical Harmonic Convolutional Networks for Cortical Surface Registration 18 Oct 2024 · 0 repositories · arXiv:2410.14805
-
Good Parenting is all you need -- Multi-agentic LLM Hallucination Mitigation 18 Oct 2024 · 0 repositories · arXiv:2410.14262
-
Graph Contrastive Learning via Cluster-refined Negative Sampling for Semi-supervised Text Classification 18 Oct 2024 · 0 repositories · arXiv:2410.18130
-
Implicit Regularization of Sharpness-Aware Minimization for Scale-Invariant Problems 18 Oct 2024 · 0 repositories · arXiv:2410.14802
-
Improving Vision Transformers by Overlapping Heads in Multi-Head Self-Attention 18 Oct 2024 · 0 repositories · arXiv:2410.14874
-
LLM The Genius Paradox: A Linguistic and Math Expert's Struggle with Simple Word-based Counting Problems 18 Oct 2024 · 0 repositories · arXiv:2410.14166
-
LUDVIG: Learning-free Uplifting of 2D Visual features to Gaussian Splatting scenes 18 Oct 2024 · 0 repositories · arXiv:2410.14462
-
Machine Learning Aided Modeling of Granular Materials: A Review 18 Oct 2024 · 0 repositories · arXiv:2410.14767
-
MambaSCI: Efficient Mamba-UNet for Quad-Bayer Patterned Video Snapshot Compressive Imaging 18 Oct 2024 · 0 repositories · arXiv:2410.14214Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Mixed Attention Transformer Enhanced Channel Estimation for Extremely Large-Scale MIMO Systems 18 Oct 2024 · 0 repositories · arXiv:2410.14439
-
MultiOrg: A Multi-rater Organoid-detection Dataset 18 Oct 2024 · 0 repositories · arXiv:2410.14612
-
Novel Development of LLM Driven mCODE Data Model for Improved Clinical Trial Matching to Enable Standardization and Interoperability in Oncology Research 18 Oct 2024 · 0 repositories · arXiv:2410.19826
-
Optimizing Attention with Mirror Descent: Generalized Max-Margin Token Selection 18 Oct 2024 · 0 repositories · arXiv:2410.14581
-
Optimizing Retrieval-Augmented Generation with Elasticsearch for Enhanced Question-Answering Systems 18 Oct 2024 · 0 repositories · arXiv:2410.14167
-
Paths-over-Graph: Knowledge Graph Empowered Large Language Model Reasoning 18 Oct 2024 · 1 repository · arXiv:2410.14211Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
Personalized Image Generation with Large Multimodal Models 18 Oct 2024 · 1 repository · arXiv:2410.14170Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Privacy for Free in the Overparameterized Regime 18 Oct 2024 · 1 repository · arXiv:2410.14787Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Provable In-context Learning for Mixture of Linear Regressions using Transformers 18 Oct 2024 · 0 repositories · arXiv:2410.14183
-
Pseudo-label Refinement for Improving Self-Supervised Learning Systems 18 Oct 2024 · 0 repositories · arXiv:2410.14242
-
ELOQ: Resources for Enhancing LLM Detection of Out-of-Scope Questions 18 Oct 2024 · 1 repository · arXiv:2410.14567
-
Real-time Fake News from Adversarial Feedback 18 Oct 2024 · 1 repository · arXiv:2410.14651
-
Rethinking Transformer for Long Contextual Histopathology Whole Slide Image Analysis 18 Oct 2024 · 1 repository · arXiv:2410.14195Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Self-Satisfied: An end-to-end framework for SAT generation and prediction 18 Oct 2024 · 0 repositories · arXiv:2410.14888
-
Sentiment Analysis Based on RoBERTa for Amazon Review: An Empirical Study on Decision Making 18 Oct 2024 · 0 repositories · arXiv:2411.00796
-
SignAttention: On the Interpretability of Transformer Models for Sign Language Translation 18 Oct 2024 · 1 repository · arXiv:2410.14506Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
SPRIG: Improving Large Language Model Performance by System Prompt Optimization 18 Oct 2024 · 1 repository · arXiv:2410.14826Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
ST-MoE-BERT: A Spatial-Temporal Mixture-of-Experts Framework for Long-Term Cross-City Mobility Prediction 18 Oct 2024 · 1 repository · arXiv:2410.14099Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)