Methods › Natural Language Processing › Autoregressive Transformers › Transformer › Papers, page 4
Transformer
Papers archive 2025-07-28
archive papers tagged: 13,999 · with a code link: 6,572 · where Syntology ran a sample: 2,248 (1,919 with a run with no instrument failure, 329 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,248 of 13,999 tagged: 1,919 with a run with no instrument failure, 329 where every run was a failure of Syntology's instrument)
Page 4 of 140: papers 301 to 400 of 13,999, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Fusion of Foundation and Vision Transformer Model Features for Dermatoscopic Image Classification 22 May 2025 · 0 repositories · arXiv:2505.16338
-
Learning Normal Patterns in Musical Loops 22 May 2025 · 0 repositories · arXiv:2505.23784
-
Native Segmentation Vision Transformers 22 May 2025 · 0 repositories · arXiv:2505.16993
-
Scalable Graph Generative Modeling via Substructure Sequences 22 May 2025 · 1 repository · arXiv:2505.16130Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Swin Transformer for Robust CGI Images Detection: Intra- and Inter-Dataset Analysis across Multiple Color Spaces 22 May 2025 · 0 repositories · arXiv:2505.16253
-
Training-Free Efficient Video Generation via Dynamic Token Carving 22 May 2025 · 1 repository · arXiv:2505.16864
-
Transformer Copilot: Learning from The Mistake Log in LLM Fine-tuning 22 May 2025 · 1 repository · arXiv:2505.16270Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Understanding Differential Transformer Unchains Pretrained Self-Attentions 22 May 2025 · 0 repositories · arXiv:2505.16333
-
An Exploratory Approach Towards Investigating and Explaining Vision Transformer and Transfer Learning for Brain Disease Detection 21 May 2025 · 0 repositories · arXiv:2505.16039
-
BountyBench: Dollar Impact of AI Agent Attackers and Defenders on Real-World Cybersecurity Systems 21 May 2025 · 0 repositories · arXiv:2505.15216
-
Leveraging Large Language Models for Command Injection Vulnerability Analysis in Python: An Empirical Study on Popular Open-Source Projects 21 May 2025 · 0 repositories · arXiv:2505.15088
-
Mechanistic Insights into Grokking from the Embedding Layer 21 May 2025 · 0 repositories · arXiv:2505.15624
-
Robo-DM: Data Management For Large Robot Datasets 21 May 2025 · 0 repositories · arXiv:2505.15558
-
SAMA-UNet: Enhancing Medical Image Segmentation with Self-Adaptive Mamba-Like Attention and Causal-Resonance Learning 21 May 2025 · 1 repository · arXiv:2505.15234
-
Scaling Diffusion Transformers Efficiently via μP 21 May 2025 · 1 repository · arXiv:2505.15270
-
Small Language Models in the Real World: Insights from Industrial Text Classification 21 May 2025 · 0 repositories · arXiv:2505.16078
-
Sonnet: Spectral Operator Neural Network for Multivariable Time Series Forecasting 21 May 2025 · 1 repository · arXiv:2505.15312
-
FlowBERT: Prompt-tuned BERT for variable flow field prediction 20 May 2025 · 0 repositories · arXiv:2506.08021
-
Articulatory Feature Prediction from Surface EMG during Speech Production 20 May 2025 · 1 repository · arXiv:2505.13814
-
CAD-Coder: An Open-Source Vision-Language Model for Computer-Aided Design Code Generation 20 May 2025 · 1 repository · arXiv:2505.14646Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Cost-Augmented Monte Carlo Tree Search for LLM-Assisted Planning 20 May 2025 · 0 repositories · arXiv:2505.14656
-
Do Language Models Use Their Depth Efficiently? 20 May 2025 · 1 repository · arXiv:2505.13898Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
DSMentor: Enhancing Data Science Agents with Curriculum Learning and Online Knowledge Accumulation 20 May 2025 · 0 repositories · arXiv:2505.14163
-
Energy-Efficient Deep Reinforcement Learning with Spiking Transformers 20 May 2025 · 0 repositories · arXiv:2505.14533
-
FlashKAT: Understanding and Addressing Performance Bottlenecks in the Kolmogorov-Arnold Transformer 20 May 2025 · 1 repository · arXiv:2505.13813
-
Latent Flow Transformer 20 May 2025 · 1 repository · arXiv:2505.14513
-
Learning Spatio-Temporal Dynamics for Trajectory Recovery via Time-Aware Transformer 20 May 2025 · 2 repositories · arXiv:2505.13857
-
Low-Cost FlashAttention with Fused Exponential and Multiplication Hardware Operators 20 May 2025 · 0 repositories · arXiv:2505.14314
-
ModRWKV: Transformer Multimodality in Linear Time 20 May 2025 · 1 repository · arXiv:2505.14505
-
MSDformer: Multi-scale Discrete Transformer For Time Series Generation 20 May 2025 · 0 repositories · arXiv:2505.14202
-
Multi-Channel Swin Transformer Framework for Bearing Remaining Useful Life Prediction 20 May 2025 · 0 repositories · arXiv:2505.14897
-
OmniStyle: Filtering High Quality Style Transfer Data at Scale 20 May 2025 · 1 repository · arXiv:2505.14028
-
ReactDiff: Latent Diffusion for Facial Reaction Generation 20 May 2025 · 1 repository · arXiv:2505.14151
-
Selective Structured State Space for Multispectral-fused Small Target Detection 20 May 2025 · 0 repositories · arXiv:2505.14043
-
Normalized Cut with Reinforcement Learning in Constrained Action Space 20 May 2025 · 0 repositories · arXiv:2505.13986
-
STree: Speculative Tree Decoding for Hybrid State-Space Models 20 May 2025 · 0 repositories · arXiv:2505.14969
-
Subquadratic Algorithms and Hardness for Attention with Any Temperature 20 May 2025 · 0 repositories · arXiv:2505.14840
-
TCSinger 2: Customizable Multilingual Zero-shot Singing Voice Synthesis 20 May 2025 · 1 repository · arXiv:2505.14910
-
A3 : an Analytical Low-Rank Approximation Framework for Attention 19 May 2025 · 0 repositories · arXiv:2505.12942
-
Adversarial Testing in LLMs: Insights into Decision-Making Vulnerabilities 19 May 2025 · 0 repositories · arXiv:2505.13195
-
Are Large Language Models Good at Detecting Propaganda? 19 May 2025 · 0 repositories · arXiv:2505.13706
-
CMLFormer: A Dual Decoder Transformer with Switching Point Learning for Code-Mixed Language Modeling 19 May 2025 · 0 repositories · arXiv:2505.12587
-
Enhancing Latent Computation in Transformers with Latent Tokens 19 May 2025 · 0 repositories · arXiv:2505.12629
-
Evaluating the Performance of RAG Methods for Conversational AI in the Airport Domain 19 May 2025 · 0 repositories · arXiv:2505.13006
-
GuidedMorph: Two-Stage Deformable Registration for Breast MRI 19 May 2025 · 0 repositories · arXiv:2505.13414
-
LiDAR MOT-DETR: A LiDAR-based Two-Stage Transformer for 3D Multiple Object Tracking 19 May 2025 · 0 repositories · arXiv:2505.12753
-
MSVIT: Improving Spiking Vision Transformer Using Multi-scale Attention Fusion 19 May 2025 · 1 repository · arXiv:2505.14719
-
Multi-head Temporal Latent Attention 19 May 2025 · 1 repository · arXiv:2505.13544Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
OMGPT: A Sequence Modeling Framework for Data-driven Operational Decision Making 19 May 2025 · 0 repositories · arXiv:2505.13580
-
PPTNet: A Hybrid Periodic Pattern-Transformer Architecture for Traffic Flow Prediction and Congestion Identification 19 May 2025 · 1 repository · arXiv:2505.13047
-
Pyramid Sparse Transformer: Enhancing Multi-Scale Feature Fusion with Dynamic Token Selection 19 May 2025 · 0 repositories · arXiv:2505.12772
-
Simplicity is Key: An Unsupervised Pretraining Approach for Sparse Radio Channels 19 May 2025 · 0 repositories · arXiv:2505.13055
-
SounDiT: Geo-Contextual Soundscape-to-Landscape Generation 19 May 2025 · 0 repositories · arXiv:2505.12734
-
The Hidden Structure -- Improving Legal Document Understanding Through Explicit Text Formatting 19 May 2025 · 0 repositories · arXiv:2505.12837
-
Unified Cross-modal Translation of Score Images, Symbolic Music, and Performance Audio 19 May 2025 · 0 repositories · arXiv:2505.12863
-
Unpacking Positional Encoding in Transformers: A Spectral Analysis of Content-Position Coupling 19 May 2025 · 0 repositories · arXiv:2505.13027Syntology 6 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
WriteViT: Handwritten Text Generation with Vision Transformer 19 May 2025 · 1 repository · arXiv:2505.13235
-
EuLearn: A 3D database for learning Euler characteristics 18 May 2025 · 1 repository · arXiv:2505.13539
-
EVALOOP: Assessing LLM Robustness in Programming from a Self-consistency Perspective 18 May 2025 · 0 repositories · arXiv:2505.12185
-
Chain-of-Model Learning for Language Model 17 May 2025 · 0 repositories · arXiv:2505.11820
-
GeoMaNO: Geometric Mamba Neural Operator for Partial Differential Equations 17 May 2025 · 0 repositories · arXiv:2505.12020
-
MedVKAN: Efficient Feature Extraction with Mamba and KAN for Medical Image Segmentation 17 May 2025 · 1 repository · arXiv:2505.11797
-
VeriReason: Reinforcement Learning with Testbench Feedback for Reasoning-Enhanced Verilog Generation 17 May 2025 · 1 repository · arXiv:2505.11849
-
Have Multimodal Large Language Models (MLLMs) Really Learned to Tell the Time on Analog Clocks? 16 May 2025 · 0 repositories · arXiv:2505.10862
-
CTP: A hybrid CNN-Transformer-PINN model for ocean front forecasting 16 May 2025 · 0 repositories · arXiv:2505.10894
-
Connecting the Dots: A Chain-of-Collaboration Prompting Framework for LLM Agents 16 May 2025 · 0 repositories · arXiv:2505.10936
-
Relational Graph Transformer 16 May 2025 · 1 repository · arXiv:2505.10960Syntology official (archive's flag): 2 ran · 3 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Illusion or Algorithm? Investigating Memorization, Emergence, and Symbolic Processing in In-Context Learning 16 May 2025 · 1 repository · arXiv:2505.11004
-
Attention on the Sphere 16 May 2025 · 1 repository · arXiv:2505.11157
-
CheX-DS: Improving Chest X-ray Image Classification with Ensemble Learning Based on DenseNet and Swin Transformer 16 May 2025 · 0 repositories · arXiv:2505.11168
-
DiCo: Revitalizing ConvNets for Scalable and Efficient Diffusion Modeling 16 May 2025 · 1 repository · arXiv:2505.11196Syntology official (archive's flag): 4 ran · 5 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Enhancing Mathematics Learning for Hard-of-Hearing Students Through Real-Time Palestinian Sign Language Recognition: A New Dataset 16 May 2025 · 0 repositories · arXiv:2505.17055
-
Heart2Mind: Human-Centered Contestable Psychiatric Disorder Diagnosis System using Wearable ECG Monitors 16 May 2025 · 1 repository · arXiv:2505.11612
-
Optimal Control for Transformer Architectures: Enhancing Generalization, Robustness and Efficiency 16 May 2025 · 0 repositories · arXiv:2505.13499
-
Transforming Decoder-Only Transformers for Accurate WiFi-Telemetry Based Indoor Localization 16 May 2025 · 0 repositories · arXiv:2505.15835
-
Vaiage: A Multi-Agent Solution to Personalized Travel Planning 16 May 2025 · 0 repositories · arXiv:2505.10922
-
Continuity and Isolation Lead to Doubts or Dilemmas in Large Language Models 15 May 2025 · 0 repositories · arXiv:2505.10606
-
A Modular Approach for Clinical SLMs Driven by Synthetic Data with Pre-Instruction Tuning, Model Merging, and Clinical-Tasks Alignment 15 May 2025 · 0 repositories · arXiv:2505.10717
-
All You Need Is Synthetic Task Augmentation 15 May 2025 · 0 repositories · arXiv:2505.10120
-
Assessing Collective Reasoning in Multi-Agent LLMs via Hidden Profile Tasks 15 May 2025 · 0 repositories · arXiv:2505.11556
-
Automating Security Audit Using Large Language Model based Agent: An Exploration Experiment 15 May 2025 · 0 repositories · arXiv:2505.10732
-
Comparing LLM Text Annotation Skills: A Study on Human Rights Violations in Social Media Data 15 May 2025 · 1 repository · arXiv:2505.10260
-
Does Scaling Law Apply in Time Series Forecasting? 15 May 2025 · 0 repositories · arXiv:2505.10172
-
On Technique Identification and Threat-Actor Attribution using LLMs and Embedding Models 15 May 2025 · 1 repository · arXiv:2505.11547
-
Pre-Act: Multi-Step Planning and Reasoning Improves Acting in LLM Agents 15 May 2025 · 0 repositories · arXiv:2505.09970
-
Private Transformer Inference in MLaaS: A Survey 15 May 2025 · 0 repositories · arXiv:2505.10315
-
Rethinking Prompt Optimizers: From Prompt Merits to Optimization 15 May 2025 · 1 repository · arXiv:2505.09930
-
SpikeVideoFormer: An Efficient Spike-Driven Video Transformer with Hamming Attention and 𝒪(T) Complexity 15 May 2025 · 1 repository · arXiv:2505.10352Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
VRU-CIPI: Crossing Intention Prediction at Intersections for Improving Vulnerable Road Users Safety 15 May 2025 · 0 repositories · arXiv:2505.09935
-
A Comprehensive Analysis of Large Language Model Outputs: Similarity, Diversity, and Bias 14 May 2025 · 0 repositories · arXiv:2505.09056
-
AdaFortiTran: An Adaptive Transformer Model for Robust OFDM Channel Estimation 14 May 2025 · 1 repository · arXiv:2505.09076
-
Beyond the Known: Decision Making with Counterfactual Reasoning Decision Transformer 14 May 2025 · 1 repository · arXiv:2505.09114
-
BrainNetMLP: An Efficient and Effective Baseline for Functional Brain Network Classification 14 May 2025 · 1 repository · arXiv:2505.11538
-
FAS-LLM: Large Language Model-Based Channel Prediction for OTFS-Enabled Satellite-FAS Links 14 May 2025 · 0 repositories · arXiv:2505.09751
-
How Hungry is AI? Benchmarking Energy, Water, and Carbon Footprint of LLM Inference 14 May 2025 · 0 repositories · arXiv:2505.09598
-
LAS: Loss-less ANN-SNN Conversion for Fully Spike-Driven Large Language Models 14 May 2025 · 1 repository · arXiv:2505.09659
-
Out-of-distribution generalisation is hard: evidence from ARC-like tasks 14 May 2025 · 0 repositories · arXiv:2505.09716
-
Quotient Complex Transformer (QCformer) for Perovskite Data Analysis 14 May 2025 · 0 repositories · arXiv:2505.09174
-
TopoDiT-3D: Topology-Aware Diffusion Transformer with Bottleneck Structure for 3D Point Cloud Generation 14 May 2025 · 1 repository · arXiv:2505.09140
-
Zero-Shot Multi-modal Large Language Model v.s. Supervised Deep Learning: A Comparative Study on CT-Based Intracranial Hemorrhage Subtyping 14 May 2025 · 1 repository · arXiv:2505.09252