Methods › Natural Language Processing › Autoregressive Transformers › Transformer › Papers, page 13
Transformer
Papers archive 2025-07-28
archive papers tagged: 13,999 · with a code link: 6,572 · where Syntology ran a sample: 2,248 (1,919 with a run with no instrument failure, 329 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,248 of 13,999 tagged: 1,919 with a run with no instrument failure, 329 where every run was a failure of Syntology's instrument)
Page 13 of 140: papers 1,201 to 1,300 of 13,999, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Self-Evaluation for Job-Shop Scheduling 12 Feb 2025 · 0 repositories · arXiv:2502.08684
-
Auditing Prompt Caching in Language Model APIs 11 Feb 2025 · 1 repository · arXiv:2502.07776Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
CausalGeD: Blending Causality and Diffusion for Spatial Gene Expression Generation 11 Feb 2025 · 0 repositories · arXiv:2502.07751
-
Deep Semantic Graph Learning via LLM based Node Enhancement 11 Feb 2025 · 0 repositories · arXiv:2502.07982
-
Dense Object Detection Based on De-homogenized Queries 11 Feb 2025 · 0 repositories · arXiv:2502.07194
-
Fast-COS: A Fast One-Stage Object Detector Based on Reparameterized Attention Vision Transformer for Autonomous Driving 11 Feb 2025 · 0 repositories · arXiv:2502.07417
-
Linear Transformers as VAR Models: Aligning Autoregressive Attention Mechanisms with Autoregressive Forecasting 11 Feb 2025 · 1 repository · arXiv:2502.07244Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
MAAT: Mamba Adaptive Anomaly Transformer with association discrepancy for time series 11 Feb 2025 · 1 repository · arXiv:2502.07858
-
Mask-Enhanced Autoregressive Prediction: Pay Less Attention to Learn More 11 Feb 2025 · 1 repository · arXiv:2502.07490
-
MIGT: Memory Instance Gated Transformer Framework for Financial Portfolio Management 11 Feb 2025 · 0 repositories · arXiv:2502.07280
-
OpenGrok: Enhancing SNS Data Processing with Distilled Knowledge and Mask-like Mechanisms 11 Feb 2025 · 1 repository · arXiv:2502.07312
-
Tractable Transformers for Flexible Conditional Generation 11 Feb 2025 · 0 repositories · arXiv:2502.07616
-
VidCRAFT3: Camera, Object, and Lighting Control for Image-to-Video Generation 11 Feb 2025 · 0 repositories · arXiv:2502.07531
-
A Simple yet Effective DDG Predictor is An Unsupervised Antibody Optimizer and Explainer 10 Feb 2025 · 1 repository · arXiv:2502.06913
-
Finding Words Associated with DIF: Predicting Differential Item Functioning using LLMs and Explainable AI 10 Feb 2025 · 0 repositories · arXiv:2502.07017
-
Foundation Model of Electronic Medical Records for Adaptive Risk Estimation 10 Feb 2025 · 1 repository · arXiv:2502.06124
-
Fully Exploiting Vision Foundation Model's Profound Prior Knowledge for Generalizable RGB-Depth Driving Scene Parsing 10 Feb 2025 · 0 repositories · arXiv:2502.06219
-
History-Guided Video Diffusion 10 Feb 2025 · 1 repository · arXiv:2502.06764
-
Powerformer: A Transformer with Weighted Causal Attention for Time-series Forecasting 10 Feb 2025 · 1 repository · arXiv:2502.06151
-
Towards bandit-based prompt-tuning for in-the-wild foundation agents 10 Feb 2025 · 0 repositories · arXiv:2502.06358
-
Unconstrained Body Recognition at Altitude and Range: Comparing Four Approaches 10 Feb 2025 · 0 repositories · arXiv:2502.07130
-
Utilizing Novelty-based Evolution Strategies to Train Transformers in Reinforcement Learning 10 Feb 2025 · 0 repositories · arXiv:2502.06301
-
ViSIR: Vision Transformer Single Image Reconstruction Method for Earth System Models 10 Feb 2025 · 0 repositories · arXiv:2502.06741
-
Emergence of Episodic Memory in Transformers: Characterizing Changes in Temporal Structure of Attention Scores During Training 9 Feb 2025 · 0 repositories · arXiv:2502.06902
-
HyLiFormer: Hyperbolic Linear Attention for Skeleton-based Human Action Recognition 9 Feb 2025 · 0 repositories
-
Investigating Compositional Reasoning in Time Series Foundation Models 9 Feb 2025 · 1 repository · arXiv:2502.06037
-
Large Language Models for In-File Vulnerability Localization Can Be "Lost in the End" 9 Feb 2025 · 0 repositories · arXiv:2502.06898
-
LM2: Large Memory Models 9 Feb 2025 · 2 repositories · arXiv:2502.06049
-
Saving 77% of the Parameters in Large Language Models Technical Report 9 Feb 2025 · 1 repository
-
ScaffoldGPT: A Scaffold-based GPT Model for Drug Optimization 9 Feb 2025 · 0 repositories · arXiv:2502.06891
-
The Curse of Depth in Large Language Models 9 Feb 2025 · 0 repositories · arXiv:2502.05795
-
VFX Creator: Animated Visual Effect Generation with Controllable Diffusion Transformer 9 Feb 2025 · 0 repositories · arXiv:2502.05979
-
Dynamic Noise Preference Optimization for LLM Self-Improvement via Synthetic Data 8 Feb 2025 · 0 repositories · arXiv:2502.05400
-
Event Stream-based Visual Object Tracking: HDETrack V2 and A High-Definition Benchmark 8 Feb 2025 · 1 repository · arXiv:2502.05574
-
Flow-based Conformal Prediction for Multi-dimensional Time Series 8 Feb 2025 · 0 repositories · arXiv:2502.05709
-
Flowing Through Layers: A Continuous Dynamical Systems Perspective on Transformers 8 Feb 2025 · 0 repositories · arXiv:2502.05656
-
GWRF: A Generalizable Wireless Radiance Field for Wireless Signal Propagation Modeling 8 Feb 2025 · 0 repositories · arXiv:2502.05708
-
Multi-scale Masked Autoencoder for Electrocardiogram Anomaly Detection 8 Feb 2025 · 0 repositories · arXiv:2502.05494
-
A Deep Learning Framework Integrating CNN and BiLSTM for Financial Systemic Risk Analysis and Prediction 7 Feb 2025 · 0 repositories · arXiv:2502.06847
-
Can Large Language Models Understand Intermediate Representations? 7 Feb 2025 · 0 repositories · arXiv:2502.06854
-
Cross-Encoder Rediscovers a Semantic Variant of BM25 7 Feb 2025 · 0 repositories · arXiv:2502.04645
-
Enhancing Pre-Trained Decision Transformers with Prompt-Tuning Bandits 7 Feb 2025 · 0 repositories · arXiv:2502.04979
-
HetSSNet: Spatial-Spectral Heterogeneous Graph Learning Network for Panchromatic and Multispectral Images Fusion 7 Feb 2025 · 0 repositories · arXiv:2502.04623
-
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation 7 Feb 2025 · 0 repositories · arXiv:2502.04847
-
Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance 7 Feb 2025 · 0 repositories · arXiv:2502.05236
-
Learning the Language of NVMe Streams for Ransomware Detection 7 Feb 2025 · 0 repositories · arXiv:2502.05011
-
MedMimic: Physician-Inspired Multimodal Fusion for Early Diagnosis of Fever of Unknown Origin 7 Feb 2025 · 0 repositories · arXiv:2502.04794
-
SelaFD:Seamless Adaptation of Vision Transformer Fine-tuning for Radar-based Human Activity 7 Feb 2025 · 1 repository · arXiv:2502.04740
-
Swin-MSTP: Swin transformer with multi-scale temporal perception for continuous sign language recognition 7 Feb 2025 · 1 repository
-
Unsafe LLM-Based Search: Quantitative Analysis and Mitigation of Safety Risks in AI Web Search 7 Feb 2025 · 0 repositories · arXiv:2502.04951
-
A Decoding Algorithm for Length-Control Summarization Based on Directed Acyclic Transformers 6 Feb 2025 · 1 repository · arXiv:2502.04535
-
A Retrospective Systematic Study on Hierarchical Sparse Query Transformer-assisted Ultrasound Screening for Early Hepatocellular Carcinoma 6 Feb 2025 · 1 repository · arXiv:2502.03772
-
A Self-supervised Multimodal Deep Learning Approach to Differentiate Post-radiotherapy Progression from Pseudoprogression in Glioblastoma 6 Feb 2025 · 0 repositories · arXiv:2502.03999
-
Beyond the Final Layer: Hierarchical Query Fusion Transformer with Agent-Interpolation Initialization for 3D Instance Segmentation 6 Feb 2025 · 0 repositories · arXiv:2502.04139
-
Building A Unified AI-centric Language System: analysis, framework and future work 6 Feb 2025 · 0 repositories · arXiv:2502.04488
-
DiTAR: Diffusion Transformer Autoregressive Modeling for Speech Generation 6 Feb 2025 · 0 repositories · arXiv:2502.03930
-
Expanding Training Data for Endoscopic Phenotyping of Eosinophilic Esophagitis 6 Feb 2025 · 0 repositories · arXiv:2502.04199
-
How vulnerable is my policy? Adversarial attacks on modern behavior cloning policies 6 Feb 2025 · 0 repositories · arXiv:2502.03698
-
ImprovNet -- Generating Controllable Musical Improvisations with Iterative Corruption Refinement 6 Feb 2025 · 1 repository · arXiv:2502.04522
-
Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis 6 Feb 2025 · 1 repository · arXiv:2502.04128
-
MedGNN: Towards Multi-resolution Spatiotemporal Graph Learning for Medical Time Series Classification 6 Feb 2025 · 1 repository · arXiv:2502.04515
-
Multilingual Non-Autoregressive Machine Translation without Knowledge Distillation 6 Feb 2025 · 1 repository · arXiv:2502.04537
-
Semantic Feature Division Multiple Access for Digital Semantic Broadcast Channels 6 Feb 2025 · 0 repositories · arXiv:2502.03949
-
SMI: An Information-Theoretic Metric for Predicting Model Knowledge Solely from Pre-Training Signals 6 Feb 2025 · 1 repository · arXiv:2502.04066
-
Vision-Integrated LLMs for Autonomous Driving Assistance : Human Performance Comparison and Trust Evaluation 6 Feb 2025 · 0 repositories · arXiv:2502.06843
-
Zero-shot Meta-learning for Tabular Prediction Tasks with Adversarially Pre-trained Transformer 6 Feb 2025 · 0 repositories · arXiv:2502.04573
-
Every Angle Is Worth A Second Glance: Mining Kinematic Skeletal Structures from Multi-view Joint Cloud 5 Feb 2025 · 0 repositories · arXiv:2502.02936
-
Label Anything: An Interpretable, High-Fidelity and Prompt-Free Annotator 5 Feb 2025 · 0 repositories · arXiv:2502.02972
-
Multimodal Transformer Models for Turn-taking Prediction: Effects on Conversational Dynamics of Human-Agent Interaction during Cooperative Gameplay 5 Feb 2025 · 0 repositories · arXiv:2503.16432
-
Omni-DNA: A Unified Genomic Foundation Model for Cross-Modal and Multi-Task Learning 5 Feb 2025 · 0 repositories · arXiv:2502.03499
-
OPTIC: Optimizing Patient-Provider Triaging & Improving Communications in Clinical Operations using GPT-4 Data Labeling and Model Distillation 5 Feb 2025 · 0 repositories · arXiv:2503.05701
-
Optimizing Robustness and Accuracy in Mixture of Experts: A Dual-Model Approach 5 Feb 2025 · 0 repositories · arXiv:2502.06832
-
Scaling laws in wearable human activity recognition 5 Feb 2025 · 0 repositories · arXiv:2502.03364
-
Distribution Transformers: Fast Approximate Bayesian Inference With On-The-Fly Prior Adaptation 4 Feb 2025 · 0 repositories · arXiv:2502.02463
-
Exploiting Ensemble Learning for Cross-View Isolated Sign Language Recognition 4 Feb 2025 · 1 repository · arXiv:2502.02196
-
Exploring the Panorama of Anxiety Levels: A Multi-Scenario Study Based on Human-Centric Anxiety Level Detection and Personalized Guidance 4 Feb 2025 · 0 repositories · arXiv:2503.15527
-
IncepFormerNet: A multi-scale multi-head attention network for SSVEP classification 4 Feb 2025 · 1 repository · arXiv:2502.13972
-
LLMER: Crafting Interactive Extended Reality Worlds with JSON Data Generated by Large Language Models 4 Feb 2025 · 1 repository · arXiv:2502.02441
-
MATCNN: Infrared and Visible Image Fusion Method Based on Multi-scale CNN with Attention Transformer 4 Feb 2025 · 1 repository · arXiv:2502.01959
-
Memory Efficient Transformer Adapter for Dense Predictions 4 Feb 2025 · 0 repositories · arXiv:2502.01962
-
MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion Paradigm 4 Feb 2025 · 0 repositories · arXiv:2502.02358
-
Open Foundation Models in Healthcare: Challenges, Paradoxes, and Opportunities with GenAI Driven Personalized Prescription 4 Feb 2025 · 0 repositories · arXiv:2502.04356
-
Peri-LN: Revisiting Layer Normalization in the Transformer Architecture 4 Feb 2025 · 0 repositories · arXiv:2502.02732
-
RIE-SenseNet: Riemannian Manifold Embedding of Multi-Source Industrial Sensor Signals for Robust Pattern Recognition 4 Feb 2025 · 0 repositories · arXiv:2502.02428
-
UniGaze: Towards Universal Gaze Estimation via Large-scale Pre-Training 4 Feb 2025 · 0 repositories · arXiv:2502.02307
-
Message-Passing GNNs Fail to Approximate Sparse Triangular Factorizations 3 Feb 2025 · 0 repositories · arXiv:2502.01397
-
GNN-DT: Graph Neural Network Enhanced Decision Transformer for Efficient Optimization in Dynamic Environments 3 Feb 2025 · 1 repository · arXiv:2502.01778
-
Joint Localization and Activation Editing for Low-Resource Fine-Tuning 3 Feb 2025 · 1 repository · arXiv:2502.01179
-
Scalable Language Models with Posterior Inference of Latent Thought Vectors 3 Feb 2025 · 0 repositories · arXiv:2502.01567
-
Toward Neurosymbolic Program Comprehension 3 Feb 2025 · 0 repositories · arXiv:2502.01806
-
Transformers trained on proteins can learn to attend to Euclidean distance 3 Feb 2025 · 1 repository · arXiv:2502.01533
-
DeepGate4: Efficient and Effective Representation Learning for Circuit Design at Scale 2 Feb 2025 · 1 repository · arXiv:2502.01681Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Estimating forest carbon stocks from high-resolution remote sensing imagery by reducing domain shift with style transfer 2 Feb 2025 · 0 repositories · arXiv:2502.00784
-
A Study on the Performance of U-Net Modifications in Retroperitoneal Tumor Segmentation 1 Feb 2025 · 1 repository · arXiv:2502.00314
-
Benchmark on Peer Review Toxic Detection: A Challenging Task with a New Dataset 1 Feb 2025 · 0 repositories · arXiv:2502.01676
-
Contrastive Forward-Forward: A Training Algorithm of Vision Transformer 1 Feb 2025 · 0 repositories · arXiv:2502.00571
-
Converting Transformers into DGNNs Form 1 Feb 2025 · 1 repository · arXiv:2502.00585
-
MambaGlue: Fast and Robust Local Feature Matching With Mamba 1 Feb 2025 · 1 repository · arXiv:2502.00462
-
Milmer: a Framework for Multiple Instance Learning based Multimodal Emotion Recognition 1 Feb 2025 · 1 repository · arXiv:2502.00547
-
Sigmoid Self-Attention has Lower Sample Complexity than Softmax Self-Attention: A Mixture-of-Experts Perspective 1 Feb 2025 · 0 repositories · arXiv:2502.00281