Methods › Natural Language Processing › Autoregressive Transformers › Transformer › Papers, page 9
Transformer
Papers archive 2025-07-28
archive papers tagged: 13,999 · with a code link: 6,572 · where Syntology ran a sample: 2,248 (1,919 with a run with no instrument failure, 329 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,248 of 13,999 tagged: 1,919 with a run with no instrument failure, 329 where every run was a failure of Syntology's instrument)
Page 9 of 140: papers 801 to 900 of 13,999, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Distil-xLSTM: Learning Attention Mechanisms through Recurrent Structures 24 Mar 2025 · 0 repositories · arXiv:2503.18565
-
Exploring the Integration of Key-Value Attention Into Pure and Hybrid Transformers for Semantic Segmentation 24 Mar 2025 · 0 repositories · arXiv:2503.18862
-
Global-Local Tree Search in VLMs for 3D Indoor Scene Generation 24 Mar 2025 · 1 repository · arXiv:2503.18476Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
LGI-DETR: Local-Global Interaction for UAV Object Detection 24 Mar 2025 · 0 repositories · arXiv:2503.18785
-
LoTUS: Large-Scale Machine Unlearning with a Taste of Uncertainty 24 Mar 2025 · 1 repository · arXiv:2503.18314
-
SPMTrack: Spatio-Temporal Parameter-Efficient Fine-Tuning with Mixture of Experts for Scalable Visual Tracking 24 Mar 2025 · 1 repository · arXiv:2503.18338Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
U-REPA: Aligning Diffusion U-Nets to ViTs 24 Mar 2025 · 1 repository · arXiv:2503.18414
-
Your ViT is Secretly an Image Segmentation Model 24 Mar 2025 · 1 repository · arXiv:2503.19108Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
ZeroLM: Data-Free Transformer Architecture Search for Language Models 24 Mar 2025 · 0 repositories · arXiv:2503.18646
-
Adaptive Rank Allocation: Speeding Up Modern Transformers with RaNA Adapters 23 Mar 2025 · 1 repository · arXiv:2503.18216Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
End-to-End Implicit Neural Representations for Classification 23 Mar 2025 · 1 repository · arXiv:2503.18123
-
ExpertRAG: Efficient RAG with Mixture of Experts -- Optimizing Context Retrieval for Adaptive LLM Responses 23 Mar 2025 · 0 repositories · arXiv:2504.08744
-
PathoHR: Breast Cancer Survival Prediction on High-Resolution Pathological Images 23 Mar 2025 · 1 repository · arXiv:2503.17970
-
SymmCompletion: High-Fidelity and High-Consistency Point Cloud Completion with Symmetry Guidance 23 Mar 2025 · 1 repository · arXiv:2503.18007
-
A Modular Dataset to Demonstrate LLM Abstraction Capability 22 Mar 2025 · 0 repositories · arXiv:2503.17645
-
Automated diagnosis of lung diseases using vision transformer: a comparative study on chest x-ray classification 22 Mar 2025 · 0 repositories · arXiv:2503.18973
-
Bandwidth Reservation for Time-Critical Vehicular Applications: A Multi-Operator Environment 22 Mar 2025 · 0 repositories · arXiv:2503.17756
-
EMPLACE: Self-Supervised Urban Scene Change Detection 22 Mar 2025 · 1 repository · arXiv:2503.17716Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 7 harvested samples) · 7 pointer-only (licence)
-
Hierarchy-Aware and Channel-Adaptive Semantic Communication for Bandwidth-Limited Data Fusion 22 Mar 2025 · 0 repositories · arXiv:2503.17777
-
TDRI: Two-Phase Dialogue Refinement and Co-Adaptation for Interactive Image Generation 22 Mar 2025 · 0 repositories · arXiv:2503.17669
-
Assessing the Reliability and Validity of GPT-4 in Annotating Emotion Appraisal Ratings 21 Mar 2025 · 0 repositories · arXiv:2503.16883
-
CoKe: Customizable Fine-Grained Story Evaluation via Chain-of-Keyword Rationalization 21 Mar 2025 · 0 repositories · arXiv:2503.17136
-
Feature-Based Dual Visual Feature Extraction Model for Compound Multimodal Emotion Recognition 21 Mar 2025 · 1 repository · arXiv:2503.17453
-
Federated Cross-Domain Click-Through Rate Prediction With Large Language Model Augmentation 21 Mar 2025 · 0 repositories · arXiv:2503.16875
-
SaudiCulture: A Benchmark for Evaluating Large Language Models Cultural Competence within Saudi Arabia 21 Mar 2025 · 0 repositories · arXiv:2503.17485
-
Vision Transformer Based Semantic Communications for Next Generation Wireless Networks 21 Mar 2025 · 0 repositories · arXiv:2503.17275
-
When Words Outperform Vision: VLMs Can Self-Improve Via Text-Only Training For Human-Centered Decision Making 21 Mar 2025 · 0 repositories · arXiv:2503.16965Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Zero-Shot Styled Text Image Generation, but Make It Autoregressive 21 Mar 2025 · 0 repositories · arXiv:2503.17074
-
Binarized Mamba-Transformer for Lightweight Quad Bayer HybridEVS Demosaicing 20 Mar 2025 · 1 repository · arXiv:2503.16134
-
Design and Implementation of an FPGA-Based Hardware Accelerator for Transformer 20 Mar 2025 · 1 repository · arXiv:2503.16731
-
EDiT: Efficient Diffusion Transformers with Linear Compressed Attention 20 Mar 2025 · 0 repositories · arXiv:2503.16726
-
FreeFlux: Understanding and Exploiting Layer-Specific Roles in RoPE-Based MMDiT for Versatile Image Editing 20 Mar 2025 · 0 repositories · arXiv:2503.16153
-
Hyperspectral Imaging for Identifying Foreign Objects on Pork Belly 20 Mar 2025 · 0 repositories · arXiv:2503.16086
-
iFlame: Interleaving Full and Linear Attention for Efficient Mesh Generation 20 Mar 2025 · 0 repositories · arXiv:2503.16653
-
Iterative Optimal Attention and Local Model for Single Image Rain Streak Removal 20 Mar 2025 · 1 repository · arXiv:2503.16165
-
Deep learning framework for action prediction reveals multi-timescale locomotor control 20 Mar 2025 · 0 repositories · arXiv:2503.16340
-
PromptHash: Affinity-Prompted Collaborative Cross-Modal Learning for Adaptive Hashing Retrieval 20 Mar 2025 · 0 repositories · arXiv:2503.16064
-
SenseExpo: Efficient Autonomous Exploration with Prediction Information from Lightweight Neural Networks 20 Mar 2025 · 0 repositories · arXiv:2503.16000
-
SpiLiFormer: Enhancing Spiking Transformers with Lateral Inhibition 20 Mar 2025 · 0 repositories · arXiv:2503.15986
-
The Lighthouse of Language: Enhancing LLM Agents via Critique-Guided Improvement 20 Mar 2025 · 0 repositories · arXiv:2503.16024
-
Transformer-based Wireless Symbol Detection Over Fading Channels 20 Mar 2025 · 0 repositories · arXiv:2503.16594
-
UniHDSA: A Unified Relation Prediction Approach for Hierarchical Document Structure Analysis 20 Mar 2025 · 1 repository · arXiv:2503.15893
-
XAttention: Block Sparse Attention with Antidiagonal Scoring 20 Mar 2025 · 1 repository · arXiv:2503.16428
-
A Novel Channel Boosted Residual CNN-Transformer with Regional-Boundary Learning for Breast Cancer Detection 19 Mar 2025 · 0 repositories · arXiv:2503.15008
-
ChatGPT or A Silent Everywhere Helper: A Survey of Large Language Models 19 Mar 2025 · 0 repositories · arXiv:2503.17403
-
ELTEX: A Framework for Domain-Driven Synthetic Data Generation 19 Mar 2025 · 1 repository · arXiv:2503.15055
-
FP4DiT: Towards Effective Floating Point Quantization for Diffusion Transformers 19 Mar 2025 · 1 repository · arXiv:2503.15465Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
GenM³: Generative Pretrained Multi-path Motion Model for Text Conditional Human Motion Generation 19 Mar 2025 · 0 repositories · arXiv:2503.14919
-
TruthLens:A Training-Free Paradigm for DeepFake Detection 19 Mar 2025 · 0 repositories · arXiv:2503.15342
-
Understanding the Generalization of In-Context Learning in Transformers: An Empirical Study 19 Mar 2025 · 1 repository · arXiv:2503.15579
-
CTSAC: Curriculum-Based Transformer Soft Actor-Critic for Goal-Oriented Robot Exploration 18 Mar 2025 · 0 repositories · arXiv:2503.14254
-
Dynamic Accumulated Attention Map for Interpreting Evolution of Decision-Making in Vision Transformer 18 Mar 2025 · 1 repository · arXiv:2503.14640
-
Fast Autoregressive Video Generation with Diagonal Decoding 18 Mar 2025 · 0 repositories · arXiv:2503.14070
-
Gricean Norms as a Basis for Effective Collaboration 18 Mar 2025 · 1 repository · arXiv:2503.14484
-
Large Language Models for Virtual Human Gesture Selection 18 Mar 2025 · 0 repositories · arXiv:2503.14408
-
Multimodal Feature-Driven Deep Learning for the Prediction of Duck Body Dimensions and Weight 18 Mar 2025 · 0 repositories · arXiv:2503.14001
-
PENCIL: Long Thoughts with Short Memory 18 Mar 2025 · 1 repository · arXiv:2503.14337Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Theoretical Foundation of Flow-Based Time Series Generation: Provable Approximation, Generalization, and Efficiency 18 Mar 2025 · 0 repositories · arXiv:2503.14076
-
A Reinforcement Learning-Driven Transformer GAN for Molecular Generation 17 Mar 2025 · 0 repositories · arXiv:2503.12796
-
A Survey on Transformer Context Extension: Approaches and Evaluation 17 Mar 2025 · 0 repositories · arXiv:2503.13299
-
Advancing Chronic Tuberculosis Diagnostics Using Vision-Language Models: A Multi modal Framework for Precision Analysis 17 Mar 2025 · 0 repositories · arXiv:2503.14536
-
An interpretable approach to automating the assessment of biofouling in video footage 17 Mar 2025 · 1 repository · arXiv:2503.12875
-
Humanoid Policy ~ Human Policy 17 Mar 2025 · 0 repositories · arXiv:2503.13441
-
SeisRDT: Latent Diffusion Model Based On Representation Learning For Seismic Data Interpolation And Reconstruction 17 Mar 2025 · 0 repositories · arXiv:2503.21791
-
Towards Scalable Foundation Model for Multi-modal and Hyperspectral Geospatial Data 17 Mar 2025 · 0 repositories · arXiv:2503.12843
-
Fourier-Based 3D Multistage Transformer for Aberration Correction in Multicellular Specimens 16 Mar 2025 · 2 repositories · arXiv:2503.12593
-
Fragile Mastery: Are Domain-Specific Trade-Offs Undermining On-Device Language Models? 16 Mar 2025 · 0 repositories · arXiv:2503.22698
-
PA-CFL: Privacy-Adaptive Clustered Federated Learning for Transformer-Based Sales Forecasting on Heterogeneous Retail Data 15 Mar 2025 · 0 repositories · arXiv:2503.12220
-
Fast Critical Clearing Time Calculation for Power Systems with Synchronous and Asynchronous Generation 15 Mar 2025 · 0 repositories · arXiv:2503.12132
-
LLM & HPC:Benchmarking DeepSeek's Performance in High-Performance Computing Tasks 15 Mar 2025 · 1 repository · arXiv:2504.03665
-
Maritime Mission Planning for Unmanned Surface Vessel using Large Language Model 15 Mar 2025 · 0 repositories · arXiv:2503.12065
-
Addressing Information Loss and Interaction Collapse: A Dual Enhanced Attention Framework for Feature Interaction 14 Mar 2025 · 0 repositories · arXiv:2503.11233
-
Alzheimer's Disease Classification Using Retinal OCT: TransnetOCT and Swin Transformer Models 14 Mar 2025 · 0 repositories · arXiv:2503.11511
-
Asynchronous Sharpness-Aware Minimization For Fast and Accurate Deep Learning 14 Mar 2025 · 0 repositories · arXiv:2503.11147
-
BEVDiffLoc: End-to-End LiDAR Global Localization in BEV View based on Diffusion Model 14 Mar 2025 · 1 repository · arXiv:2503.11372
-
Context-Aware Rule Mining Using a Dynamic Transformer-Based Framework 14 Mar 2025 · 0 repositories · arXiv:2503.11125
-
DynRsl-VLM: Enhancing Autonomous Driving Perception with Dynamic Resolution Vision-Language Models 14 Mar 2025 · 0 repositories · arXiv:2503.11265
-
Exploring the Potential of Large Multimodal Models as Effective Alternatives for Pronunciation Assessment 14 Mar 2025 · 0 repositories · arXiv:2503.11229
-
Time and Memory Trade-off of KV-Cache Compression in Tensor Transformer Decoding 14 Mar 2025 · 0 repositories · arXiv:2503.11108
-
MEET: A Million-Scale Dataset for Fine-Grained Geospatial Scene Classification with Zoom-Free Remote Sensing Imagery 14 Mar 2025 · 0 repositories · arXiv:2503.11219
-
Prompt Sentiment: The Catalyst for LLM Change 14 Mar 2025 · 0 repositories · arXiv:2503.13510
-
RESPONSE: Benchmarking the Ability of Language Models to Undertake Commonsense Reasoning in Crisis Situation 14 Mar 2025 · 0 repositories · arXiv:2503.11348
-
Solution for 8th Competition on Affective & Behavior Analysis in-the-wild 14 Mar 2025 · 0 repositories · arXiv:2503.11115
-
TransiT: Transient Transformer for Non-line-of-sight Videography 14 Mar 2025 · 0 repositories · arXiv:2503.11328
-
TreeMeshGPT: Artistic Mesh Generation with Autoregressive Tree Sequencing 14 Mar 2025 · 1 repository · arXiv:2503.11629Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 6 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
When Do Transformers Outperform Feedforward and Recurrent Networks? A Statistical Perspective 14 Mar 2025 · 1 repository · arXiv:2503.11272
-
A Frustratingly Simple Yet Highly Effective Attack Baseline: Over 90% Success Rate Against the Strong Black-box Models of GPT-4.5/4o/o1 13 Mar 2025 · 1 repository · arXiv:2503.10635Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
A Hybrid Architecture with Efficient Fine Tuning for Abstractive Patent Document Summarization 13 Mar 2025 · 0 repositories · arXiv:2503.10354
-
Advanced Tool Learning and Selection System (ATLASS): A Closed-Loop Framework Using LLM 13 Mar 2025 · 0 repositories · arXiv:2503.10071
-
AudioX: Diffusion Transformer for Anything-to-Audio Generation 13 Mar 2025 · 0 repositories · arXiv:2503.10522
-
ChatGPT Encounters Morphing Attack Detection: Zero-Shot MAD with Multi-Modal Large Language Models and General Vision Models 13 Mar 2025 · 0 repositories · arXiv:2503.10937
-
CoCMT: Communication-Efficient Cross-Modal Transformer for Collaborative Perception 13 Mar 2025 · 0 repositories · arXiv:2503.13504
-
Cosh-DiT: Co-Speech Gesture Video Synthesis via Hybrid Audio-Visual Diffusion Transformers 13 Mar 2025 · 0 repositories · arXiv:2503.09942
-
CountPath: Automating Fragment Counting in Digital Pathology 13 Mar 2025 · 0 repositories · arXiv:2503.10520
-
Do I look like a `cat.n.01` to you? A Taxonomy Image Generation Benchmark 13 Mar 2025 · 0 repositories · arXiv:2503.10357
-
Emotion Recognition with CLIP and Sequential Learning 13 Mar 2025 · 0 repositories · arXiv:2503.09929
-
Fixed-Point RNNs: From Diagonal to Dense in a Few Iterations 13 Mar 2025 · 0 repositories · arXiv:2503.10799
-
Gumiho: A Hybrid Architecture to Prioritize Early Tokens in Speculative Decoding 13 Mar 2025 · 0 repositories · arXiv:2503.10135Syntology 12 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 8 pointer-only (licence)
-
KV-Distill: Nearly Lossless Learnable Context Compression for LLMs 13 Mar 2025 · 0 repositories · arXiv:2503.10337
-
Radar: Fast Long-Context Decoding for Any Transformer 13 Mar 2025 · 1 repository · arXiv:2503.10571Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)