Methods › Natural Language Processing › Autoregressive Transformers › Transformer › Papers, page 20
Transformer
Papers archive 2025-07-28
archive papers tagged: 13,999 · with a code link: 6,572 · where Syntology ran a sample: 2,248 (1,919 with a run with no instrument failure, 329 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,248 of 13,999 tagged: 1,919 with a run with no instrument failure, 329 where every run was a failure of Syntology's instrument)
Page 20 of 140: papers 1,901 to 2,000 of 13,999, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Training and Evaluating Language Models with Template-based Data Generation 27 Nov 2024 · 1 repository · arXiv:2411.18104Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Unpacking the Individual Components of Diffusion Policy 27 Nov 2024 · 0 repositories · arXiv:2412.00084
-
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models 26 Nov 2024 · 0 repositories · arXiv:2411.17182
-
Can artificial intelligence predict clinical trial outcomes? 26 Nov 2024 · 0 repositories · arXiv:2411.17595
-
Distributed Sign Momentum with Local Steps for Training Transformers 26 Nov 2024 · 1 repository · arXiv:2411.17866
-
ER2Score: LLM-based Explainable and Customizable Metric for Assessing Radiology Reports with Reward-Control Loss 26 Nov 2024 · 0 repositories · arXiv:2411.17301
-
Geometric Point Attention Transformer for 3D Shape Reassembly 26 Nov 2024 · 0 repositories · arXiv:2411.17788
-
"Give me the code" -- Log Analysis of First-Year CS Students' Interactions With GPT 26 Nov 2024 · 0 repositories · arXiv:2411.17855
-
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning 26 Nov 2024 · 1 repository · arXiv:2411.17100
-
Leveraging Large Language Models and Topic Modeling for Toxicity Classification 26 Nov 2024 · 1 repository · arXiv:2411.17876
-
MARVEL-40M+: Multi-Level Visual Elaboration for High-Fidelity Text-to-3D Content Creation 26 Nov 2024 · 2 repositories · arXiv:2411.17945Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
MAT: Multi-Range Attention Transformer for Efficient Image Super-Resolution 26 Nov 2024 · 1 repository · arXiv:2411.17214Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 3 pointer-only (licence)
-
MWFormer: Multi-Weather Image Restoration Using Degradation-Aware Transformers 26 Nov 2024 · 1 repository · arXiv:2411.17226
-
ΩSFormer: Dual-Modal Ω-like Super-Resolution Transformer Network for Cross-scale and High-accuracy Terraced Field Vectorization Extraction 26 Nov 2024 · 0 repositories · arXiv:2411.17088
-
Pretrained LLM Adapted with LoRA as a Decision Transformer for Offline RL in Quantitative Trading 26 Nov 2024 · 1 repository · arXiv:2411.17900
-
Push the Limit of Multi-modal Emotion Recognition by Prompting LLMs with Receptive-Field-Aware Attention Weighting 26 Nov 2024 · 0 repositories · arXiv:2411.17674
-
SCASeg: Strip Cross-Attention for Efficient Semantic Segmentation 26 Nov 2024 · 0 repositories · arXiv:2411.17061
-
TED-VITON: Transformer-Empowered Diffusion Models for Virtual Try-On 26 Nov 2024 · 1 repository · arXiv:2411.17017
-
TinyViM: Frequency Decoupling for Tiny Hybrid Vision Mamba 26 Nov 2024 · 1 repository · arXiv:2411.17473
-
What Differentiates Educational Literature? A Multimodal Fusion Approach of Transformers and Computational Linguistics 26 Nov 2024 · 0 repositories · arXiv:2411.17593
-
Can AI grade your essays? A comparative analysis of large language models and teacher ratings in multidimensional essay scoring 25 Nov 2024 · 0 repositories · arXiv:2411.16337
-
CATP-LLM: Empowering Large Language Models for Cost-Aware Tool Planning 25 Nov 2024 · 0 repositories · arXiv:2411.16313
-
CMAViT: Integrating Climate, Managment, and Remote Sensing Data for Crop Yield Estimation with Multimodel Vision Transformers 25 Nov 2024 · 0 repositories · arXiv:2411.16989
-
DF-GNN: Dynamic Fusion Framework for Attention Graph Neural Networks on GPUs 25 Nov 2024 · 1 repository · arXiv:2411.16127
-
Enhancing Answer Reliability Through Inter-Model Consensus of Large Language Models 25 Nov 2024 · 0 repositories · arXiv:2411.16797
-
Enhancing Fluorescence Lifetime Parameter Estimation Accuracy with Differential Transformer Based Deep Learning Model Incorporating Pixelwise Instrument Response Function 25 Nov 2024 · 0 repositories · arXiv:2411.16896
-
NormXLogit: The Head-on-Top Never Lies 25 Nov 2024 · 0 repositories · arXiv:2411.16252
-
Scaling Spike-driven Transformer with Efficient Spike Firing Approximation Training 25 Nov 2024 · 1 repository · arXiv:2411.16061Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Soft-TransFormers for Continual Learning 25 Nov 2024 · 1 repository · arXiv:2411.16073
-
Solaris: A Foundation Model of the Sun 25 Nov 2024 · 0 repositories · arXiv:2411.16339
-
Swin fMRI Transformer Predicts Early Neurodevelopmental Outcomes from Neonatal fMRI 25 Nov 2024 · 0 repositories · arXiv:2412.07783
-
Tree Transformers are an Ineffective Model of Syntactic Constituency 25 Nov 2024 · 0 repositories · arXiv:2411.16993
-
UltraSam: A Foundation Model for Ultrasound using Large Open-Access Segmentation Datasets 25 Nov 2024 · 1 repository · arXiv:2411.16222
-
VQ-SGen: A Vector Quantized Stroke Representation for Creative Sketch Generation 25 Nov 2024 · 0 repositories · arXiv:2411.16446
-
A Method for Building Large Language Models with Predefined KV Cache Capacity 24 Nov 2024 · 0 repositories · arXiv:2411.15785
-
Fixing the Perspective: A Critical Examination of Zero-1-to-3 24 Nov 2024 · 0 repositories · arXiv:2411.15706
-
Investigating Factuality in Long-Form Text Generation: The Roles of Self-Known and Self-Unknown 24 Nov 2024 · 0 repositories · arXiv:2411.15993
-
LTCF-Net: A Transformer-Enhanced Dual-Channel Fourier Framework for Low-Light Image Restoration 24 Nov 2024 · 0 repositories · arXiv:2411.15740
-
Medical Slice Transformer: Improved Diagnosis and Explainability on 3D Medical Images with DINOv2 24 Nov 2024 · 1 repository · arXiv:2411.15802
-
Nimbus: Secure and Efficient Two-Party Inference for Transformers 24 Nov 2024 · 1 repository · arXiv:2411.15707Syntology official: harvested, nothing ran · 0 ran · 5 unverified (of 5 harvested samples)
-
Best of Both Worlds: Advantages of Hybrid Graph Sequence Models 23 Nov 2024 · 0 repositories · arXiv:2411.15671
-
TANGNN: a Concise, Scalable and Effective Graph Neural Networks with Top-m Attention Mechanism for Graph Representation Learning 23 Nov 2024 · 1 repository · arXiv:2411.15458
-
A Real-Time DETR Approach to Bangladesh Road Object Detection for Autonomous Vehicles 22 Nov 2024 · 0 repositories · arXiv:2411.15110
-
Don't Mesh with Me: Generating Constructive Solid Geometry Instead of Meshes by Fine-Tuning a Code-Generation LLM 22 Nov 2024 · 0 repositories · arXiv:2411.15279
-
ElastiFormer: Learned Redundancy Reduction in Transformer via Self-Distillation 22 Nov 2024 · 0 repositories · arXiv:2411.15281
-
AI Foundation Models for Wearable Movement Data in Mental Health Research 22 Nov 2024 · 1 repository · arXiv:2411.15240
-
Multiset Transformer: Advancing Representation Learning in Persistence Diagrams 22 Nov 2024 · 1 repository · arXiv:2411.14662
-
OminiControl: Minimal and Universal Control for Diffusion Transformer 22 Nov 2024 · 2 repositories · arXiv:2411.15098Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Point Cloud Understanding via Attention-Driven Contrastive Learning 22 Nov 2024 · 0 repositories · arXiv:2411.14744
-
Purrfessor: A Fine-tuned Multimodal LLaVA Diet Health Chatbot 22 Nov 2024 · 0 repositories · arXiv:2411.14925
-
RED: Effective Trajectory Representation Learning with Comprehensive Information 22 Nov 2024 · 0 repositories · arXiv:2411.15096
-
Resolution-Agnostic Transformer-based Climate Downscaling 22 Nov 2024 · 0 repositories · arXiv:2411.14774
-
ScribeAgent: Towards Specialized Web Agents Using Production-Scale Workflow Data 22 Nov 2024 · 1 repository · arXiv:2411.15004Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
When Spatial meets Temporal in Action Recognition 22 Nov 2024 · 0 repositories · arXiv:2411.15284
-
Benchmarking GPT-4 against Human Translators: A Comprehensive Evaluation Across Languages, Domains, and Expertise Levels 21 Nov 2024 · 1 repository · arXiv:2411.13775Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Explaining GPT-4's Schema of Depression Using Machine Behavior Analysis 21 Nov 2024 · 0 repositories · arXiv:2411.13800
-
Generative Fuzzy System for Sequence Generation 21 Nov 2024 · 0 repositories · arXiv:2411.13867
-
Global and Local Attention-Based Transformer for Hyperspectral Image Change Detection 21 Nov 2024 · 1 repository · arXiv:2411.14109
-
GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI 21 Nov 2024 · 1 repository · arXiv:2411.14522Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
Learning from "Silly" Questions Improves Large Language Models, But Only Slightly 21 Nov 2024 · 0 repositories · arXiv:2411.14121
-
Parameter Efficient Mamba Tuning via Projector-targeted Diagonal-centric Linear Transformation 21 Nov 2024 · 0 repositories · arXiv:2411.15224
-
Stable Flow: Vital Layers for Training-Free Image Editing 21 Nov 2024 · 1 repository · arXiv:2411.14430Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Understanding World or Predicting Future? A Comprehensive Survey of World Models 21 Nov 2024 · 0 repositories · arXiv:2411.14499
-
BIPro: Zero-shot Chinese Poem Generation via Block Inverse Prompting Constrained Generation Framework 20 Nov 2024 · 0 repositories · arXiv:2411.13237
-
DrugGen: Advancing Drug Discovery with Large Language Models and Reinforcement Learning Feedback 20 Nov 2024 · 4 repositories · arXiv:2411.14157
-
Exploring Large Language Models for Climate Forecasting 20 Nov 2024 · 0 repositories · arXiv:2411.13724
-
Scaling Laws for Online Advertisement Retrieval 20 Nov 2024 · 0 repositories · arXiv:2411.13322
-
The Impossible Test: A 2024 Unsolvable Dataset and A Chance for an AGI Quiz 20 Nov 2024 · 0 repositories · arXiv:2411.14486
-
Transformers with Sparse Attention for Granger Causality 20 Nov 2024 · 0 repositories · arXiv:2411.13264
-
Comparing Prior and Learned Time Representations in Transformer Models of Timeseries 19 Nov 2024 · 0 repositories · arXiv:2411.12476
-
Evaluating Tokenizer Performance of Large Language Models Across Official Indian Languages 19 Nov 2024 · 0 repositories · arXiv:2411.12240
-
Faster Multi-GPU Training with PPLL: A Pipeline Parallelism Framework Leveraging Local Learning 19 Nov 2024 · 0 repositories · arXiv:2411.12780
-
Multi-Grained Preference Enhanced Transformer for Multi-Behavior Sequential Recommendation 19 Nov 2024 · 1 repository · arXiv:2411.12179
-
Residual Vision Transformer (ResViT) Based Self-Supervised Learning Model for Brain Tumor Classification 19 Nov 2024 · 0 repositories · arXiv:2411.12874
-
Transformer Neural Processes -- Kernel Regression 19 Nov 2024 · 0 repositories · arXiv:2411.12502
-
Ultra-Sparse Memory Network 19 Nov 2024 · 0 repositories · arXiv:2411.12364
-
Advacheck at GenAI Detection Task 1: AI Detection Powered by Domain-Aware Multi-Tasking 18 Nov 2024 · 1 repository · arXiv:2411.11736
-
Chapter 7 Review of Data-Driven Generative AI Models for Knowledge Extraction from Scientific Literature in Healthcare 18 Nov 2024 · 0 repositories · arXiv:2411.11635
-
Enhancing Decision Transformer with Diffusion-Based Trajectory Branch Generation 18 Nov 2024 · 0 repositories · arXiv:2411.11327
-
FCC: Fully Connected Correlation for Few-Shot Segmentation 18 Nov 2024 · 0 repositories · arXiv:2411.11917
-
In-Situ Melt Pool Characterization via Thermal Imaging for Defect Detection in Directed Energy Deposition Using Vision Transformers 18 Nov 2024 · 0 repositories · arXiv:2411.12028
-
LaVin-DiT: Large Vision Diffusion Transformer 18 Nov 2024 · 0 repositories · arXiv:2411.11505
-
LiTformer: Efficient Modeling and Analysis of High-Speed Link Transmitters Using Non-Autoregressive Transformer 18 Nov 2024 · 0 repositories · arXiv:2411.11699
-
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback 18 Nov 2024 · 1 repository · arXiv:2412.03578Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Popular LLMs Amplify Race and Gender Disparities in Human Mobility 18 Nov 2024 · 0 repositories · arXiv:2411.14469
-
ST-Tree with Interpretability for Multivariate Time Series Classification 18 Nov 2024 · 0 repositories · arXiv:2411.11620
-
TimeFormer: Capturing Temporal Relationships of Deformable 3D Gaussians for Robust Reconstruction 18 Nov 2024 · 1 repository · arXiv:2411.11941
-
Unveiling the Inflexibility of Adaptive Embedding in Traffic Forecasting 18 Nov 2024 · 1 repository · arXiv:2411.11448
-
Exploiting VLM Localizability and Semantics for Open Vocabulary Action Detection 17 Nov 2024 · 1 repository · arXiv:2411.10922
-
IVE: Enhanced Probabilistic Forecasting of Intraday Volume Ratio with Transformers 17 Nov 2024 · 0 repositories · arXiv:2411.10956
-
Knowledge-enhanced Transformer for Multivariate Long Sequence Time-series Forecasting 17 Nov 2024 · 0 repositories · arXiv:2411.11046
-
A Wearable Gait Monitoring System for 17 Gait Parameters Based on Computer Vision 16 Nov 2024 · 0 repositories · arXiv:2411.10739
-
AllRestorer: All-in-One Transformer for Image Restoration under Composite Degradations 16 Nov 2024 · 0 repositories · arXiv:2411.10708
-
Bag of Design Choices for Inference of High-Resolution Masked Generative Transformer 16 Nov 2024 · 1 repository · arXiv:2411.10781
-
IntentGPT: Few-shot Intent Discovery with Large Language Models 16 Nov 2024 · 0 repositories · arXiv:2411.10670
-
MetaLA: Unified Optimal Linear Approximation to Softmax Attention Map 16 Nov 2024 · 1 repository · arXiv:2411.10741Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
MpoxVLM: A Vision-Language Model for Diagnosing Skin Lesions from Mpox Virus Infection 16 Nov 2024 · 1 repository · arXiv:2411.10888
-
TDSM: Triplet Diffusion for Skeleton-Text Matching in Zero-Shot Action Recognition 16 Nov 2024 · 1 repository · arXiv:2411.10745
-
A Multi-Scale Spatial-Temporal Network for Wireless Video Transmission 15 Nov 2024 · 0 repositories · arXiv:2411.09936
-
Building 6G Radio Foundation Models with Transformer Architectures 15 Nov 2024 · 0 repositories · arXiv:2411.09996