Methods › Natural Language Processing › Autoregressive Transformers › Transformer › Papers, page 5
Transformer
Papers archive 2025-07-28
archive papers tagged: 13,999 · with a code link: 6,572 · where Syntology ran a sample: 2,248 (1,919 with a run with no instrument failure, 329 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,248 of 13,999 tagged: 1,919 with a run with no instrument failure, 329 where every run was a failure of Syntology's instrument)
Page 5 of 140: papers 401 to 500 of 13,999, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
A Deep Learning-Driven Inhalation Injury Grading Assistant Using Bronchoscopy Images 13 May 2025 · 0 repositories · arXiv:2505.08517
-
A Head to Predict and a Head to Question: Pre-trained Uncertainty Quantification Heads for Hallucination Detection in LLM Outputs 13 May 2025 · 1 repository · arXiv:2505.08200Syntology 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
A suite of LMs comprehend puzzle statements as well as humans 13 May 2025 · 0 repositories · arXiv:2505.08996
-
AC-Reason: Towards Theory-Guided Actual Causality Reasoning with Large Language Models 13 May 2025 · 1 repository · arXiv:2505.08750
-
CNN and ViT Efficiency Study on Tiny ImageNet and DermaMNIST Datasets 13 May 2025 · 0 repositories · arXiv:2505.08259
-
Constrained Edge AI Deployment: Fine-Tuning vs Distillation for LLM Compression 13 May 2025 · 0 repositories · arXiv:2505.18166
-
For GPT-4 as with Humans: Information Structure Predicts Acceptability of Long-Distance Dependencies 13 May 2025 · 0 repositories · arXiv:2505.09005
-
HealthBench: Evaluating Large Language Models Towards Improved Human Health 13 May 2025 · 1 repository · arXiv:2505.08775Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning? 13 May 2025 · 1 repository · arXiv:2505.08468
-
Lost in Transmission: When and Why LLMs Fail to Reason Globally 13 May 2025 · 0 repositories · arXiv:2505.08140Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
OpenThinkIMG: Learning to Think with Images via Visual Tool Reinforcement Learning 13 May 2025 · 1 repository · arXiv:2505.08617
-
SAR-GTR: Attributed Scattering Information Guided SAR Graph Transformer Recognition Algorithm 13 May 2025 · 0 repositories · arXiv:2505.08547
-
Structural-Temporal Coupling Anomaly Detection with Dynamic Graph Transformer 13 May 2025 · 1 repository · arXiv:2505.08330
-
A Generative Re-ranking Model for List-level Multi-objective Optimization at Taobao 12 May 2025 · 0 repositories · arXiv:2505.07197
-
AIS Data-Driven Maritime Monitoring Based on Transformer: A Comprehensive Review 12 May 2025 · 1 repository · arXiv:2505.07374
-
An Extra RMSNorm is All You Need for Fine Tuning to 1.58 Bits 12 May 2025 · 0 repositories · arXiv:2505.08823
-
Fused3S: Fast Sparse Attention on Tensor Cores 12 May 2025 · 1 repository · arXiv:2505.08098
-
Generative Pre-trained Autoregressive Diffusion Transformer 12 May 2025 · 0 repositories · arXiv:2505.07344
-
Hybrid Spiking Vision Transformer for Object Detection with Event Cameras 12 May 2025 · 0 repositories · arXiv:2505.07715
-
LAMM-ViT: AI Face Detection via Layer-Aware Modulation of Region-Guided Attention 12 May 2025 · 0 repositories · arXiv:2505.07734
-
MAIS: Memory-Attention for Interactive Segmentation 12 May 2025 · 0 repositories · arXiv:2505.07511
-
Sleep Position Classification using Transfer Learning for Bed-based Pressure Sensors 12 May 2025 · 0 repositories · arXiv:2505.08111
-
Statistical CSI-Based Distributed Precoding Design for OFDM-Cooperative Multi-Satellite Systems 12 May 2025 · 0 repositories · arXiv:2505.08038
-
The Geography of Transportation Cybersecurity: Visitor Flows, Industry Clusters, and Spatial Dynamics 12 May 2025 · 0 repositories · arXiv:2505.08822
-
Topology-Guided Knowledge Distillation for Efficient Point Cloud Processing 12 May 2025 · 1 repository · arXiv:2505.08101
-
UMoE: Unifying Attention and FFN with Shared Experts 12 May 2025 · 0 repositories · arXiv:2505.07260
-
Image Classification Using a Diffusion Model as a Pre-Training Model 11 May 2025 · 0 repositories · arXiv:2505.06890
-
Matrix Is All You Need 11 May 2025 · 0 repositories · arXiv:2506.01966
-
NeuRN: Neuro-inspired Domain Generalization for Image Classification 11 May 2025 · 0 repositories · arXiv:2505.06881
-
Technical Report for ICRA 2025 GOOSE 2D Semantic Segmentation Challenge: Leveraging Color Shift Correction, RoPE-Swin Backbone, and Quantile-based Label Denoising Strategy for Robust Outdoor Scene Understanding 11 May 2025 · 0 repositories · arXiv:2505.06991
-
Boosting Neural Language Inference via Cascaded Interactive Reasoning 10 May 2025 · 0 repositories · arXiv:2505.06607
-
Probing In-Context Learning: Impact of Task Complexity and Model Architecture on Generalization and Efficiency 10 May 2025 · 1 repository · arXiv:2505.06475
-
QoS-Efficient Serving of Multiple Mixture-of-Expert LLMs Using Partial Runtime Reconfiguration 10 May 2025 · 0 repositories · arXiv:2505.06481
-
Underwater object detection in sonar imagery with detection transformer and Zero-shot neural architecture search 10 May 2025 · 0 repositories · arXiv:2505.06694
-
Utilizing LLMs to Investigate the Disputed Role of Evidence in Electronic Cigarette Health Policy Formation in Australia and the UK 10 May 2025 · 0 repositories · arXiv:2505.06782
-
xGen-small Technical Report 10 May 2025 · 0 repositories · arXiv:2505.06496
-
Accurate and Efficient Multivariate Time Series Forecasting via Offline Clustering 9 May 2025 · 0 repositories · arXiv:2505.05738
-
Camera Control at the Edge with Language Models for Scene Understanding 9 May 2025 · 0 repositories · arXiv:2505.06402
-
DFEN: Dual Feature Equalization Network for Medical Image Segmentation 9 May 2025 · 1 repository · arXiv:2505.05913
-
Graph Laplacian Wavelet Transformer via Learnable Spectral Decomposition 9 May 2025 · 0 repositories · arXiv:2505.07862
-
Healthy LLMs? Benchmarking LLM Knowledge of UK Government Public Health Information 9 May 2025 · 0 repositories · arXiv:2505.06046
-
Towards Robust Few-Shot Text Classification Using Transformer Architectures and Dual Loss Strategies 9 May 2025 · 0 repositories · arXiv:2505.06145
-
Turbo-ICL: In-Context Learning-Based Turbo Equalization 9 May 2025 · 0 repositories · arXiv:2505.06175
-
UniSymNet: A Unified Symbolic Network Guided by Transformer 9 May 2025 · 0 repositories · arXiv:2505.06091
-
AI Approaches to Qualitative and Quantitative News Analytics on NATO Unity 8 May 2025 · 0 repositories · arXiv:2505.06313
-
Benchmarking Vision, Language, & Action Models in Procedurally Generated, Open Ended Action Environments 8 May 2025 · 1 repository · arXiv:2505.05540
-
Cardioformer: Advancing AI in ECG Analysis with Multi-Granularity Patching and ResNet 8 May 2025 · 1 repository · arXiv:2505.05538
-
Performance Evaluation of Large Language Models in Bangla Consumer Health Query Summarization 8 May 2025 · 0 repositories · arXiv:2505.05070
-
Pro2SAM: Mask Prompt to SAM with Grid Points for Weakly Supervised Object Localization 8 May 2025 · 0 repositories · arXiv:2505.04905
-
Progressive Inertial Poser: Progressive Real-Time Kinematic Chain Estimation for 3D Full-Body Pose from Three IMU Sensors 8 May 2025 · 0 repositories · arXiv:2505.05336
-
SSH-Net: A Self-Supervised and Hybrid Network for Noisy Image Watermark Removal 8 May 2025 · 1 repository · arXiv:2505.05088
-
Trading Under Uncertainty: A Distribution-Based Strategy for Futures Markets Using FutureQuant Transformer 8 May 2025 · 0 repositories · arXiv:2505.05595
-
Balancing Accuracy, Calibration, and Efficiency in Active Learning with Vision Transformers Under Label Noise 7 May 2025 · 0 repositories · arXiv:2505.04375
-
DOTA: Deformable Optimized Transformer Architecture for End-to-End Text Recognition with Retrieval-Augmented Generation 7 May 2025 · 0 repositories · arXiv:2505.04175
-
HDiffTG: A Lightweight Hybrid Diffusion-Transformer-GCN Architecture for 3D Human Pose Estimation 7 May 2025 · 1 repository · arXiv:2505.04276
-
HiPerRAG: High-Performance Retrieval Augmented Generation for Scientific Insights 7 May 2025 · 0 repositories · arXiv:2505.04846
-
Image Restoration via Multi-domain Learning 7 May 2025 · 1 repository · arXiv:2505.05504
-
Lay-Your-Scene: Natural Scene Layout Generation with Diffusion Transformers 7 May 2025 · 0 repositories · arXiv:2505.04718
-
LLM-e Guess: Can LLMs Capabilities Advance Without Hardware Progress? 7 May 2025 · 1 repository · arXiv:2505.04075
-
M2Rec: Multi-scale Mamba for Efficient Sequential Recommendation 7 May 2025 · 0 repositories · arXiv:2505.04445
-
ORBIT-2: Scaling Exascale Vision Foundation Models for Weather and Climate Downscaling 7 May 2025 · 0 repositories · arXiv:2505.04802
-
Personalized Risks and Regulatory Strategies of Large Language Models in Digital Advertising 7 May 2025 · 0 repositories · arXiv:2505.04665
-
Pose Estimation for Intra-cardiac Echocardiography Catheter via AI-Based Anatomical Understanding 7 May 2025 · 0 repositories · arXiv:2505.07851
-
Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs 7 May 2025 · 0 repositories · arXiv:2505.04806
-
SwinLip: An Efficient Visual Speech Encoder for Lip Reading Using Swin Transformer 7 May 2025 · 0 repositories · arXiv:2505.04394
-
Theoretical Guarantees for LT-TTD: A Unified Transformer-based Architecture for Two-Level Ranking Systems 7 May 2025 · 0 repositories · arXiv:2505.04434
-
Image Recognition with Online Lightweight Vision Transformer: A Survey 6 May 2025 · 0 repositories · arXiv:2505.03113
-
MergeGuard: Efficient Thwarting of Trojan Attacks in Machine Learning Models 6 May 2025 · 1 repository · arXiv:2505.04015
-
Physics-inspired Energy Transition Neural Network for Sequence Learning 6 May 2025 · 0 repositories · arXiv:2505.03281
-
Rethinking Boundary Detection in Deep Learning-Based Medical Image Segmentation 6 May 2025 · 1 repository · arXiv:2505.04652
-
Large Language Model Partitioning for Low-Latency Inference at the Edge 5 May 2025 · 0 repositories · arXiv:2505.02533
-
LLM4FTS: Enhancing Large Language Models for Financial Time Series Prediction 5 May 2025 · 0 repositories · arXiv:2505.02880
-
Low-Loss Space in Neural Networks is Continuous and Fully Connected 5 May 2025 · 0 repositories · arXiv:2505.02604Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
Rapid yet accurate Tile-circuit and device modeling for Analog In-Memory Computing 5 May 2025 · 0 repositories · arXiv:2506.00004
-
SCFormer: Structured Channel-wise Transformer with Cumulative Historical State for Multivariate Time Series Forecasting 5 May 2025 · 1 repository · arXiv:2505.02655
-
T2S: High-resolution Time Series Generation with Text-to-Series Diffusion Models 5 May 2025 · 1 repository · arXiv:2505.02417
-
Voila: Voice-Language Foundation Models for Real-Time Autonomous Interaction and Voice Role-Play 5 May 2025 · 1 repository · arXiv:2505.02707Syntology official (archive's flag): 7 ran · 7 ran (of which 6 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples)
-
CASA: CNN Autoencoder-based Score Attention for Efficient Multivariate Long-term Time-series Forecasting 4 May 2025 · 1 repository · arXiv:2505.02011
-
DualReal: Adaptive Joint Training for Lossless Identity-Motion Fusion in Video Customization 4 May 2025 · 0 repositories · arXiv:2505.02192
-
Learning Local Causal World Models with State Space Models and Attention 4 May 2025 · 0 repositories · arXiv:2505.02074
-
LLM-OptiRA: LLM-Driven Optimization of Resource Allocation for Non-Convex Problems in Wireless Communications 4 May 2025 · 1 repository · arXiv:2505.02091
-
Local Herb Identification Using Transfer Learning: A CNN-Powered Mobile Application for Nepalese Flora 4 May 2025 · 0 repositories · arXiv:2505.02147
-
SEval-Ex: A Statement-Level Framework for Explainable Summarization Evaluation 4 May 2025 · 0 repositories · arXiv:2505.02235
-
Securing 5G and Beyond-Enabled UAV Networks: Resilience Through Multiagent Learning and Transformers Detection 3 May 2025 · 0 repositories · arXiv:2505.01885
-
Semantic Intelligence: Integrating GPT-4 with A Planning in Low-Cost Robotics 3 May 2025 · 0 repositories · arXiv:2505.01931
-
Toward Onboard AI-Enabled Solutions to Space Object Detection for Space Sustainability 3 May 2025 · 0 repositories · arXiv:2505.01650
-
3D Human Pose Estimation via Spatial Graph Order Attention and Temporal Body Aware Transformer 2 May 2025 · 1 repository · arXiv:2505.01003
-
A Self-Supervised Transformer for Unusable Shared Bike Detection 2 May 2025 · 0 repositories · arXiv:2505.00932
-
A Transformer-based Neural Architecture Search Method 2 May 2025 · 1 repository · arXiv:2505.01314
-
Asset Pricing in Pre-trained Transformer 2 May 2025 · 0 repositories · arXiv:2505.01575
-
Compact Recurrent Transformer with Persistent Memory 2 May 2025 · 0 repositories · arXiv:2505.00929
-
Enhancing SPARQL Query Rewriting for Complex Ontology Alignments 2 May 2025 · 0 repositories · arXiv:2505.01309
-
FalconWing: An Open-Source Platform for Ultra-Light Fixed-Wing Aircraft Research 2 May 2025 · 0 repositories · arXiv:2505.01383
-
FreCT: Frequency-augmented Convolutional Transformer for Robust Time Series Anomaly Detection 2 May 2025 · 0 repositories · arXiv:2505.00941
-
Good News for Script Kiddies? Evaluating Large Language Models for Automated Exploit Generation 2 May 2025 · 0 repositories · arXiv:2505.01065
-
Multimodal Transformers are Hierarchical Modal-wise Heterogeneous Graphs 2 May 2025 · 0 repositories · arXiv:2505.01068Syntology 0 ran · 2 unverified (of 2 harvested samples)
-
Token-free Models for Sarcasm Detection 2 May 2025 · 0 repositories · arXiv:2505.01006
-
Zero-Shot Document-Level Biomedical Relation Extraction via Scenario-based Prompt Design in Two-Stage with LLM 2 May 2025 · 0 repositories · arXiv:2505.01077
-
DARTer: Dynamic Adaptive Representation Tracker for Nighttime UAV Tracking 1 May 2025 · 0 repositories · arXiv:2505.00752
-
A Time-Series Data Augmentation Model through Diffusion and Transformer Integration 1 May 2025 · 0 repositories · arXiv:2505.03790