Methods › Natural Language Processing › Autoregressive Transformers › Transformer › Papers, page 67
Transformer
Papers archive 2025-07-28
archive papers tagged: 13,999 · with a code link: 6,572 · where Syntology ran a sample: 2,248 (1,919 with a run with no instrument failure, 329 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,248 of 13,999 tagged: 1,919 with a run with no instrument failure, 329 where every run was a failure of Syntology's instrument)
Page 67 of 140: papers 6,601 to 6,700 of 13,999, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Evaluating, Understanding, and Improving Constrained Text Generation for Large Language Models 25 Oct 2023 · 0 repositories · arXiv:2310.16343
-
A No-Reference Quality Assessment Method for Digital Human Head 25 Oct 2023 · 0 repositories · arXiv:2310.16732
-
An Early Evaluation of GPT-4V(ision) 25 Oct 2023 · 1 repository · arXiv:2310.16534
-
Can GPT models Follow Human Summarization Guidelines? Evaluating ChatGPT and GPT-4 for Dialogue Summarization 25 Oct 2023 · 0 repositories · arXiv:2310.16810
-
CLEX: Continuous Length Extrapolation for Large Language Models 25 Oct 2023 · 1 repository · arXiv:2310.16450Syntology official (archive's flag): 9 ran · 9 ran (of which 2 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
GraFT: Gradual Fusion Transformer for Multimodal Re-Identification 25 Oct 2023 · 0 repositories · arXiv:2310.16856
-
Is ChatGPT a Good Multi-Party Conversation Solver? 25 Oct 2023 · 1 repository · arXiv:2310.16301
-
LLM-FP4: 4-Bit Floating-Point Quantized Transformers 25 Oct 2023 · 1 repository · arXiv:2310.16836Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
LLM Performance Predictors are good initializers for Architecture Search 25 Oct 2023 · 1 repository · arXiv:2310.16712Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Modality-Agnostic Self-Supervised Learning with Meta-Learned Masked Auto-Encoder 25 Oct 2023 · 1 repository · arXiv:2310.16318Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
netFound: Foundation Model for Network Security 25 Oct 2023 · 1 repository · arXiv:2310.17025
-
OccuQuest: Mitigating Occupational Bias for Inclusive Large Language Models 25 Oct 2023 · 1 repository · arXiv:2310.16517
-
Prompt-Driven Building Footprint Extraction in Aerial Images with Offset-Building Model 25 Oct 2023 · 0 repositories · arXiv:2310.16717
-
SMURF-THP: Score Matching-based UnceRtainty quantiFication for Transformer Hawkes Process 25 Oct 2023 · 1 repository · arXiv:2310.16336
-
SuperHF: Supervised Iterative Learning from Human Feedback 25 Oct 2023 · 1 repository · arXiv:2310.16763Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
TransPose: 6D Object Pose Estimation with Geometry-Aware Transformer 25 Oct 2023 · 0 repositories · arXiv:2310.16279
-
Using GPT-4 to Augment Unbalanced Data for Automatic Scoring 25 Oct 2023 · 0 repositories · arXiv:2310.18365
-
Confounder Balancing in Adversarial Domain Adaptation for Pre-Trained Large Models Fine-Tuning 24 Oct 2023 · 0 repositories · arXiv:2310.16062
-
Decoupled DETR: Spatially Disentangling Localization and Classification for Improved End-to-End Object Detection 24 Oct 2023 · 0 repositories · arXiv:2310.15955
-
Dynamic Convolutional Neural Networks as Efficient Pre-trained Audio Models 24 Oct 2023 · 1 repository · arXiv:2310.15648
-
Mixture of Tokens: Continuous MoE through Cross-Example Aggregation 24 Oct 2023 · 1 repository · arXiv:2310.15961
-
MuSR: Testing the Limits of Chain-of-thought with Multistep Soft Reasoning 24 Oct 2023 · 3 repositories · arXiv:2310.16049Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
NoteChat: A Dataset of Synthetic Doctor-Patient Conversations Conditioned on Clinical Notes 24 Oct 2023 · 1 repository · arXiv:2310.15959
-
Octopus: A Multitask Model and Toolkit for Arabic Natural Language Generation 24 Oct 2023 · 0 repositories · arXiv:2310.16127
-
Practical Computational Power of Linear Transformers and Their Recurrent and Self-Referential Extensions 24 Oct 2023 · 1 repository · arXiv:2310.16076Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
TRAMS: Training-free Memory Selection for Long-range Language Modeling 24 Oct 2023 · 1 repository · arXiv:2310.15494Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
UI Layout Generation with LLMs Guided by UI Grammar 24 Oct 2023 · 0 repositories · arXiv:2310.15455
-
What Algorithms can Transformers Learn? A Study in Length Generalization 24 Oct 2023 · 0 repositories · arXiv:2310.16028
-
AlpaCare:Instruction-tuned Large Language Models for Medical Application 23 Oct 2023 · 1 repository · arXiv:2310.14558Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Analyzing Multilingual Competency of LLMs in Multi-Turn Instruction Following: A Case Study of Arabic 23 Oct 2023 · 0 repositories · arXiv:2310.14819
-
Branch-Solve-Merge Improves Large Language Model Evaluation and Generation 23 Oct 2023 · 0 repositories · arXiv:2310.15123
-
Calibration of Time-Series Forecasting: Detecting and Adapting Context-Driven Distribution Shift 23 Oct 2023 · 2 repositories · arXiv:2310.14838Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
Causal Inference Using LLM-Guided Discovery 23 Oct 2023 · 0 repositories · arXiv:2310.15117
-
Evaluating Spatial Understanding of Large Language Models 23 Oct 2023 · 1 repository · arXiv:2310.14540Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Evaluating the Knowledge Base Completion Potential of GPT 23 Oct 2023 · 0 repositories · arXiv:2310.14771
-
Exploring the Boundaries of GPT-4 in Radiology 23 Oct 2023 · 0 repositories · arXiv:2310.14573
-
Generative Pre-trained Transformer for Vietnamese Community-based COVID-19 Question Answering 23 Oct 2023 · 0 repositories · arXiv:2310.14602
-
GPT-4 as an Effective Zero-Shot Evaluator for Scientific Figure Captions 23 Oct 2023 · 0 repositories · arXiv:2310.15405
-
Hallucination Detection for Grounded Instruction Generation 23 Oct 2023 · 0 repositories · arXiv:2310.15319
-
InstructExcel: A Benchmark for Natural Language Instruction in Excel 23 Oct 2023 · 0 repositories · arXiv:2310.14495
-
Large Language Models can Share Images, Too! 23 Oct 2023 · 2 repositories · arXiv:2310.14804
-
LINC: A Neurosymbolic Approach for Logical Reasoning by Combining Language Models with First-Order Logic Provers 23 Oct 2023 · 1 repository · arXiv:2310.15164Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Location-Aware Visual Question Generation with Lightweight Models 23 Oct 2023 · 1 repository · arXiv:2310.15129
-
Non-autoregressive Streaming Transformer for Simultaneous Translation 23 Oct 2023 · 1 repository · arXiv:2310.14883Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
P2AT: Pyramid Pooling Axial Transformer for Real-time Semantic Segmentation 23 Oct 2023 · 1 repository · arXiv:2310.15025
-
PartialFormer: Modeling Part Instead of Whole for Machine Translation 23 Oct 2023 · 1 repository · arXiv:2310.14921
-
SLOG: A Structural Generalization Benchmark for Semantic Parsing 23 Oct 2023 · 1 repository · arXiv:2310.15040Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
TeleQnA: A Benchmark Dataset to Assess Large Language Models Telecommunications Knowledge 23 Oct 2023 · 1 repository · arXiv:2310.15051
-
Text2Topic: Multi-Label Text Classification System for Efficient Topic Detection in User Generated Content with Zero-Shot Capabilities 23 Oct 2023 · 1 repository · arXiv:2310.14817
-
Unleashing the potential of prompt engineering for large language models 23 Oct 2023 · 0 repositories · arXiv:2310.14735
-
A Pytorch Reproduction of Masked Generative Image Transformer 22 Oct 2023 · 1 repository · arXiv:2310.14400
-
Affine-Consistent Transformer for Multi-Class Cell Nuclei Detection 22 Oct 2023 · 1 repository · arXiv:2310.14154
-
An overview of text-to-speech systems and media applications 22 Oct 2023 · 0 repositories · arXiv:2310.14301
-
ConViViT -- A Deep Neural Network Combining Convolutions and Factorized Self-Attention for Human Activity Recognition 22 Oct 2023 · 0 repositories · arXiv:2310.14416
-
CXR-LLAVA: a multimodal large language model for interpreting chest X-ray images 22 Oct 2023 · 1 repository · arXiv:2310.18341Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Hierarchical Vector Quantized Transformer for Multi-class Unsupervised Anomaly Detection 22 Oct 2023 · 1 repository · arXiv:2310.14228Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Large Language Models are biased to overestimate profoundness 22 Oct 2023 · 1 repository · arXiv:2310.14422
-
Manifold-Preserving Transformers are Effective for Short-Long Range Encoding 22 Oct 2023 · 1 repository · arXiv:2310.14206
-
MMTF-DES: A Fusion of Multimodal Transformer Models for Desire, Emotion, and Sentiment Analysis of Social Media Data 22 Oct 2023 · 0 repositories · arXiv:2310.14143
-
UniMAP: Universal SMILES-Graph Representation Learning 22 Oct 2023 · 1 repository · arXiv:2310.14216
-
Visual-Attribute Prompt Learning for Progressive Mild Cognitive Impairment Prediction 22 Oct 2023 · 1 repository · arXiv:2310.14158
-
Concept-based Anomaly Detection in Retail Stores for Automatic Correction using Mobile Robots 21 Oct 2023 · 0 repositories · arXiv:2310.14063
-
Exploring Driving Behavior for Autonomous Vehicles Based on Gramian Angular Field Vision Transformer 21 Oct 2023 · 0 repositories · arXiv:2310.13906
-
GEMBA-MQM: Detecting Translation Quality Error Spans with GPT-4 21 Oct 2023 · 1 repository · arXiv:2310.13988Syntology 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Small Language Models Fine-tuned to Coordinate Larger Language Models improve Complex Reasoning 21 Oct 2023 · 1 repository · arXiv:2310.18338Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
AllTogether: Investigating the Efficacy of Spliced Prompt for Web Navigation using Large Language Models 20 Oct 2023 · 0 repositories · arXiv:2310.18331
-
Ask Language Model to Clean Your Noisy Translation Data 20 Oct 2023 · 0 repositories · arXiv:2310.13469
-
Auxiliary Features-Guided Super Resolution for Monte Carlo Rendering 20 Oct 2023 · 0 repositories · arXiv:2310.13235
-
BotChat: Evaluating LLMs' Capabilities of Having Multi-Turn Dialogues 20 Oct 2023 · 1 repository · arXiv:2310.13650Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Bridging the Gap between Synthetic and Authentic Images for Multimodal Machine Translation 20 Oct 2023 · 1 repository · arXiv:2310.13361
-
Cache me if you Can: an Online Cost-aware Teacher-Student framework to Reduce the Calls to Large Language Models 20 Oct 2023 · 1 repository · arXiv:2310.13395Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
She had Cobalt Blue Eyes: Prompt Testing to Create Aligned and Sustainable Language Models 20 Oct 2023 · 0 repositories · arXiv:2310.18333
-
Evaluation Metrics in the Era of GPT-4: Reliably Evaluating Large Language Models on Sequence to Sequence Tasks 20 Oct 2023 · 1 repository · arXiv:2310.13800Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
FMRT: Learning Accurate Feature Matching with Reconciliatory Transformer 20 Oct 2023 · 0 repositories · arXiv:2310.13605
-
MoqaGPT : Zero-Shot Multi-modal Open-domain Question Answering with Large Language Model 20 Oct 2023 · 1 repository · arXiv:2310.13265
-
Plausibility Processing in Transformer Language Models: Focusing on the Role of Attention Heads in GPT 20 Oct 2023 · 1 repository · arXiv:2310.13824
-
POTLoc: Pseudo-Label Oriented Transformer for Point-Supervised Temporal Action Localization 20 Oct 2023 · 0 repositories · arXiv:2310.13585
-
Skin Lesion Segmentation Improved by Transformer-based Networks with Inter-scale Dependency Modeling 20 Oct 2023 · 1 repository · arXiv:2310.13604
-
The Impact of Performance Expectancy, Workload, Risk, and Satisfaction on Trust in ChatGPT: Cross-sectional Survey Analysis 20 Oct 2023 · 0 repositories · arXiv:2311.05632
-
The Perils & Promises of Fact-checking with Large Language Models 20 Oct 2023 · 0 repositories · arXiv:2310.13549
-
Tuna: Instruction Tuning using Feedback from Large Language Models 20 Oct 2023 · 1 repository · arXiv:2310.13385
-
2D-3D Interlaced Transformer for Point Cloud Segmentation with Scene-Level Supervision 19 Oct 2023 · 0 repositories · arXiv:2310.12817
-
AgentTuning: Enabling Generalized Agent Abilities for LLMs 19 Oct 2023 · 1 repository · arXiv:2310.12823Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ReEval: Automatic Hallucination Evaluation for Retrieval-Augmented Large Language Models via Transferable Adversarial Attacks 19 Oct 2023 · 0 repositories · arXiv:2310.12516
-
AutoMix: Automatically Mixing Language Models 19 Oct 2023 · 1 repository · arXiv:2310.12963Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
AVTENet: Audio-Visual Transformer-based Ensemble Network Exploiting Multiple Experts for Video Deepfake Detection 19 Oct 2023 · 0 repositories · arXiv:2310.13103
-
Character-level Chinese Backpack Language Models 19 Oct 2023 · 1 repository · arXiv:2310.12751
-
DA-TransUNet: Integrating Spatial and Channel Dual Attention with Transformer U-Net for Medical Image Segmentation 19 Oct 2023 · 1 repository · arXiv:2310.12570
-
Eureka: Human-Level Reward Design via Coding Large Language Models 19 Oct 2023 · 1 repository · arXiv:2310.12931Syntology official (archive's flag): 8 ran · 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
Experimental Narratives: A Comparison of Human Crowdsourced Storytelling and AI Storytelling 19 Oct 2023 · 0 repositories · arXiv:2310.12902
-
LeTFuser: Light-weight End-to-end Transformer-Based Sensor Fusion for Autonomous Driving with Multi-Task Learning 19 Oct 2023 · 1 repository · arXiv:2310.13135
-
Minimalist and High-Performance Semantic Segmentation with Plain Vision Transformers 19 Oct 2023 · 1 repository · arXiv:2310.12755
-
Multi-granularity Backprojection Transformer for Remote Sensing Image Super-Resolution 19 Oct 2023 · 0 repositories · arXiv:2310.12507
-
Non-Autoregressive Sentence Ordering 19 Oct 2023 · 1 repository · arXiv:2310.12640
-
Not All Countries Celebrate Thanksgiving: On the Cultural Dominance in Large Language Models 19 Oct 2023 · 0 repositories · arXiv:2310.12481
-
Predicting Ovarian Cancer Treatment Response in Histopathology using Hierarchical Vision Transformers and Multiple Instance Learning 19 Oct 2023 · 1 repository · arXiv:2310.12866
-
ExtractGPT: Exploring the Potential of Large Language Models for Product Attribute Value Extraction 19 Oct 2023 · 1 repository · arXiv:2310.12537
-
Real-Time Motion Prediction via Heterogeneous Polyline Transformer with Relative Pose Encoding 19 Oct 2023 · 2 repositories · arXiv:2310.12970Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Sequence Length Independent Norm-Based Generalization Bounds for Transformers 19 Oct 2023 · 1 repository · arXiv:2310.13088Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
The Foundation Model Transparency Index 19 Oct 2023 · 1 repository · arXiv:2310.12941