Methods › General › Fine-Tuning › Discriminative Fine-Tuning › Papers, page 15
Discriminative Fine-Tuning
Papers archive 2025-07-28
archive papers tagged: 1,990 · with a code link: 794 · where Syntology ran a sample: 271 (223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (271 of 1,990 tagged: 223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument)
Page 15 of 20: papers 1,401 to 1,500 of 1,990, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Can ChatGPT Replace Traditional KBQA Models? An In-depth Analysis of the Question Answering Performance of the GPT LLM Family 14 Mar 2023 · 2 repositories · arXiv:2303.07992
-
Learning Combinatorial Prompts for Universal Controllable Image Captioning 11 Mar 2023 · 0 repositories · arXiv:2303.06338
-
Algorithmic Ghost in the Research Shell: Large Language Models and Academic Knowledge Creation in Management Research 10 Mar 2023 · 0 repositories · arXiv:2303.07304
-
Large Language Models (GPT) Struggle to Answer Multiple-Choice Questions about Code 9 Mar 2023 · 0 repositories · arXiv:2303.08033
-
Stealing the Decoding Algorithms of Language Models 8 Mar 2023 · 1 repository · arXiv:2303.04729
-
A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT 7 Mar 2023 · 1 repository · arXiv:2303.04226
-
Towards Zero-Shot Functional Compositionality of Language Models 6 Mar 2023 · 1 repository · arXiv:2303.03103
-
Industry Risk Assessment via Hierarchical Financial Data Using Stock Market Sentiment Indicators 5 Mar 2023 · 0 repositories · arXiv:2303.02707
-
Zero-Shot Cross-Lingual Summarization via Large Language Models 28 Feb 2023 · 0 repositories · arXiv:2302.14229
-
Information-Restricted Neural Language Models Reveal Different Brain Regions' Sensitivity to Semantics, Syntax and Context 28 Feb 2023 · 1 repository · arXiv:2302.14389Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Large Language Models Are State-of-the-Art Evaluators of Translation Quality 28 Feb 2023 · 4 repositories · arXiv:2302.14520
-
Inseq: An Interpretability Toolkit for Sequence Generation Models 27 Feb 2023 · 2 repositories · arXiv:2302.13942
-
Fast Attention Requires Bounded Entries 26 Feb 2023 · 0 repositories · arXiv:2302.13214
-
Large-scale Multi-Modal Pre-trained Models: A Comprehensive Survey 20 Feb 2023 · 1 repository · arXiv:2302.10035
-
How Good Are GPT Models at Machine Translation? A Comprehensive Evaluation 18 Feb 2023 · 1 repository · arXiv:2302.09210Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Conveying the Predicted Future to Users: A Case Study of Story Plot Prediction 17 Feb 2023 · 1 repository · arXiv:2302.09122
-
PAC Prediction Sets for Large Language Models of Code 17 Feb 2023 · 1 repository · arXiv:2302.08703
-
Foundation Models for Natural Language Processing -- Pre-trained Language Models Integrating Media 16 Feb 2023 · 0 repositories · arXiv:2302.08575
-
Commonsense Reasoning for Conversational AI: A Survey of the State of the Art 15 Feb 2023 · 0 repositories · arXiv:2302.07926
-
Tree-Based Representation and Generation of Natural and Mathematical Language 15 Feb 2023 · 1 repository · arXiv:2302.07974
-
Diminished Diversity-of-Thought in a Standard Large Language Model 13 Feb 2023 · 0 repositories · arXiv:2302.07267
-
Academic Writing with GPT-3.5: Reflections on Practices, Efficacy and Transparency 12 Feb 2023 · 0 repositories · arXiv:2304.11079
-
Combat AI With AI: Counteract Machine-Generated Fake Restaurant Reviews on Social Media 10 Feb 2023 · 1 repository · arXiv:2302.07731
-
FairPy: A Toolkit for Evaluation of Prediction Biases and their Mitigation in Large Language Models 10 Feb 2023 · 1 repository · arXiv:2302.05508
-
The Wisdom of Hindsight Makes Language Models Better Instruction Followers 10 Feb 2023 · 1 repository · arXiv:2302.05206
-
Translating Natural Language to Planning Goals with Large-Language Models 10 Feb 2023 · 1 repository · arXiv:2302.05128
-
Better by you, better than me, chatgpt3 as writing assistance in students essays 9 Feb 2023 · 0 repositories · arXiv:2302.04536
-
Nationality Bias in Text Generation 5 Feb 2023 · 0 repositories · arXiv:2302.02463
-
Quantized Distributed Training of Large Models with Convergence Guarantees 5 Feb 2023 · 0 repositories · arXiv:2302.02390
-
REaLTabFormer: Generating Realistic Relational and Tabular Data using Transformers 4 Feb 2023 · 3 repositories · arXiv:2302.02041Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Semantic Coherence Markers for the Early Diagnosis of the Alzheimer Disease 2 Feb 2023 · 1 repository · arXiv:2302.01025
-
Large language models predict human sensory judgments across six modalities 2 Feb 2023 · 0 repositories · arXiv:2302.01308
-
Analyzing Leakage of Personally Identifiable Information in Language Models 1 Feb 2023 · 1 repository · arXiv:2302.00539
-
The Exploration of Knowledge-Preserving Prompts for Document Summarisation 27 Jan 2023 · 0 repositories · arXiv:2301.11719
-
Large Language Models Are Latent Variable Models: Explaining and Finding Good Demonstrations for In-Context Learning 27 Jan 2023 · 1 repository · arXiv:2301.11916Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
A Stability Analysis of Fine-Tuning a Pre-Trained Model 24 Jan 2023 · 0 repositories · arXiv:2301.09820
-
Audience-Centric Natural Language Generation via Style Infusion 24 Jan 2023 · 1 repository · arXiv:2301.10283
-
T2M-GPT: Generating Human Motion from Textual Descriptions with Discrete Representations 15 Jan 2023 · 1 repository · arXiv:2301.06052Syntology official (archive's flag): 5 ran · 5 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
GPT as Knowledge Worker: A Zero-Shot Evaluation of (AI)CPA Capabilities 11 Jan 2023 · 1 repository · arXiv:2301.04408
-
Automatic Generation of German Drama Texts Using Fine Tuned GPT-2 Models 8 Jan 2023 · 0 repositories · arXiv:2301.03119
-
Sequentially Controlled Text Generation 5 Jan 2023 · 0 repositories · arXiv:2301.02299
-
Generating Human Motion From Textual Descriptions With Discrete Representations 1 Jan 2023 · 0 repositories
-
Targeted Phishing Campaigns using Large Scale Language Models 30 Dec 2022 · 0 repositories · arXiv:2301.00665
-
GPT Takes the Bar Exam 29 Dec 2022 · 5 repositories · arXiv:2212.14402
-
TegFormer: Topic-to-Essay Generation with Good Topic Coverage and High Text Coherence 27 Dec 2022 · 0 repositories · arXiv:2212.13456
-
Benchmark for Uncertainty & Robustness in Self-Supervised Learning 23 Dec 2022 · 1 repository · arXiv:2212.12411
-
Why Does Surprisal From Larger Transformer-Based Language Models Provide a Poorer Fit to Human Reading Times? 23 Dec 2022 · 0 repositories · arXiv:2212.12131
-
Entropy- and Distance-Based Predictors From GPT-2 Attention Patterns Predict Reading Times Over and Above GPT-2 Surprisal 21 Dec 2022 · 1 repository · arXiv:2212.11185Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 8 pointer-only (licence)
-
KL Regularized Normalization Framework for Low Resource Tasks 21 Dec 2022 · 0 repositories · arXiv:2212.11275
-
ByGPT5: End-to-End Style-conditioned Poetry Generation with Token-free Language Models 20 Dec 2022 · 1 repository · arXiv:2212.10474Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
True Detective: A Deep Abductive Reasoning Benchmark Undoable for GPT-3 and Challenging for GPT-4 20 Dec 2022 · 0 repositories · arXiv:2212.10114
-
Why Can GPT Learn In-Context? Language Models Implicitly Perform Gradient Descent as Meta-Optimizers 20 Dec 2022 · 1 repository · arXiv:2212.10559
-
Reasoning with Language Model Prompting: A Survey 19 Dec 2022 · 2 repositories · arXiv:2212.09597
-
The case for 4-bit precision: k-bit Inference Scaling Laws 19 Dec 2022 · 1 repository · arXiv:2212.09720
-
MURMUR: Modular Multi-Step Reasoning for Semi-Structured Data-to-Text Generation 16 Dec 2022 · 0 repositories · arXiv:2212.08607
-
Elixir: Train a Large Language Model on a Small GPU Cluster 10 Dec 2022 · 2 repositories · arXiv:2212.05339
-
The Turing Deception 9 Dec 2022 · 0 repositories · arXiv:2212.06721
-
TRBLLmaker -- Transformer Reads Between Lyrics Lines maker 9 Dec 2022 · 0 repositories · arXiv:2212.04917
-
Explain to me like I am five -- Sentence Simplification Using Transformers 8 Dec 2022 · 1 repository · arXiv:2212.04595
-
Modern French Poetry Generation with RoBERTa and GPT-2 6 Dec 2022 · 0 repositories · arXiv:2212.02911
-
Audio-Driven Co-Speech Gesture Video Generation 5 Dec 2022 · 0 repositories · arXiv:2212.02350
-
Automatic Generation of Factual News Headlines in Finnish 5 Dec 2022 · 0 repositories · arXiv:2212.02170
-
Exploring the Limits of Differentially Private Deep Learning with Group-wise Clipping 3 Dec 2022 · 0 repositories · arXiv:2212.01539
-
Distilling Reasoning Capabilities into Smaller Language Models 1 Dec 2022 · 1 repository · arXiv:2212.00193
-
Quadapter: Adapter for GPT-2 Quantization 30 Nov 2022 · 0 repositories · arXiv:2211.16912
-
Outfit Generation and Recommendation -- An Experimental Study 29 Nov 2022 · 0 repositories · arXiv:2211.16353
-
GPT-Neo for commonsense reasoning -- a theoretical and practical lens 28 Nov 2022 · 1 repository · arXiv:2211.15593
-
Scientific and Creative Analogies in Pretrained Language Models 28 Nov 2022 · 2 repositories · arXiv:2211.15268
-
Understanding BLOOM: An empirical study on diverse NLP tasks 27 Nov 2022 · 0 repositories · arXiv:2211.14865
-
Exploring the Efficacy of Pre-trained Checkpoints in Text-to-Music Generation Task 21 Nov 2022 · 2 repositories · arXiv:2211.11216
-
PointCLIP V2: Prompting CLIP and GPT for Powerful 3D Open-world Learning 21 Nov 2022 · 2 repositories · arXiv:2211.11682Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 4 pointer-only (licence)
-
Conceptor-Aided Debiasing of Large Language Models 20 Nov 2022 · 0 repositories · arXiv:2211.11087
-
A survey on knowledge-enhanced multimodal learning 19 Nov 2022 · 0 repositories · arXiv:2211.12328
-
Random-LTD: Random and Layerwise Token Dropping Brings Efficient Training for Large-scale Transformers 17 Nov 2022 · 1 repository · arXiv:2211.11586
-
TSMind: Alibaba and Soochow University's Submission to the WMT22 Translation Suggestion Task 16 Nov 2022 · 0 repositories · arXiv:2211.08987
-
Textual Data Augmentation for Patient Outcomes Prediction 13 Nov 2022 · 0 repositories · arXiv:2211.06778
-
Large Language Models Meet Harry Potter: A Bilingual Dataset for Aligning Dialogue Agents with Characters 13 Nov 2022 · 1 repository · arXiv:2211.06869
-
Collateral facilitation in humans and language models 9 Nov 2022 · 1 repository · arXiv:2211.05198Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Active Example Selection for In-Context Learning 8 Nov 2022 · 1 repository · arXiv:2211.04486Syntology official (archive's flag): 10 ran · 10 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 6 where Syntology's instrument failed) · 6 unverified (of 16 harvested samples)
-
Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small 1 Nov 2022 · 7 repositories · arXiv:2211.00593Syntology community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
Text-Only Training for Image Captioning using Noise-Injected CLIP 1 Nov 2022 · 4 repositories · arXiv:2211.00575Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers 31 Oct 2022 · 17 repositories · arXiv:2210.17323Syntology official (archive's flag): 1 ran · 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 10 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
SSD-LM: Semi-autoregressive Simplex-based Diffusion Language Model for Text Generation and Modular Control 31 Oct 2022 · 2 repositories · arXiv:2210.17432
-
Probing for targeted syntactic knowledge through grammatical error detection 28 Oct 2022 · 1 repository · arXiv:2210.16228
-
TRScore: A Novel GPT-based Readability Scorer for ASR Segmentation and Punctuation model evaluation and selection 27 Oct 2022 · 0 repositories · arXiv:2210.15104
-
Exploring Robustness of Prefix Tuning in Noisy Data: A Case Study in Financial Sentiment Analysis 26 Oct 2022 · 0 repositories · arXiv:2211.05584
-
IELM: An Open Information Extraction Benchmark for Pre-Trained Language Models 25 Oct 2022 · 0 repositories · arXiv:2210.14128
-
Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task 24 Oct 2022 · 4 repositories · arXiv:2210.13382Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Perfectly Secure Steganography Using Minimum Entropy Coupling 24 Oct 2022 · 2 repositories · arXiv:2210.14889Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)
-
Meta-learning Pathologies from Radiology Reports using Variance Aware Prototypical Networks 22 Oct 2022 · 0 repositories · arXiv:2210.13979
-
A Causal Framework to Quantify the Robustness of Mathematical Reasoning with Language Models 21 Oct 2022 · 1 repository · arXiv:2210.12023Syntology official (archive's flag): 5 ran · 5 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
General Image Descriptors for Open World Image Retrieval using ViT CLIP 20 Oct 2022 · 1 repository · arXiv:2210.11141Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
BioGPT: Generative Pre-trained Transformer for Biomedical Text Generation and Mining 19 Oct 2022 · 4 repositories · arXiv:2210.10341
-
Towards a neural architecture of language: Deep learning versus logistics of access in neural architectures for compositional processing 19 Oct 2022 · 0 repositories · arXiv:2210.10543
-
Team Flow at DRC2022: Pipeline System for Travel Destination Recommendation Task in Spoken Dialogue 18 Oct 2022 · 0 repositories · arXiv:2210.09518
-
A Generative User Simulator with GPT-based Architecture and Goal State Tracking for Reinforced Multi-Domain Dialog Systems 17 Oct 2022 · 1 repository · arXiv:2210.08692Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
DyLoRA: Parameter Efficient Tuning of Pre-trained Models using Dynamic Search-Free Low-Rank Adaptation 14 Oct 2022 · 2 repositories · arXiv:2210.07558Syntology official: not harvested · 0 ran · 1 unverified (of 1 harvested sample)
-
"John is 50 years old, can his son be 65?" Evaluating NLP Models' Understanding of Feasibility 14 Oct 2022 · 1 repository · arXiv:2210.07471
-
Jointly Reinforced User Simulator and Task-oriented Dialog System with Simplified Generative Architecture 13 Oct 2022 · 0 repositories · arXiv:2210.06706
-
Foundation Transformers 12 Oct 2022 · 4 repositories · arXiv:2210.06423Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)