Methods › General › Fine-Tuning › Discriminative Fine-Tuning › Papers, page 16
Discriminative Fine-Tuning
Papers archive 2025-07-28
archive papers tagged: 1,990 · with a code link: 794 · where Syntology ran a sample: 271 (223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (271 of 1,990 tagged: 223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument)
Page 16 of 20: papers 1,501 to 1,600 of 1,990, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Predictive Querying for Autoregressive Neural Sequence Models 12 Oct 2022 · 1 repository · arXiv:2210.06464Syntology official (archive's flag): 7 ran · 7 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 5 where Syntology's instrument failed) · 6 unverified (of 13 harvested samples)
-
SUMBot: Summarizing Context in Open-Domain Dialogue Systems 12 Oct 2022 · 0 repositories · arXiv:2210.06496
-
Fine-Tuning Pre-trained Transformers into Decaying Fast Weights 9 Oct 2022 · 1 repository · arXiv:2210.04243Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
AlphaTuning: Quantization-Aware Parameter-Efficient Adaptation of Large-Scale Pre-Trained Language Models 8 Oct 2022 · 0 repositories · arXiv:2210.03858
-
Underspecification in Language Modeling Tasks: A Causality-Informed Study of Gendered Pronoun Resolution 30 Sep 2022 · 2 repositories · arXiv:2210.00131
-
SmallCap: Lightweight Image Captioning Prompted with Retrieval Augmentation 30 Sep 2022 · 1 repository · arXiv:2209.15323
-
DFX: A Low-latency Multi-FPGA Appliance for Accelerating Transformer-based Text Generation 22 Sep 2022 · 0 repositories · arXiv:2209.10797
-
Text Revealer: Private Text Reconstruction via Model Inversion Attacks against Transformers 21 Sep 2022 · 0 repositories · arXiv:2209.10505
-
Chain of Explanation: New Prompting Method to Generate Higher Quality Natural Language Explanation for Implicit Hate Speech 11 Sep 2022 · 0 repositories · arXiv:2209.04889
-
Why So Toxic? Measuring and Triggering Toxic Behavior in Open-Domain Chatbots 7 Sep 2022 · 0 repositories · arXiv:2209.03463
-
Every picture tells a story: Image-grounded controllable stylistic story generation 4 Sep 2022 · 0 repositories · arXiv:2209.01638
-
Using Large Language Models to Simulate Multiple Humans and Replicate Human Subject Studies 18 Aug 2022 · 2 repositories · arXiv:2208.10264Syntology community repositories only · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 4 harvested samples)
-
Neural Embeddings for Text 17 Aug 2022 · 1 repository · arXiv:2208.08386
-
MoCapAct: A Multi-Task Dataset for Simulated Humanoid Control 15 Aug 2022 · 1 repository · arXiv:2208.07363Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Adan: Adaptive Nesterov Momentum Algorithm for Faster Optimizing Deep Models 13 Aug 2022 · 9 repositories · arXiv:2208.06677Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Interacting with next-phrase suggestions: How suggestion systems aid and influence the cognitive processes of writing 1 Aug 2022 · 0 repositories · arXiv:2208.00636
-
Zero-Shot Video Captioning with Evolving Pseudo-Tokens 22 Jul 2022 · 1 repository · arXiv:2207.11100
-
Word Play for Playing Othello (Reverses) 18 Jul 2022 · 0 repositories · arXiv:2207.08766
-
Combing for Credentials: Active Pattern Extraction from Smart Reply 14 Jul 2022 · 0 repositories · arXiv:2207.10802
-
Few-shot training LLMs for project-specific code-summarization 9 Jul 2022 · 0 repositories · arXiv:2207.04237
-
Hidden Schema Networks 8 Jul 2022 · 0 repositories · arXiv:2207.03777
-
Neural Language Models are not Born Equal to Fit Brain Data, but Training Helps 7 Jul 2022 · 0 repositories · arXiv:2207.03380
-
Sensitivity Analysis on Transferred Neural Architectures of BERT and GPT-2 for Financial Sentiment Analysis 7 Jul 2022 · 0 repositories · arXiv:2207.03037
-
Materials Transformers Language Models for Generative Materials Design: a benchmark study 27 Jun 2022 · 1 repository · arXiv:2206.13578
-
Explainable and High-Performance Hate and Offensive Speech Detection 26 Jun 2022 · 0 repositories · arXiv:2206.12983
-
CoCoPIE XGen: A Full-Stack AI-Oriented Optimizing Framework 21 Jun 2022 · 0 repositories · arXiv:2206.10620
-
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models 9 Jun 2022 · 6 repositories · arXiv:2206.04615Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Differentially Private Model Compression 3 Jun 2022 · 0 repositories · arXiv:2206.01838
-
Romantic-Computing 1 Jun 2022 · 0 repositories · arXiv:2206.11864
-
Multi-Agent Reinforcement Learning is a Sequence Modeling Problem 30 May 2022 · 1 repository · arXiv:2205.14953
-
CPED: A Large-Scale Chinese Personalized and Emotional Dialogue Dataset for Conversational AI 29 May 2022 · 1 repository · arXiv:2205.14727
-
FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness 27 May 2022 · 13 repositories · arXiv:2205.14135Syntology official: no sample here; runs from other or unrecorded repositories · 24 ran (of which 2 constructed an object rather than computing a result; 18 with no instrument failure: 0 honoured, 1 violated, 17 with no contract checked; 6 where Syntology's instrument failed) · 6 unverified (of 30 harvested samples) · 1 pointer-only (licence)
-
kNN-Prompt: Nearest Neighbor Zero-Shot Inference 27 May 2022 · 1 repository · arXiv:2205.13792
-
Do we need Label Regularization to Fine-tune Pre-trained Language Models? 25 May 2022 · 0 repositories · arXiv:2205.12428
-
Transcormer: Transformer for Sentence Scoring with Sliding Language Modeling 25 May 2022 · 1 repository · arXiv:2205.12986Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Garden-Path Traversal in GPT-2 24 May 2022 · 1 repository · arXiv:2205.12302Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
On the Role of Bidirectionality in Language Model Pre-Training 24 May 2022 · 0 repositories · arXiv:2205.11726
-
GraphMAE: Self-Supervised Masked Graph Autoencoders 22 May 2022 · 3 repositories · arXiv:2205.10803Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Life after BERT: What do Other Muppets Understand about Language? 21 May 2022 · 1 repository · arXiv:2205.10696
-
Prototypical Calibration for Few-shot Learning of Language Models 20 May 2022 · 1 repository · arXiv:2205.10183
-
Automated Scoring for Reading Comprehension via In-context BERT Tuning 19 May 2022 · 1 repository · arXiv:2205.09864
-
Towards Understanding Gender-Seniority Compound Bias in Natural Language Generation 19 May 2022 · 1 repository · arXiv:2205.09830
-
Evaluation of Transfer Learning for Polish with a Text-to-Text Model 18 May 2022 · 0 repositories · arXiv:2205.08808
-
What GPT Knows About Who is Who 16 May 2022 · 1 repository · arXiv:2205.07407
-
Naturalistic Causal Probing for Morpho-Syntax 14 May 2022 · 1 repository · arXiv:2205.07043
-
Ratatouille: A tool for Novel Recipe Generation 10 May 2022 · 0 repositories · arXiv:2206.08267
-
Multi-segment preserving sampling for deep manifold sampler 9 May 2022 · 0 repositories · arXiv:2205.04259
-
When a sentence does not introduce a discourse entity, Transformer-based models still sometimes refer to it 6 May 2022 · 1 repository · arXiv:2205.03472
-
Provably Confidential Language Modelling 4 May 2022 · 1 repository · arXiv:2205.01863Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
HiNER: A Large Hindi Named Entity Recognition Dataset 28 Apr 2022 · 1 repository · arXiv:2204.13743
-
Tailor: A Prompt-Based Approach to Attribute-Based Controlled Text Generation 28 Apr 2022 · 0 repositories · arXiv:2204.13362
-
Measuring artificial intelligence: a systematic assessment and implications for governance 21 Apr 2022 · 0 repositories · arXiv:2204.10304
-
Impact of Tokenization on Language Models: An Analysis for Turkish 19 Apr 2022 · 0 repositories · arXiv:2204.08832
-
Zero-shot Entity and Tweet Characterization with Designed Conditional Prompts and Contexts 18 Apr 2022 · 0 repositories · arXiv:2204.08405
-
mGPT: Few-Shot Learners Go Multilingual 15 Apr 2022 · 1 repository · arXiv:2204.07580
-
Polling Latent Opinions: A Method for Computational Sociolinguistics Using Transformer Language Models 15 Apr 2022 · 1 repository · arXiv:2204.07483
-
Analysing similarities between legal court documents using natural language processing approaches based on Transformers 14 Apr 2022 · 0 repositories · arXiv:2204.07182
-
Uniform Complexity for Text Generation 11 Apr 2022 · 1 repository · arXiv:2204.05185
-
FoundationLayerNorm: Scaling BERT and GPT to 1,000 Layers 9 Apr 2022 · 0 repositories · arXiv:2204.04477
-
Accelerating Attention through Gradient-Based Learned Runtime Pruning 7 Apr 2022 · 0 repositories · arXiv:2204.03227
-
Testing the limits of natural language models for predicting human language judgments 7 Apr 2022 · 1 repository · arXiv:2204.03592
-
Knowledge Infused Decoding 6 Apr 2022 · 1 repository · arXiv:2204.03084Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Data Augmentation for Intent Classification with Off-the-shelf Large Language Models 5 Apr 2022 · 1 repository · arXiv:2204.01959Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
Effect and Analysis of Large-scale Language Model Rescoring on Competitive ASR Systems 1 Apr 2022 · 0 repositories · arXiv:2204.00212
-
Monarch: Expressive Structured Matrices for Efficient and Accurate Training 1 Apr 2022 · 2 repositories · arXiv:2204.00595Syntology community repositories only · 15 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 1 violated, 13 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 15 harvested samples)
-
CREATE: A Benchmark for Chinese Short Video Retrieval and Title Generation 31 Mar 2022 · 0 repositories · arXiv:2203.16763
-
Bailando: 3D Dance Generation by Actor-Critic GPT with Choreographic Memory 24 Mar 2022 · 1 repository · arXiv:2203.13055Syntology official (archive's flag): 4 ran · 4 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Self-supervision through Random Segments with Autoregressive Coding (RandSAC) 22 Mar 2022 · 0 repositories · arXiv:2203.12054
-
A Slot Is Not Built in One Utterance: Spoken Language Dialogs with Sub-Slots 21 Mar 2022 · 1 repository · arXiv:2203.10759
-
Compression of Generative Pre-trained Language Models via Quantization 21 Mar 2022 · 0 repositories · arXiv:2203.10705
-
Dependency-based Mixture Language Models 19 Mar 2022 · 1 repository · arXiv:2203.10256Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Are You Robert or RoBERTa? Deceiving Online Authorship Attribution Models Using Neural Text Generators 18 Mar 2022 · 0 repositories · arXiv:2203.09813
-
Do Language Models Plagiarize? 15 Mar 2022 · 1 repository · arXiv:2203.07618
-
Contrastive Visual Semantic Pretraining Magnifies the Semantics of Natural Language Representations 14 Mar 2022 · 0 repositories · arXiv:2203.07511
-
GrIPS: Gradient-free, Edit-based Instruction Search for Prompting Large Language Models 14 Mar 2022 · 2 repositories · arXiv:2203.07281Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
VAST: The Valence-Assessing Semantics Test for Contextualizing Language Models 14 Mar 2022 · 1 repository · arXiv:2203.07504
-
ELLE: Efficient Lifelong Pre-training for Emerging Data 12 Mar 2022 · 1 repository · arXiv:2203.06311Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Block-Sparse Adversarial Attack to Fool Transformer-Based Text Classifiers 11 Mar 2022 · 1 repository · arXiv:2203.05948
-
When classifying grammatical role, BERT doesn't care about word order... except when it matters 11 Mar 2022 · 1 repository · arXiv:2203.06204
-
NLX-GPT: A Model for Natural Language Explanations in Vision and Vision-Language Tasks 9 Mar 2022 · 1 repository · arXiv:2203.05081
-
LiteTransformerSearch: Training-free Neural Architecture Search for Efficient Language Models 4 Mar 2022 · 1 repository · arXiv:2203.02094Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Parameter-Efficient Mixture-of-Experts Architecture for Pre-trained Language Models 2 Mar 2022 · 2 repositories · arXiv:2203.01104Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Exploring and Adapting Chinese GPT to Pinyin Input Method 1 Mar 2022 · 1 repository · arXiv:2203.00249
-
A Systematic Evaluation of Large Language Models of Code 26 Feb 2022 · 3 repositories · arXiv:2202.13169Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Consistent Dropout for Policy Gradient Reinforcement Learning 23 Feb 2022 · 0 repositories · arXiv:2202.11818
-
SGPT: GPT Sentence Embeddings for Semantic Search 17 Feb 2022 · 1 repository · arXiv:2202.08904Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Defending against Reconstruction Attacks with Rényi Differential Privacy 15 Feb 2022 · 0 repositories · arXiv:2202.07623
-
Maximizing Communication Efficiency for Large-scale Training via 0/1 Adam 12 Feb 2022 · 1 repository · arXiv:2202.06009
-
What are the best systems? New perspectives on NLP Benchmarking 8 Feb 2022 · 1 repository · arXiv:2202.03799
-
A Benchmark Corpus for the Detection of Automatically Generated Text in Academic Publications 4 Feb 2022 · 1 repository · arXiv:2202.02013
-
L3Cube-MahaCorpus and MahaBERT: Marathi Monolingual Corpus, Marathi BERT Language Models, and Resources 2 Feb 2022 · 1 repository · arXiv:2202.01159
-
A Frustratingly Simple Approach for End-to-End Image Captioning 30 Jan 2022 · 0 repositories · arXiv:2201.12723
-
DNNFuser: Generative Pre-Trained Transformer as a Generalized Mapper for Layer Fusion in DNN Accelerators 26 Jan 2022 · 0 repositories · arXiv:2201.11218
-
Pre-Trained Language Transformers are Universal Image Classifiers 25 Jan 2022 · 0 repositories · arXiv:2201.10182
-
Synthetic Books 24 Jan 2022 · 0 repositories · arXiv:2201.09518
-
Auto-regressive Text Generation with Pre-Trained Language Models: An Empirical Study on Question-type Short Text Generation 16 Jan 2022 · 0 repositories
-
Efficient Hierarchical Domain Adaptation for Pretrained Language Models 16 Jan 2022 · 0 repositories
-
Elastic Weight Consolidation for Reduction of Catastrophic Forgetting in GPT-2 16 Jan 2022 · 0 repositories
-
Jointly Reinforced User Simulator and Task-oriented Dialog System with Simplified Generative Architecture 16 Jan 2022 · 0 repositories
-
Polling Latent Opinions: A Method for Computational Sociolinguistics Using Transformer Language Models 16 Jan 2022 · 0 repositories