Methods › General › Fine-Tuning › Discriminative Fine-Tuning › Papers, page 11
Discriminative Fine-Tuning
Papers archive 2025-07-28
archive papers tagged: 1,990 · with a code link: 794 · where Syntology ran a sample: 271 (223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (271 of 1,990 tagged: 223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument)
Page 11 of 20: papers 1,001 to 1,100 of 1,990, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
TimelyGPT: Extrapolatable Transformer Pre-training for Long-term Time-Series Forecasting in Healthcare 29 Nov 2023 · 0 repositories · arXiv:2312.00817
-
CharacterGLM: Customizing Chinese Conversational AI Characters with Large Language Models 28 Nov 2023 · 1 repository · arXiv:2311.16832Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
ChatGPT's One-year Anniversary: Are Open-Source Large Language Models Catching up? 28 Nov 2023 · 1 repository · arXiv:2311.16989
-
COLE: A Hierarchical Generation Framework for Multi-Layered and Editable Graphic Design 28 Nov 2023 · 0 repositories · arXiv:2311.16974
-
Comparing Generative Chatbots Based on Process Requirements 28 Nov 2023 · 0 repositories · arXiv:2312.03741
-
SEED-Bench-2: Benchmarking Multimodal Large Language Models 28 Nov 2023 · 2 repositories · arXiv:2311.17092
-
Real Customization or Just Marketing: Are Customized Versions of Chat GPT Useful? 27 Nov 2023 · 0 repositories · arXiv:2312.03728
-
Leveraging AI-derived Data for Carbon Accounting: Information Extraction from Alternative Sources 26 Nov 2023 · 0 repositories · arXiv:2312.03722
-
UHGEval: Benchmarking the Hallucination of Chinese Large Language Models via Unconstrained Generation 26 Nov 2023 · 1 repository · arXiv:2311.15296Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
CMed-GPT: Prompt Tuning for Entity-Aware Chinese Medical Dialogue Generation 24 Nov 2023 · 0 repositories · arXiv:2311.14539
-
Data-to-Text Bilingual Generation 24 Nov 2023 · 2 repositories · arXiv:2311.14808
-
GPT Struct Me: Probing GPT Models on Narrative Entity Extraction 24 Nov 2023 · 1 repository · arXiv:2311.14583
-
Towards Auditing Large Language Models: Improving Text-based Stereotype Detection 23 Nov 2023 · 0 repositories · arXiv:2311.14126
-
Comparison of pipeline, sequence-to-sequence, and GPT models for end-to-end relation extraction: experiments with the rare disease use-case 22 Nov 2023 · 1 repository · arXiv:2311.13729
-
Drilling Down into the Discourse Structure with LLMs for Long Document Question Answering 22 Nov 2023 · 0 repositories · arXiv:2311.13565
-
AcademicGPT: Empowering Academic Research 21 Nov 2023 · 0 repositories · arXiv:2311.12315
-
ALPHA: AnomaLous Physiological Health Assessment Using Large Language Models 21 Nov 2023 · 1 repository · arXiv:2311.12524
-
Descriptor and Word Soups: Overcoming the Parameter Efficiency Accuracy Tradeoff for Out-of-Distribution Few-shot Learning 21 Nov 2023 · 1 repository · arXiv:2311.13612
-
Extracting Definienda in Mathematical Scholarly Articles with Transformers 21 Nov 2023 · 2 repositories · arXiv:2311.12448
-
GPT4Motion: Scripting Physical Motions in Text-to-Video Generation via Blender-Oriented GPT Planning 21 Nov 2023 · 0 repositories · arXiv:2311.12631
-
Assessing Prompt Injection Risks in 200+ Custom GPTs 20 Nov 2023 · 1 repository · arXiv:2311.11538
-
Towards Human-Level Text Coding with LLMs: The Case of Fatherhood Roles in Public Policy Documents 20 Nov 2023 · 1 repository · arXiv:2311.11844
-
MemoryCompanion: A Smart Healthcare Solution to Empower Efficient Alzheimer's Care Via Unleashing Generative AI 20 Nov 2023 · 0 repositories · arXiv:2311.14730
-
Bias A-head? Analyzing Bias in Transformer-Based Language Model Attention Heads 17 Nov 2023 · 0 repositories · arXiv:2311.10395
-
DynaPipe: Optimizing Multi-task Training through Dynamic Pipelines 17 Nov 2023 · 2 repositories · arXiv:2311.10418
-
Event Causality Is Key to Computational Story Understanding 16 Nov 2023 · 1 repository · arXiv:2311.09648Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
On Retrieval Augmentation and the Limitations of Language Model Training 16 Nov 2023 · 0 repositories · arXiv:2311.09615
-
Predictive Minds: LLMs As Atypical Active Inference Agents 16 Nov 2023 · 0 repositories · arXiv:2311.10215
-
LOKE: Linked Open Knowledge Extraction for Automated Knowledge Graph Construction 15 Nov 2023 · 0 repositories · arXiv:2311.09366
-
Unifying the Perspectives of NLP and Software Engineering: A Survey on Language Models for Code 14 Nov 2023 · 1 repository · arXiv:2311.07989
-
Evaluating LLMs on Document-Based QA: Exact Answer Selection and Numerical Extraction using Cogtale dataset 14 Nov 2023 · 0 repositories · arXiv:2311.07878
-
Fair Abstractive Summarization of Diverse Perspectives 14 Nov 2023 · 1 repository · arXiv:2311.07884
-
Large Language Model-Driven Classroom Flipping: Empowering Student-Centric Peer Questioning with Flipped Interaction 14 Nov 2023 · 0 repositories · arXiv:2311.14708
-
In-context Learning Generalizes, But Not Always Robustly: The Case of Syntax 13 Nov 2023 · 1 repository · arXiv:2311.07811
-
Language Model-In-The-Loop: Data Optimal Approach to Learn-To-Recommend Actions in Text Games 13 Nov 2023 · 0 repositories · arXiv:2311.07687
-
On The Truthfulness of 'Surprisingly Likely' Responses of Large Language Models 13 Nov 2023 · 0 repositories · arXiv:2311.07692
-
Establishing Performance Baselines in Fine-Tuning, Retrieval-Augmented Generation and Soft-Prompting for Non-Specialist LLM Users 10 Nov 2023 · 0 repositories · arXiv:2311.05903
-
Exploring Fine-tuning ChatGPT for News Recommendation 10 Nov 2023 · 0 repositories · arXiv:2311.05850
-
Smart Agent-Based Modeling: On the Use of Large Language Models in Computer Simulations 10 Nov 2023 · 4 repositories · arXiv:2311.06330Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
GeoFormer: Predicting Human Mobility using Generative Pre-trained Transformer (GPT) 9 Nov 2023 · 0 repositories · arXiv:2311.05092
-
Leveraging Artificial Intelligence Technology for Mapping Research to Sustainable Development Goals: A Case Study 9 Nov 2023 · 0 repositories · arXiv:2311.16162
-
Agent Lumos: Unified and Modular Training for Open-Source Language Agents 9 Nov 2023 · 2 repositories · arXiv:2311.05657Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Vision Encoder-Decoder Models for AI Coaching 9 Nov 2023 · 2 repositories · arXiv:2311.16161
-
Massive Editing for Large Language Models via Meta Learning 8 Nov 2023 · 1 repository · arXiv:2311.04661Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Towards Interpretable Sequence Continuation: Analyzing Shared Circuits in Large Language Models 7 Nov 2023 · 1 repository · arXiv:2311.04131Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Neuro-GPT: Towards A Foundation Model for EEG 7 Nov 2023 · 1 repository · arXiv:2311.03764Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Unraveling Downstream Gender Bias from Large Language Models: A Study on AI Educational Writing Assistance 6 Nov 2023 · 1 repository · arXiv:2311.03311
-
Measuring Five Accountable Talk Moves to Improve Instruction at Scale 2 Nov 2023 · 0 repositories · arXiv:2311.10749
-
Is GPT Powerful Enough to Analyze the Emotions of Memes? 1 Nov 2023 · 0 repositories · arXiv:2311.00223
-
Theory of Mind in Large Language Models: Examining Performance of 11 State-of-the-Art models vs. Children Aged 7-10 on Advanced Tests 31 Oct 2023 · 0 repositories · arXiv:2310.20320
-
Herd: Using multiple, smaller LLMs to match the performances of proprietary, large LLMs via an intelligent composer 30 Oct 2023 · 0 repositories · arXiv:2310.19902
-
LitCab: Lightweight Language Model Calibration over Short- and Long-form Responses 30 Oct 2023 · 1 repository · arXiv:2310.19208
-
Synthetic Imitation Edit Feedback for Factual Alignment in Clinical Summarization 30 Oct 2023 · 1 repository · arXiv:2310.20033Syntology official (archive's flag): 7 ran · 7 ran (of which 3 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
From Chatbots to PhishBots? -- Preventing Phishing scams created using ChatGPT, Google Bard and Claude 29 Oct 2023 · 0 repositories · arXiv:2310.19181
-
Efficient kernel surrogates for neural network-based regression 28 Oct 2023 · 0 repositories · arXiv:2310.18612
-
The Synergy of Speculative Decoding and Batching in Serving Large Language Models 28 Oct 2023 · 0 repositories · arXiv:2310.18813
-
Lost in Translation, Found in Spans: Identifying Claims in Multilingual Social Media 27 Oct 2023 · 1 repository · arXiv:2310.18205
-
Can LLMs Grade Short-Answer Reading Comprehension Questions : An Empirical Study with a Novel Dataset 26 Oct 2023 · 0 repositories · arXiv:2310.18373
-
From Transcripts to Insights: Uncovering Corporate Risks Using Generative AI 26 Oct 2023 · 0 repositories · arXiv:2310.17721
-
LightLM: A Lightweight Deep and Narrow Language Model for Generative Recommendation 26 Oct 2023 · 1 repository · arXiv:2310.17488Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 1 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
DecoderTracker: Decoder-Only Method for Multiple-Object Tracking 26 Oct 2023 · 0 repositories · arXiv:2310.17170
-
Sliceformer: Make Multi-head Attention as Simple as Sorting in Discriminative Tasks 26 Oct 2023 · 1 repository · arXiv:2310.17683
-
ZeroQuant-HERO: Hardware-Enhanced Robust Optimized Post-Training Quantization Framework for W8A8 Transformers 26 Oct 2023 · 0 repositories · arXiv:2310.17723
-
BabyStories: Can Reinforcement Learning Teach Baby Language Models to Write Better Stories? 25 Oct 2023 · 1 repository · arXiv:2310.16681Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Can GPT models Follow Human Summarization Guidelines? Evaluating ChatGPT and GPT-4 for Dialogue Summarization 25 Oct 2023 · 0 repositories · arXiv:2310.16810
-
Discrete Diffusion Modeling by Estimating the Ratios of the Data Distribution 25 Oct 2023 · 4 repositories · arXiv:2310.16834Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 1 honoured, 0 violated, 13 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 18 harvested samples) · 15 pointer-only (licence)
-
RCAgent: Cloud Root Cause Analysis by Autonomous Agents with Tool-Augmented Large Language Models 25 Oct 2023 · 0 repositories · arXiv:2310.16340
-
A Language Model with Limited Memory Capacity Captures Interference in Human Sentence Processing 24 Oct 2023 · 0 repositories · arXiv:2310.16142
-
Learning From Free-Text Human Feedback -- Collect New Datasets Or Extend Existing Ones? 24 Oct 2023 · 1 repository · arXiv:2310.15758Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Generative Pre-trained Transformer for Vietnamese Community-based COVID-19 Question Answering 23 Oct 2023 · 0 repositories · arXiv:2310.14602
-
Prefix-Tuning Based Unsupervised Text Style Transfer 23 Oct 2023 · 0 repositories · arXiv:2310.14599
-
Establishing Vocabulary Tests as a Benchmark for Evaluating Large Language Models 23 Oct 2023 · 1 repository · arXiv:2310.14703
-
Towards a Mechanistic Interpretation of Multi-Step Reasoning Capabilities of Language Models 23 Oct 2023 · 2 repositories · arXiv:2310.14491Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Is ChatGPT a game changer for geocoding -- a benchmark for geocoding address parsing techniques 22 Oct 2023 · 1 repository · arXiv:2310.14360
-
GEMBA-MQM: Detecting Translation Quality Error Spans with GPT-4 21 Oct 2023 · 1 repository · arXiv:2310.13988Syntology 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
AllTogether: Investigating the Efficacy of Spliced Prompt for Web Navigation using Large Language Models 20 Oct 2023 · 0 repositories · arXiv:2310.18331
-
Challenges and Contributing Factors in the Utilization of Large Language Models (LLMs) 20 Oct 2023 · 0 repositories · arXiv:2310.13343
-
Equivariant Transformer is all you need 20 Oct 2023 · 0 repositories · arXiv:2310.13222
-
Foundation Model's Embedded Representations May Detect Distribution Shift 20 Oct 2023 · 0 repositories · arXiv:2310.13836
-
Robust Training for Conversational Question Answering Models with Reinforced Reformulation Generation 20 Oct 2023 · 0 repositories · arXiv:2310.13505
-
Identifying and Adapting Transformer-Components Responsible for Gender Bias in an English Language Model 19 Oct 2023 · 1 repository · arXiv:2310.12611Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
LASER: Linear Compression in Wireless Distributed Optimization 19 Oct 2023 · 0 repositories · arXiv:2310.13033
-
Not All Countries Celebrate Thanksgiving: On the Cultural Dominance in Large Language Models 19 Oct 2023 · 0 repositories · arXiv:2310.12481
-
The Shifted and The Overlooked: A Task-oriented Investigation of User-GPT Interactions 19 Oct 2023 · 1 repository · arXiv:2310.12418Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Solving the multiplication problem of a large language model system using a graph-based method 18 Oct 2023 · 0 repositories · arXiv:2310.13016
-
Emergent AI-Assisted Discourse: Case Study of a Second Language Writer Authoring with ChatGPT 17 Oct 2023 · 0 repositories · arXiv:2310.10903
-
Data Contamination Through the Lens of Time 16 Oct 2023 · 1 repository · arXiv:2310.10628Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
MoConVQ: Unified Physics-Based Motion Control via Scalable Discrete Representations 16 Oct 2023 · 0 repositories · arXiv:2310.10198
-
Prediction of Arabic Legal Rulings using Large Language Models 16 Oct 2023 · 0 repositories · arXiv:2310.10260
-
Configuration Validation with Large Language Models 15 Oct 2023 · 0 repositories · arXiv:2310.09690
-
Image Augmentation with Controlled Diffusion for Weakly-Supervised Semantic Segmentation 15 Oct 2023 · 0 repositories · arXiv:2310.09760
-
Efficient Model-Agnostic Multi-Group Equivariant Networks 14 Oct 2023 · 0 repositories · arXiv:2310.09675
-
From Words and Exercises to Wellness: Farsi Chatbot for Self-Attachment Technique 13 Oct 2023 · 0 repositories · arXiv:2310.09362
-
QUIK: Towards End-to-End 4-Bit Inference on Generative Large Language Models 13 Oct 2023 · 1 repository · arXiv:2310.09259Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
Multiclass Classification of Policy Documents with Large Language Models 12 Oct 2023 · 0 repositories · arXiv:2310.08167
-
Promptor: A Conversational and Autonomous Prompt Generation Agent for Intelligent Text Entry Techniques 12 Oct 2023 · 0 repositories · arXiv:2310.08101
-
QASiNa: Religious Domain Question Answering using Sirah Nabawiyah 12 Oct 2023 · 1 repository · arXiv:2310.08102
-
Training Generative Question-Answering on Synthetic Data Obtained from an Instruct-tuned Model 12 Oct 2023 · 0 repositories · arXiv:2310.08072
-
InstructRetro: Instruction Tuning post Retrieval-Augmented Pretraining 11 Oct 2023 · 1 repository · arXiv:2310.07713
-
Uncovering Hidden Connections: Iterative Search and Reasoning for Video-grounded Dialog 11 Oct 2023 · 2 repositories · arXiv:2310.07259Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)