Methods › General › Fine-Tuning › Discriminative Fine-Tuning › Papers, page 12
Discriminative Fine-Tuning
Papers archive 2025-07-28
archive papers tagged: 1,990 · with a code link: 794 · where Syntology ran a sample: 271 (223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (271 of 1,990 tagged: 223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument)
Page 12 of 20: papers 1,101 to 1,200 of 1,990, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Discriminative Speech Recognition Rescoring with Pre-trained Language Models 10 Oct 2023 · 0 repositories · arXiv:2310.06248
-
GPT-4 as an Agronomist Assistant? Answering Agriculture Exams Using Large Language Models 10 Oct 2023 · 0 repositories · arXiv:2310.06225
-
Humans and language models diverge when predicting repeating text 10 Oct 2023 · 1 repository · arXiv:2310.06408Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Foundation Models Meet Visualizations: Challenges and Opportunities 9 Oct 2023 · 0 repositories · arXiv:2310.05771
-
Distantly-Supervised Joint Extraction with Noise-Robust Learning 8 Oct 2023 · 1 repository · arXiv:2310.04994
-
Do self-supervised speech and language models extract similar representations as human brain? 7 Oct 2023 · 0 repositories · arXiv:2310.04645
-
LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT 7 Oct 2023 · 2 repositories · arXiv:2310.04673Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Question-focused Summarization by Decomposing Articles into Facts and Opinions and Retrieving Entities 7 Oct 2023 · 0 repositories · arXiv:2310.04880
-
Copy Suppression: Comprehensively Understanding an Attention Head 6 Oct 2023 · 1 repository · arXiv:2310.04625Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
Keyword Augmented Retrieval: Novel framework for Information Retrieval integrated with speech interface 6 Oct 2023 · 0 repositories · arXiv:2310.04205
-
SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks 5 Oct 2023 · 1 repository · arXiv:2310.03684
-
Discriminative Training of VBx Diarization 4 Oct 2023 · 1 repository · arXiv:2310.02732
-
Memoria: Resolving Fateful Forgetting Problem through Human-Inspired Memory Architecture 4 Oct 2023 · 1 repository · arXiv:2310.03052Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 9 unverified (of 10 harvested samples)
-
NOLA: Compressing LoRA using Linear Combination of Random Basis 4 Oct 2023 · 1 repository · arXiv:2310.02556Syntology official (archive's flag): 5 ran · 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Retrieval meets Long Context Large Language Models 4 Oct 2023 · 0 repositories · arXiv:2310.03025
-
PolySketchFormer: Fast Transformers via Sketching Polynomial Kernels 2 Oct 2023 · 0 repositories · arXiv:2310.01655
-
Analyzing and Mitigating Object Hallucination in Large Vision-Language Models 1 Oct 2023 · 1 repository · arXiv:2310.00754Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
RoleLLM: Benchmarking, Eliciting, and Enhancing Role-Playing Abilities of Large Language Models 1 Oct 2023 · 2 repositories · arXiv:2310.00746
-
An evaluation of GPT models for phenotype concept recognition 29 Sep 2023 · 0 repositories · arXiv:2309.17169
-
Revolutionizing Mobile Interaction: Enabling a 3 Billion Parameter GPT LLM on Mobile 29 Sep 2023 · 0 repositories · arXiv:2310.01434
-
Split and Merge: Aligning Position Biases in LLM-based Evaluators 29 Sep 2023 · 0 repositories · arXiv:2310.01432
-
Training and inference of large language models using 8-bit floating point 29 Sep 2023 · 0 repositories · arXiv:2309.17224
-
AE-GPT: Using Large Language Models to Extract Adverse Events from Surveillance Reports-A Use Case with Influenza Vaccine Adverse Events 28 Sep 2023 · 0 repositories · arXiv:2309.16150
-
Large Language Model Soft Ideologization via AI-Self-Consciousness 28 Sep 2023 · 0 repositories · arXiv:2309.16167
-
MindGPT: Interpreting What You See with Non-invasive Brain Recordings 27 Sep 2023 · 1 repository · arXiv:2309.15729Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Legal Question-Answering in the Indian Context: Efficacy, Challenges, and Potential of Modern AI Models 26 Sep 2023 · 0 repositories · arXiv:2309.14735
-
LogGPT: Log Anomaly Detection via GPT 25 Sep 2023 · 1 repository · arXiv:2309.14482
-
Seeing Is Not Always Believing: Invisible Collision Attack and Defence on Pre-Trained Models 24 Sep 2023 · 1 repository · arXiv:2309.13579
-
AMPLIFY:Attention-based Mixup for Performance Improvement and Label Smoothing in Transformer 22 Sep 2023 · 1 repository · arXiv:2309.12689
-
Investigating Large Language Models and Control Mechanisms to Improve Text Readability of Biomedical Abstracts 22 Sep 2023 · 1 repository · arXiv:2309.13202
-
Bad Actor, Good Advisor: Exploring the Role of Large Language Models in Fake News Detection 21 Sep 2023 · 1 repository · arXiv:2309.12247
-
Constraints First: A New MDD-based Model to Generate Sentences Under Constraints 21 Sep 2023 · 0 repositories · arXiv:2309.12415
-
Sequence-to-Sequence Spanish Pre-trained Language Models 20 Sep 2023 · 1 repository · arXiv:2309.11259
-
The Languini Kitchen: Enabling Language Modelling Research at Different Scales of Compute 20 Sep 2023 · 1 repository · arXiv:2309.11197
-
Rigorously Assessing Natural Language Explanations of Neurons 19 Sep 2023 · 0 repositories · arXiv:2309.10312
-
RECAP: Retrieval-Augmented Audio Captioning 18 Sep 2023 · 1 repository · arXiv:2309.09836Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Towards Ontology Construction with Language Models 18 Sep 2023 · 0 repositories · arXiv:2309.09898
-
A Modern Turkish Poet: Fine-Tuned GPT-2 15 Sep 2023 · 1 repository
-
Indian-BhED: A Dataset for Measuring India-Centric Biases in Large Language Models 15 Sep 2023 · 1 repository · arXiv:2309.08573
-
CoCA: Fusing Position Embedding with Collinear Constrained Attention in Transformers for Long Context Window Extending 15 Sep 2023 · 1 repository · arXiv:2309.08646
-
GPT-Lab: Next Generation Of Optimal Chemistry Discovery By GPT Driven Robotic Lab 15 Sep 2023 · 0 repositories · arXiv:2309.16721
-
Assessing the nature of large language models: A caution against anthropocentrism 14 Sep 2023 · 0 repositories · arXiv:2309.07683
-
ChatGPT MT: Competitive for High- (but not Low-) Resource Languages 14 Sep 2023 · 2 repositories · arXiv:2309.07423Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Traveling Words: A Geometric Interpretation of Transformers 13 Sep 2023 · 1 repository · arXiv:2309.07315
-
Circuit Breaking: Removing Model Behaviors with Targeted Ablation 12 Sep 2023 · 1 repository · arXiv:2309.05973Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Characterizing Latent Perspectives of Media Houses Towards Public Figures 12 Sep 2023 · 0 repositories · arXiv:2309.06112
-
Exploring Large Language Models for Ontology Alignment 12 Sep 2023 · 1 repository · arXiv:2309.07172
-
Unveiling the potential of large language models in generating semantic and cross-language clones 12 Sep 2023 · 0 repositories · arXiv:2309.06424
-
Memory Injections: Correcting Multi-Hop Reasoning Failures during Inference in Transformer-Based Language Models 11 Sep 2023 · 1 repository · arXiv:2309.05605Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Zero-shot Learning with Minimum Instruction to Extract Social Determinants and Family History from Clinical Notes using GPT Model 11 Sep 2023 · 0 repositories · arXiv:2309.05475
-
Supervised Learning and Large Language Model Benchmarks on Mental Health Datasets: Cognitive Distortions and Suicidal Risks in Chinese Social Media 7 Sep 2023 · 2 repositories · arXiv:2309.03564
-
Zero-Shot Audio Captioning via Audibility Guidance 7 Sep 2023 · 0 repositories · arXiv:2309.03884
-
CodeApex: A Bilingual Programming Evaluation Benchmark for Large Language Models 5 Sep 2023 · 1 repository · arXiv:2309.01940
-
Do You Trust ChatGPT? -- Perceived Credibility of Human and AI-Generated Content 5 Sep 2023 · 0 repositories · arXiv:2309.02524
-
Do androids dream of fictional references? A bibliographic dialogue with ChatGPT3.5 4 Sep 2023 · 0 repositories · arXiv:2312.00789
-
Large Language Models for Semantic Monitoring of Corporate Disclosures: A Case Study on Korea's Top 50 KOSPI Companies 1 Sep 2023 · 0 repositories · arXiv:2309.00208
-
Why do universal adversarial attacks work on large language models?: Geometry might be the answer 1 Sep 2023 · 0 repositories · arXiv:2309.00254
-
GPT has become financially literate: Insights from financial literacy tests of GPT and a preliminary test of how people use it as a source of advice 31 Aug 2023 · 0 repositories · arXiv:2309.00649
-
Large Language Models as Data Preprocessors 30 Aug 2023 · 0 repositories · arXiv:2308.16361
-
Quantifying Uncertainty in Answers from any Language Model and Enhancing their Trustworthiness 30 Aug 2023 · 0 repositories · arXiv:2308.16175
-
Breaking the Bank with ChatGPT: Few-Shot Text Classification for Finance 28 Aug 2023 · 0 repositories · arXiv:2308.14634
-
Target-independent XLA optimization using Reinforcement Learning 28 Aug 2023 · 0 repositories · arXiv:2308.14364
-
TextrolSpeech: A Text Style Control Speech Corpus With Codec Language Text-to-Speech Models 28 Aug 2023 · 1 repository · arXiv:2308.14430
-
Examining User-Friendly and Open-Sourced Large GPT Models: A Survey on Language, Multimodal, and Scientific GPT Models 27 Aug 2023 · 1 repository · arXiv:2308.14149
-
Improving Knowledge Distillation for BERT Models: Loss Functions, Mapping Methods, and Weight Tuning 26 Aug 2023 · 0 repositories · arXiv:2308.13958
-
MLLM-DataEngine: An Iterative Refinement Approach for MLLM 25 Aug 2023 · 1 repository · arXiv:2308.13566Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Transforming the Output of Generative Pre-trained Transformer: The Influence of the PGI Framework on Attention Dynamics 25 Aug 2023 · 0 repositories · arXiv:2308.13317
-
Financial News Analytics Using Fine-Tuned Llama 2 GPT Model 24 Aug 2023 · 0 repositories · arXiv:2308.13032
-
Evaluating Large Language Models on Graphs: Performance Insights and Comparative Analysis 22 Aug 2023 · 1 repository · arXiv:2308.11224Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Exploring the Effectiveness of GPT Models in Test-Taking: A Case Study of the Driver's License Knowledge Test 22 Aug 2023 · 0 repositories · arXiv:2308.11827
-
MulMarker: a comprehensive framework for identifying multi-gene prognostic signatures 22 Aug 2023 · 1 repository · arXiv:2308.11349
-
Tryage: Real-time, intelligent Routing of User Prompts to Large Language Models 22 Aug 2023 · 0 repositories · arXiv:2308.11601
-
GPT-in-the-Loop: Adaptive Decision-Making for Multiagent Systems 21 Aug 2023 · 0 repositories · arXiv:2308.10435
-
GradientCoin: A Peer-to-Peer Decentralized Large Language Models 21 Aug 2023 · 0 repositories · arXiv:2308.10502
-
Large Language Models on Wikipedia-Style Survey Generation: an Evaluation in NLP Concepts 21 Aug 2023 · 1 repository · arXiv:2308.10410
-
Steering Language Models With Activation Engineering 20 Aug 2023 · 2 repositories · arXiv:2308.10248Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
How Good Are LLMs at Out-of-Distribution Detection? 20 Aug 2023 · 1 repository · arXiv:2308.10261Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Data-to-text Generation for Severely Under-Resourced Languages with GPT-3.5: A Bit of Help Needed from Google Translate 19 Aug 2023 · 1 repository · arXiv:2308.09957
-
A tailored Handwritten-Text-Recognition System for Medieval Latin 18 Aug 2023 · 0 repositories · arXiv:2308.09368
-
A Preliminary Study on a Conceptual Game Feature Generation and Recommendation System 16 Aug 2023 · 0 repositories · arXiv:2308.13538
-
Approximating Human-Like Few-shot Learning with GPT-based Compression 14 Aug 2023 · 0 repositories · arXiv:2308.06942
-
Generating Individual Trajectories Using GPT-2 Trained from Scratch on Encoded Spatiotemporal Data 14 Aug 2023 · 0 repositories · arXiv:2308.07940
-
LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked 14 Aug 2023 · 1 repository · arXiv:2308.07308
-
Playing with Words: Comparing the Vocabulary and Lexical Richness of ChatGPT and Humans 14 Aug 2023 · 0 repositories · arXiv:2308.07462
-
Semantic Similarity Loss for Neural Source Code Summarization 14 Aug 2023 · 1 repository · arXiv:2308.07429
-
Enhancing Phenotype Recognition in Clinical Notes Using Large Language Models: PhenoBCBERT and PhenoGPT 11 Aug 2023 · 1 repository · arXiv:2308.06294
-
Large Language Models to Identify Social Determinants of Health in Electronic Health Records 11 Aug 2023 · 1 repository · arXiv:2308.06354Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
AudioLDM 2: Learning Holistic Audio Generation with Self-supervised Pretraining 10 Aug 2023 · 2 repositories · arXiv:2308.05734Syntology official (archive's flag): 8 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 2 honoured, 3 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 11 unverified (of 27 harvested samples) · 19 pointer-only (licence)
-
Testing GPT-4 with Wolfram Alpha and Code Interpreter plug-ins on math and science problems 10 Aug 2023 · 0 repositories · arXiv:2308.05713
-
WeaverBird: Empowering Financial Decision-Making with Large Language Model, Knowledge Base, and Search Engine 10 Aug 2023 · 1 repository · arXiv:2308.05361
-
An Empirical Study on Using Large Language Models to Analyze Software Supply Chain Security Failures 9 Aug 2023 · 0 repositories · arXiv:2308.04898
-
LLMeBench: A Flexible Framework for Accelerating LLMs Benchmarking 9 Aug 2023 · 1 repository · arXiv:2308.04945
-
Gromov-Wasserstein unsupervised alignment reveals structural correspondences between the color similarity structures of humans and large language models 8 Aug 2023 · 0 repositories · arXiv:2308.04381
-
I-WAS: a Data Augmentation Method with GPT-2 for Simile Detection 8 Aug 2023 · 0 repositories · arXiv:2308.04109
-
Fact-Checking Generative AI: Ontology-Driven Biological Graphs for Disease-Gene Link Verification 7 Aug 2023 · 0 repositories · arXiv:2308.03929
-
GPTScan: Detecting Logic Vulnerabilities in Smart Contracts by Combining GPT with Program Analysis 7 Aug 2023 · 1 repository · arXiv:2308.03314
-
Explaining Relation Classification Models with Semantic Extents 4 Aug 2023 · 2 repositories · arXiv:2308.02193
-
GEMRec: Towards Generative Model Recommendation 4 Aug 2023 · 1 repository · arXiv:2308.02205
-
Baby Llama: knowledge distillation from an ensemble of teachers trained on a small dataset with no performance penalty 3 Aug 2023 · 1 repository · arXiv:2308.02019
-
Does Correction Remain A Problem For Large Language Models? 3 Aug 2023 · 0 repositories · arXiv:2308.01776