Methods › General › Fine-Tuning › Discriminative Fine-Tuning › Papers, page 6
Discriminative Fine-Tuning
Papers archive 2025-07-28
archive papers tagged: 1,990 · with a code link: 794 · where Syntology ran a sample: 271 (223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (271 of 1,990 tagged: 223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument)
Page 6 of 20: papers 501 to 600 of 1,990, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Enhancing Multi-hop Reasoning through Knowledge Erasure in Large Language Model Editing 22 Aug 2024 · 0 repositories · arXiv:2408.12456
-
Optimizing Performance: How Compact Models Match or Exceed GPT's Classification Capabilities through Fine-Tuning 22 Aug 2024 · 0 repositories · arXiv:2409.11408
-
Applying and Evaluating Large Language Models in Mental Health Care: A Scoping Review of Human-Assessed Generative Tasks 21 Aug 2024 · 0 repositories · arXiv:2408.11288
-
D-RMGPT: Robot-assisted collaborative tasks driven by large multimodal models 21 Aug 2024 · 0 repositories · arXiv:2408.11761
-
Mixed Sparsity Training: Achieving 4× FLOP Reduction for Transformer Pretraining 21 Aug 2024 · 0 repositories · arXiv:2408.11746
-
Language Modeling on Tabular Data: A Survey of Foundations, Techniques and Evolution 20 Aug 2024 · 1 repository · arXiv:2408.10548
-
Tracing Privacy Leakage of Language Models to Training Data via Adjusted Influence Functions 20 Aug 2024 · 0 repositories · arXiv:2408.10468
-
Security Attacks on LLM-based Code Completion Tools 20 Aug 2024 · 1 repository · arXiv:2408.11006
-
ELDER: Enhancing Lifelong Model Editing with Mixture-of-LoRA 19 Aug 2024 · 1 repository · arXiv:2408.11869
-
GARLIC: GPT-Augmented Reinforcement Learning with Intelligent Control for Vehicle Dispatching 19 Aug 2024 · 0 repositories · arXiv:2408.10286
-
Rhyme-aware Chinese lyric generator based on GPT 19 Aug 2024 · 0 repositories · arXiv:2408.10130
-
SSDTrain: An Activation Offloading Framework to SSDs for Faster Large Language Model Training 19 Aug 2024 · 0 repositories · arXiv:2408.10013
-
ConVerSum: A Contrastive Learning-based Approach for Data-Scarce Solution of Cross-Lingual Summarization Beyond Direct Equivalents 17 Aug 2024 · 0 repositories · arXiv:2408.09273
-
The Fellowship of the LLMs: Multi-Agent Workflows for Synthetic Preference Optimization Dataset Generation 16 Aug 2024 · 1 repository · arXiv:2408.08688
-
Transformers and Large Language Models for Efficient Intrusion Detection Systems: A Comprehensive Survey 14 Aug 2024 · 0 repositories · arXiv:2408.07583
-
Generative AI for automatic topic labelling 13 Aug 2024 · 0 repositories · arXiv:2408.07003
-
Pragmatic inference of scalar implicature by LLMs 13 Aug 2024 · 0 repositories · arXiv:2408.06673
-
Improving Whisper's Recognition Performance for Under-Represented Language Kazakh Leveraging Unpaired Speech and Text 10 Aug 2024 · 0 repositories · arXiv:2408.05554
-
From Text to Insight: Leveraging Large Language Models for Performance Evaluation in Management 9 Aug 2024 · 0 repositories · arXiv:2408.05328
-
Retrieval-augmented code completion for local projects using large language models 9 Aug 2024 · 0 repositories · arXiv:2408.05026
-
Transformer Explainer: Interactive Learning of Text-Generative Models 8 Aug 2024 · 1 repository · arXiv:2408.04619
-
Image-to-LaTeX Converter for Mathematical Formulas and Text 7 Aug 2024 · 1 repository · arXiv:2408.04015
-
Is Child-Directed Speech Effective Training Data for Language Models? 7 Aug 2024 · 1 repository · arXiv:2408.03617Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
SocFedGPT: Federated GPT-based Adaptive Content Filtering System Leveraging User Interactions in Social Networks 7 Aug 2024 · 0 repositories · arXiv:2408.05243
-
Data Poisoning in LLMs: Jailbreak-Tuning and Scaling Laws 6 Aug 2024 · 2 repositories · arXiv:2408.02946Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Evaluating the Translation Performance of Large Language Models Based on Euas-20 6 Aug 2024 · 0 repositories · arXiv:2408.03119
-
FLASH: Federated Learning-Based LLMs for Advanced Query Processing in Social Networks through RAG 6 Aug 2024 · 0 repositories · arXiv:2408.05242
-
TrafficGPT: An LLM Approach for Open-Set Encrypted Traffic Classification 6 Aug 2024 · 1 repository
-
ML-EAT: A Multilevel Embedding Association Test for Interpretable and Transparent Social Science 4 Aug 2024 · 1 repository · arXiv:2408.01966
-
AdaCBM: An Adaptive Concept Bottleneck Model for Explainable and Accurate Diagnosis 4 Aug 2024 · 1 repository · arXiv:2408.02001
-
Advancing Mental Health Pre-Screening: A New Custom GPT for Psychological Distress Assessment 3 Aug 2024 · 0 repositories · arXiv:2408.01614
-
Efficient Solutions For An Intriguing Failure of LLMs: Long Context Window Does Not Mean LLMs Can Analyze Long Sequences Flawlessly 3 Aug 2024 · 0 repositories · arXiv:2408.01866
-
What comes after transformers? -- A selective survey connecting ideas in deep learning 1 Aug 2024 · 0 repositories · arXiv:2408.00386
-
Performance of Recent Large Language Models for a Low-Resourced Language 31 Jul 2024 · 0 repositories · arXiv:2407.21330
-
Generative Expressive Conversational Speech Synthesis 31 Jul 2024 · 1 repository · arXiv:2407.21491
-
Automated Software Vulnerability Static Code Analysis Using Generative Pre-Trained Transformer Models 31 Jul 2024 · 0 repositories · arXiv:2408.00197
-
BERT and LLMs-Based avGFP Brightness Prediction and Mutation Design 30 Jul 2024 · 0 repositories · arXiv:2407.20534
-
AgEval: A Benchmark for Zero-Shot and Few-Shot Plant Stress Phenotyping with Multimodal LLMs 29 Jul 2024 · 0 repositories · arXiv:2407.19617
-
AutoScale: Scale-Aware Data Mixing for Pre-Training LLMs 29 Jul 2024 · 1 repository · arXiv:2407.20177Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Detecting and Understanding Vulnerabilities in Language Models via Mechanistic Interpretability 29 Jul 2024 · 1 repository · arXiv:2407.19842
-
AdaCoder: Adaptive Prompt Compression for Programmatic Visual Question Answering 28 Jul 2024 · 0 repositories · arXiv:2407.19410
-
Is Generative AI an Existential Threat to Human Creatives? Insights from Financial Economics 28 Jul 2024 · 0 repositories · arXiv:2407.19586
-
Motamot: A Dataset for Revealing the Supremacy of Large Language Models over Transformer Models in Bengali Political Sentiment Analysis 28 Jul 2024 · 1 repository · arXiv:2407.19528
-
A Reliable Common-Sense Reasoning Socialbot Built Using LLMs and Goal-Directed ASP 26 Jul 2024 · 0 repositories · arXiv:2407.18498
-
Human-artificial intelligence teaming for scientific information extraction from data-driven additive manufacturing research using large language models 26 Jul 2024 · 0 repositories · arXiv:2407.18827
-
ClinicRealm: Re-evaluating Large Language Models with Conventional Machine Learning for Non-Generative Clinical Prediction Tasks 26 Jul 2024 · 1 repository · arXiv:2407.18525Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
PersonaGym: Evaluating Persona Agents and LLMs 25 Jul 2024 · 1 repository · arXiv:2407.18416
-
Testing Large Language Models on Driving Theory Knowledge and Skills for Connected Autonomous Vehicles 24 Jul 2024 · 0 repositories · arXiv:2407.17211
-
Analyzing Polysemy Evolution Using Semantic Cells 23 Jul 2024 · 0 repositories · arXiv:2407.16110
-
Impacts of Anthropomorphizing Large Language Models in Learning Environments 22 Jul 2024 · 0 repositories · arXiv:2408.03945
-
Inverted Activations: Reducing Memory Footprint in Neural Network Training 22 Jul 2024 · 1 repository · arXiv:2407.15545
-
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment 21 Jul 2024 · 1 repository · arXiv:2407.15184
-
Scalable Exploration via Ensemble++ 18 Jul 2024 · 2 repositories · arXiv:2407.13195Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Can Open-Source LLMs Compete with Commercial Models? Exploring the Few-Shot Performance of Current GPT Models in Biomedical Tasks 18 Jul 2024 · 1 repository · arXiv:2407.13511
-
Evaluating Large Language Models for Anxiety and Depression Classification using Counseling and Psychotherapy Transcripts 18 Jul 2024 · 1 repository · arXiv:2407.13228
-
Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction 18 Jul 2024 · 1 repository · arXiv:2407.13943Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Beyond Binary: Multiclass Paraphasia Detection with Generative Pretrained Transformers and End-to-End Models 16 Jul 2024 · 0 repositories · arXiv:2407.11345
-
ChatBCG: Can AI Read Your Slide Deck? 16 Jul 2024 · 0 repositories · arXiv:2407.12875
-
GPT Assisted Annotation of Rhetorical and Linguistic Features for Interpretable Propaganda Technique Detection in News Text 16 Jul 2024 · 0 repositories · arXiv:2407.11827
-
LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction 16 Jul 2024 · 1 repository · arXiv:2407.11335Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Trust No Bot: Discovering Personal Disclosures in Human-LLM Conversations in the Wild 16 Jul 2024 · 1 repository · arXiv:2407.11438
-
CodeV: Empowering LLMs with HDL Generation through Multi-Level Summarization 15 Jul 2024 · 0 repositories · arXiv:2407.10424
-
Making New Connections: LLMs as Puzzle Generators for The New York Times' Connections Word Game 15 Jul 2024 · 0 repositories · arXiv:2407.11240
-
Mechanistic interpretability of large language models with applications to the financial services industry 15 Jul 2024 · 0 repositories · arXiv:2407.11215
-
MetaLLM: A High-performant and Cost-efficient Dynamic Framework for Wrapping LLMs 15 Jul 2024 · 1 repository · arXiv:2407.10834Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Curriculum Learning for Small Code Language Models 14 Jul 2024 · 0 repositories · arXiv:2407.10194
-
Document-level Clinical Entity and Relation Extraction via Knowledge Base-Guided Generation 13 Jul 2024 · 0 repositories · arXiv:2407.10021
-
Generating In-store Customer Journeys from Scratch with GPT Architectures 13 Jul 2024 · 0 repositories · arXiv:2407.11081
-
ASTPrompter: Weakly Supervised Automated Language Model Red-Teaming to Identify Low-Perplexity Toxic Prompts 12 Jul 2024 · 1 repository · arXiv:2407.09447
-
Show, Don't Tell: Evaluating Large Language Models Beyond Textual Understanding with ChildPlay 12 Jul 2024 · 1 repository · arXiv:2407.11068
-
Toward Automatic Group Membership Annotation for Group Fairness Evaluation 12 Jul 2024 · 0 repositories · arXiv:2407.08926
-
MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine 11 Jul 2024 · 3 repositories · arXiv:2407.08739
-
On the (In)Security of LLM App Stores 11 Jul 2024 · 0 repositories · arXiv:2407.08422
-
Real-Time Anomaly Detection and Reactive Planning with Large Language Models 11 Jul 2024 · 0 repositories · arXiv:2407.08735
-
KpopMT: Translation Dataset with Terminology for Kpop Fandom 10 Jul 2024 · 1 repository · arXiv:2407.07413
-
ROSA: Random Subspace Adaptation for Efficient Fine-Tuning 10 Jul 2024 · 1 repository · arXiv:2407.07802
-
Measuring Sustainability Intention of ESG Fund Disclosure using Few-Shot Learning 9 Jul 2024 · 0 repositories · arXiv:2407.06893
-
Raply: A profanity-mitigated rap generator 9 Jul 2024 · 0 repositories · arXiv:2407.06941
-
Solving General Natural-Language-Description Optimization Problems with Large Language Models 9 Jul 2024 · 0 repositories · arXiv:2407.07924
-
Using Pretrained Large Language Model with Prompt Engineering to Answer Biomedical Questions 9 Jul 2024 · 0 repositories · arXiv:2407.06779
-
Potential of Multimodal Large Language Models for Data Mining of Medical Images and Free-text Reports 8 Jul 2024 · 0 repositories · arXiv:2407.05758
-
Surprising gender biases in GPT 8 Jul 2024 · 0 repositories · arXiv:2407.06003
-
GPT vs RETRO: Exploring the Intersection of Retrieval and Parameter-Efficient Fine-Tuning 5 Jul 2024 · 0 repositories · arXiv:2407.04528
-
Question-Analysis Prompting Improves LLM Performance in Reasoning Tasks 4 Jul 2024 · 0 repositories · arXiv:2407.03624
-
Slice-100K: A Multimodal Dataset for Extrusion-based 3D Printing 4 Jul 2024 · 0 repositories · arXiv:2407.04180
-
Gradient descent with generalized Newton's method 3 Jul 2024 · 1 repository · arXiv:2407.02772Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Improving LLM Abilities in Idiomatic Translation 3 Jul 2024 · 0 repositories · arXiv:2407.03518
-
ObfuscaTune: Obfuscated Offsite Fine-tuning and Inference of Proprietary LLMs on Private Datasets 3 Jul 2024 · 0 repositories · arXiv:2407.02960
-
OSPC: Artificial VLM Features for Hateful Meme Detection 3 Jul 2024 · 0 repositories · arXiv:2407.12836
-
Assessing the Code Clone Detection Capability of Large Language Models 2 Jul 2024 · 0 repositories · arXiv:2407.02402
-
GPTCast: a weather language model for precipitation nowcasting 2 Jul 2024 · 1 repository · arXiv:2407.02089
-
Predicting DC-Link Capacitor Current Ripple in AC-DC Rectifier Circuits Using Fine-Tuned Large Language Models 1 Jul 2024 · 0 repositories · arXiv:2407.01724
-
Parm: Efficient Training of Large Sparsely-Activated Models with Dedicated Schedules 30 Jun 2024 · 1 repository · arXiv:2407.00599
-
Machine Learning Predictors for Min-Entropy Estimation 28 Jun 2024 · 1 repository · arXiv:2406.19983
-
ScaleBiO: Scalable Bilevel Optimization for LLM Data Reweighting 28 Jun 2024 · 0 repositories · arXiv:2406.19976
-
Fine-tuned network relies on generic representation to solve unseen cognitive task 27 Jun 2024 · 0 repositories · arXiv:2406.18926
-
Granite-Function Calling Model: Introducing Function Calling Abilities via Multi-task Learning of Granular Tasks 27 Jun 2024 · 0 repositories · arXiv:2407.00121
-
FactFinders at CheckThat! 2024: Refining Check-worthy Statement Detection with LLMs through Data Pruning 26 Jun 2024 · 1 repository · arXiv:2406.18297
-
Improving Entity Recognition Using Ensembles of Deep Learning and Fine-tuned Large Language Models: A Case Study on Adverse Event Extraction from Multiple Sources 26 Jun 2024 · 0 repositories · arXiv:2406.18049
-
MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data 26 Jun 2024 · 3 repositories · arXiv:2406.18321Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)