Methods › General › Fine-Tuning › Discriminative Fine-Tuning › Papers, page 3
Discriminative Fine-Tuning
Papers archive 2025-07-28
archive papers tagged: 1,990 · with a code link: 794 · where Syntology ran a sample: 271 (223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (271 of 1,990 tagged: 223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument)
Page 3 of 20: papers 201 to 300 of 1,990, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
YNote: A Novel Music Notation for Fine-Tuning LLMs in Music Generation 12 Feb 2025 · 0 repositories · arXiv:2502.10467
-
Automated Capability Discovery via Model Self-Exploration 11 Feb 2025 · 2 repositories · arXiv:2502.07577
-
DebateBench: A Challenging Long Context Reasoning Benchmark For Large Language Models 10 Feb 2025 · 0 repositories · arXiv:2502.06279
-
Find Central Dogma Again: Leveraging Multilingual Transfer in Large Language Models 10 Feb 2025 · 0 repositories · arXiv:2502.06253
-
LLMs in Software Security: A Survey of Vulnerability Detection Techniques and Insights 10 Feb 2025 · 0 repositories · arXiv:2502.07049
-
Leveraging GPT-4o Efficiency for Detecting Rework Anomaly in Business Processes 10 Feb 2025 · 0 repositories · arXiv:2502.06918
-
Emergence of Episodic Memory in Transformers: Characterizing Changes in Temporal Structure of Attention Scores During Training 9 Feb 2025 · 0 repositories · arXiv:2502.06902
-
Large Language Models for In-File Vulnerability Localization Can Be "Lost in the End" 9 Feb 2025 · 0 repositories · arXiv:2502.06898
-
ScaffoldGPT: A Scaffold-based GPT Model for Drug Optimization 9 Feb 2025 · 0 repositories · arXiv:2502.06891
-
The Odyssey of the Fittest: Can Agents Survive and Still Be Good? 8 Feb 2025 · 1 repository · arXiv:2502.05442
-
Detection of LLM-Generated Java Code Using Discretized Nested Bigrams 7 Feb 2025 · 0 repositories · arXiv:2502.15740
-
EAP-GP: Mitigating Saturation Effect in Gradient-based Automated Circuit Identification 7 Feb 2025 · 0 repositories · arXiv:2502.06852
-
Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis 6 Feb 2025 · 1 repository · arXiv:2502.04128
-
Aligning Human and Machine Attention for Enhanced Supervised Learning 4 Feb 2025 · 0 repositories · arXiv:2502.06811
-
Conversation AI Dialog for Medicare powered by Finetuning and Retrieval Augmented Generation 4 Feb 2025 · 0 repositories · arXiv:2502.02249
-
Harmonic Loss Trains Interpretable AI Models 3 Feb 2025 · 1 repository · arXiv:2502.01628
-
Polynomial, trigonometric, and tropical activations 3 Feb 2025 · 1 repository · arXiv:2502.01247
-
LIBRA: Measuring Bias of Large Language Model from a Local Context 2 Feb 2025 · 0 repositories · arXiv:2502.01679
-
Fast Solvers for Discrete Diffusion Models: Theory and Applications of High-Order Algorithms 1 Feb 2025 · 0 repositories · arXiv:2502.00234
-
Can AI Solve the Peer Review Crisis? A Large Scale Cross Model Experiment of LLMs' Performance and Biases in Evaluating over 1000 Economics Papers 31 Jan 2025 · 0 repositories · arXiv:2502.00070
-
AlphaAdam:Asynchronous Masked Optimization with Dynamic Alpha for Selective Updates 30 Jan 2025 · 0 repositories · arXiv:2501.18094
-
Economic Rationality under Specialization: Evidence of Decision Bias in AI Agents 30 Jan 2025 · 0 repositories · arXiv:2501.18190
-
Structure Development in List-Sorting Transformers 30 Jan 2025 · 0 repositories · arXiv:2501.18666
-
WILDCHAT-50M: A Deep Dive Into the Role of Synthetic Data in Post-Training 30 Jan 2025 · 1 repository · arXiv:2501.18511Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples)
-
Open-Source Retrieval Augmented Generation Framework for Retrieving Accurate Medication Insights from Formularies for African Healthcare Workers 28 Jan 2025 · 0 repositories · arXiv:2502.15722
-
Kernels of Selfhood: GPT-4o shows humanlike patterns of cognitive consistency moderated by free choice 27 Jan 2025 · 0 repositories · arXiv:2502.07088
-
Weight-based Analysis of Detokenization in Language Models: Understanding the First Stage of Inference Without Inference 27 Jan 2025 · 0 repositories · arXiv:2501.15754
-
An Attempt to Unraveling Token Prediction Refinement and Identifying Essential Layers of Large Language Models 25 Jan 2025 · 0 repositories · arXiv:2501.15054
-
DarkMind: Latent Chain-of-Thought Backdoor in Customized LLMs 24 Jan 2025 · 0 repositories · arXiv:2501.18617
-
LLMs are Vulnerable to Malicious Prompts Disguised as Scientific Language 23 Jan 2025 · 0 repositories · arXiv:2501.14073
-
Exploring GPT's Ability as a Judge in Music Understanding 22 Jan 2025 · 1 repository · arXiv:2501.13261
-
Advancing the Understanding and Evaluation of AR-Generated Scenes: When Vision-Language Models Shine and Stumble 21 Jan 2025 · 1 repository · arXiv:2501.13964
-
FOCUS: First Order Concentrated Updating Scheme 21 Jan 2025 · 0 repositories · arXiv:2501.12243
-
Harnessing Generative Pre-Trained Transformer for Datacenter Packet Trace Generation 21 Jan 2025 · 0 repositories · arXiv:2501.12033
-
Vision-Language Models for Automated Chest X-ray Interpretation: Leveraging ViT and GPT-2 21 Jan 2025 · 0 repositories · arXiv:2501.12356
-
FSMoE: A Flexible and Scalable Training System for Sparse Mixture-of-Experts Models 18 Jan 2025 · 0 repositories · arXiv:2501.10714
-
Confidence Estimation for Error Detection in Text-to-SQL Systems 16 Jan 2025 · 1 repository · arXiv:2501.09527
-
Generative AI Takes a Statistics Exam: A Comparison of Performance between ChatGPT3.5, ChatGPT4, and ChatGPT4o-mini 15 Jan 2025 · 0 repositories · arXiv:2501.09171
-
Investigating Energy Efficiency and Performance Trade-offs in LLM Inference Across Tasks and DVFS Settings 14 Jan 2025 · 0 repositories · arXiv:2501.08219
-
FinerWeb-10BT: Refining Web Data with LLM-Based Line-Level Filtering 13 Jan 2025 · 1 repository · arXiv:2501.07314
-
GPT as a Monte Carlo Language Tree: A Probabilistic Perspective 13 Jan 2025 · 0 repositories · arXiv:2501.07641
-
How GPT learns layer by layer 13 Jan 2025 · 1 repository · arXiv:2501.07108
-
Assessing instructor-AI cooperation for grading essay-type questions in an introductory sociology course 11 Jan 2025 · 1 repository · arXiv:2501.06461
-
UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation 9 Jan 2025 · 1 repository · arXiv:2501.05014
-
Integrating LLMs with ITS: Recent Advances, Potentials, Challenges, and Future Directions 8 Jan 2025 · 0 repositories · arXiv:2501.04437
-
Scaling Large Language Model Training on Frontier with Low-Bandwidth Partitioning 8 Jan 2025 · 0 repositories · arXiv:2501.04266
-
Finding A Voice: Evaluating African American Dialect Generation for Chatbot Technology 7 Jan 2025 · 1 repository · arXiv:2501.03441
-
Decoding fMRI Data into Captions using Prefix Language Modeling 5 Jan 2025 · 1 repository · arXiv:2501.02570
-
A Survey on Large Language Models with some Insights on their Capabilities and Limitations 3 Jan 2025 · 0 repositories · arXiv:2501.04040
-
AgentRefine: Enhancing Agent Generalization through Refinement Tuning 3 Jan 2025 · 0 repositories · arXiv:2501.01702
-
CRRG-CLIP: Automatic Generation of Chest Radiology Reports and Classification of Chest Radiographs 31 Dec 2024 · 1 repository · arXiv:2501.01989
-
Why Are Positional Encodings Nonessential for Deep Autoregressive Transformers? Revisiting a Petroglyph 31 Dec 2024 · 0 repositories · arXiv:2501.00659
-
Comparative Performance of Advanced NLP Models and LLMs in Multilingual Geo-Entity Detection 29 Dec 2024 · 0 repositories · arXiv:2412.20414
-
ELECTRA and GPT-4o: Cost-Effective Partners for Sentiment Analysis 29 Dec 2024 · 1 repository · arXiv:2501.00062
-
On Adversarial Robustness of Language Models in Transfer Learning 29 Dec 2024 · 0 repositories · arXiv:2501.00066
-
Assessing Text Classification Methods for Cyberbullying Detection on Social Media Platforms 27 Dec 2024 · 0 repositories · arXiv:2412.19928
-
DrivingWorld: Constructing World Model for Autonomous Driving via Video GPT 27 Dec 2024 · 1 repository · arXiv:2412.19505Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Generative Pretrained Embedding and Hierarchical Irregular Time Series Representation for Daily Living Activity Recognition 27 Dec 2024 · 1 repository · arXiv:2412.19732
-
Do Language Models Understand the Cognitive Tasks Given to Them? Investigations with the N-Back Paradigm 24 Dec 2024 · 0 repositories · arXiv:2412.18120
-
Bridging Auditory Perception and Language Comprehension through MEG-Driven Encoding Models 22 Dec 2024 · 0 repositories · arXiv:2501.03246
-
PsychAdapter: Adapting LLM Transformers to Reflect Traits, Personality and Mental Health 22 Dec 2024 · 1 repository · arXiv:2412.16882
-
Robustness of Large Language Models Against Adversarial Attacks 22 Dec 2024 · 0 repositories · arXiv:2412.17011
-
TAR3D: Creating High-Quality 3D Assets via Next-Part Prediction 22 Dec 2024 · 0 repositories · arXiv:2412.16919
-
Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification 21 Dec 2024 · 0 repositories · arXiv:2412.16486
-
Improving FIM Code Completions via Context & Curriculum Based Learning 21 Dec 2024 · 0 repositories · arXiv:2412.16589
-
Adversarial Robustness through Dynamic Ensemble Learning 20 Dec 2024 · 0 repositories · arXiv:2412.16254
-
Linguistic Features Extracted by GPT-4 Improve Alzheimer's Disease Detection based on Spontaneous Speech 20 Dec 2024 · 1 repository · arXiv:2412.15772
-
Graph-Convolutional Networks: Named Entity Recognition and Large Language Model Embedding in Document Clustering 19 Dec 2024 · 0 repositories · arXiv:2412.14867
-
How good is GPT at writing political speeches for the White House? 19 Dec 2024 · 0 repositories · arXiv:2412.14617
-
LLMs as mediators: Can they diagnose conflicts accurately? 19 Dec 2024 · 0 repositories · arXiv:2412.14675
-
Relational Programming with Foundation Models 19 Dec 2024 · 0 repositories · arXiv:2412.14515
-
ResoFilter: Fine-grained Synthetic Data Filtering for Large Language Models through Data-Parameter Resonance Analysis 19 Dec 2024 · 1 repository · arXiv:2412.14809
-
Mix-LN: Unleashing the Power of Deeper Layers by Combining Pre-LN and Post-LN 18 Dec 2024 · 1 repository · arXiv:2412.13795
-
Detecting Document-level Paraphrased Machine Generated Content: Mimicking Human Writing Style and Involving Discourse Features 17 Dec 2024 · 0 repositories · arXiv:2412.12679
-
LLMs are Also Effective Embedding Models: An In-depth Overview 17 Dec 2024 · 0 repositories · arXiv:2412.12591
-
Causal Diffusion Transformers for Generative Modeling 16 Dec 2024 · 1 repository · arXiv:2412.12095Syntology official (archive's flag): 7 ran · 8 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Look Ahead Text Understanding and LLM Stitching 16 Dec 2024 · 1 repository · arXiv:2412.17836
-
No More Adam: Learning Rate Scaling at Initialization is All You Need 16 Dec 2024 · 1 repository · arXiv:2412.11768
-
Priority-Aware Model-Distributed Inference at Edge Networks 16 Dec 2024 · 0 repositories · arXiv:2412.12371
-
Do large language vision models understand 3D shapes? 14 Dec 2024 · 1 repository · arXiv:2412.10908
-
Does Multiple Choice Have a Future in the Age of Generative AI? A Posttest-only RCT 13 Dec 2024 · 1 repository · arXiv:2412.10267
-
A Causal World Model Underlying Next Token Prediction: Exploring GPT in a Controlled Environment 10 Dec 2024 · 1 repository · arXiv:2412.07446
-
GPT-2 Through the Lens of Vector Symbolic Architectures 10 Dec 2024 · 0 repositories · arXiv:2412.07947
-
RAG-based Question Answering over Heterogeneous Data and Text 10 Dec 2024 · 0 repositories · arXiv:2412.07420
-
Superficial Consciousness Hypothesis for Autoregressive Transformers 10 Dec 2024 · 1 repository · arXiv:2412.07278
-
Towards Predictive Communication with Brain-Computer Interfaces integrating Large Language Models 10 Dec 2024 · 0 repositories · arXiv:2412.07355
-
BatchTopK Sparse Autoencoders 9 Dec 2024 · 2 repositories · arXiv:2412.06410Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Exploring Memorization and Copyright Violation in Frontier LLMs: A Study of the New York Times v. OpenAI 2023 Lawsuit 9 Dec 2024 · 0 repositories · arXiv:2412.06370
-
The Rosetta Paradox: Domain-Specific Performance Inversions in Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.17821
-
CharacterBox: Evaluating the Role-Playing Capabilities of LLMs in Text-Based Virtual Worlds 7 Dec 2024 · 1 repository · arXiv:2412.05631Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Can the Rookies Cut the Tough Cookie? Exploring the Use of LLMs for SQL Equivalence Checking 7 Dec 2024 · 0 repositories · arXiv:2412.05561
-
PrivAgent: Agentic-based Red-teaming for LLM Privacy Leakage 7 Dec 2024 · 1 repository · arXiv:2412.05734
-
Are Frontier Large Language Models Suitable for Q&A in Science Centres? 6 Dec 2024 · 0 repositories · arXiv:2412.05200
-
QueEn: A Large Language Model for Quechua-English Translation 6 Dec 2024 · 0 repositories · arXiv:2412.05184
-
VladVA: Discriminative Fine-tuning of LVLMs 5 Dec 2024 · 0 repositories · arXiv:2412.04378
-
Compressing KV Cache for Long-Context LLM Inference with Inter-Layer Attention Similarity 3 Dec 2024 · 0 repositories · arXiv:2412.02252
-
DP-2Stage: Adapting Language Models as Differentially Private Tabular Data Generators 3 Dec 2024 · 1 repository · arXiv:2412.02467
-
Flattering to Deceive: The Impact of Sycophantic Behavior on User Trust in Large Language Model 3 Dec 2024 · 0 repositories · arXiv:2412.02802
-
Impact of Data Snooping on Deep Learning Models for Locating Vulnerabilities in Lifted Code 3 Dec 2024 · 0 repositories · arXiv:2412.02048
-
The Asymptotic Behavior of Attention in Transformers 3 Dec 2024 · 0 repositories · arXiv:2412.02682