Methods › General › Fine-Tuning › Discriminative Fine-Tuning › Papers, page 4
Discriminative Fine-Tuning
Papers archive 2025-07-28
archive papers tagged: 1,990 · with a code link: 794 · where Syntology ran a sample: 271 (223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (271 of 1,990 tagged: 223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument)
Page 4 of 20: papers 301 to 400 of 1,990, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Su-RoBERTa: A Semi-supervised Approach to Predicting Suicide Risk through Social Media using Base Language Models 2 Dec 2024 · 0 repositories · arXiv:2412.01353
-
The Promise and Peril of Generative AI: Evidence from GPT-4 as Sell-Side Analysts 2 Dec 2024 · 0 repositories · arXiv:2412.01069
-
Tokenizing 3D Molecule Structure with Quantized Spherical Coordinates 2 Dec 2024 · 0 repositories · arXiv:2412.01564
-
A Comprehensive Guide to Explainable AI: From Classical Models to LLMs 1 Dec 2024 · 1 repository · arXiv:2412.00800
-
EventGPT: Event Stream Understanding with Multimodal Large Language Models 1 Dec 2024 · 0 repositories · arXiv:2412.00832
-
CDEMapper: Enhancing NIH Common Data Element Normalization using Large Language Models 30 Nov 2024 · 0 repositories · arXiv:2412.00491
-
Beautimeter: Harnessing GPT for Assessing Architectural and Urban Beauty based on the 15 Properties of Living Structure 28 Nov 2024 · 0 repositories · arXiv:2411.19094
-
Habit Coach: Customising RAG-based chatbots to support behavior change 28 Nov 2024 · 0 repositories · arXiv:2411.19229
-
The Impact of Example Selection in Few-Shot Prompting on Automated Essay Scoring Using GPT Models 28 Nov 2024 · 0 repositories · arXiv:2411.18924
-
Can bidirectional encoder become the ultimate winner for downstream applications of foundation models? 27 Nov 2024 · 0 repositories · arXiv:2411.18021
-
ChatGPT as speechwriter for the French presidents 27 Nov 2024 · 0 repositories · arXiv:2411.18382
-
Streamlining Prediction in Bayesian Deep Learning 27 Nov 2024 · 1 repository · arXiv:2411.18425Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
Advancing Content Moderation: Evaluating Large Language Models for Detecting Sensitive Content Across Text, Images, and Videos 26 Nov 2024 · 0 repositories · arXiv:2411.17123
-
CLOVER: Cross-Layer Orthogonal Vectors Pruning and Fine-Tuning 26 Nov 2024 · 1 repository · arXiv:2411.17426
-
Distributed Sign Momentum with Local Steps for Training Transformers 26 Nov 2024 · 1 repository · arXiv:2411.17866
-
"Give me the code" -- Log Analysis of First-Year CS Students' Interactions With GPT 26 Nov 2024 · 0 repositories · arXiv:2411.17855
-
On Limitations of LLM as Annotator for Low Resource Languages 26 Nov 2024 · 0 repositories · arXiv:2411.17637
-
Pretrained LLM Adapted with LoRA as a Decision Transformer for Offline RL in Quantitative Trading 26 Nov 2024 · 1 repository · arXiv:2411.17900
-
Adaptive Circuit Behavior and Generalization in Mechanistic Interpretability 25 Nov 2024 · 0 repositories · arXiv:2411.16105
-
Are Transformers Truly Foundational for Robotics? 25 Nov 2024 · 0 repositories · arXiv:2411.16917
-
Can AI grade your essays? A comparative analysis of large language models and teacher ratings in multidimensional essay scoring 25 Nov 2024 · 0 repositories · arXiv:2411.16337
-
MarketGPT: Developing a Pre-trained transformer (GPT) for Modeling Financial Time Series 25 Nov 2024 · 1 repository · arXiv:2411.16585Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Predictive Power of LLMs in Financial Markets 25 Nov 2024 · 0 repositories · arXiv:2411.16569
-
Development of Pre-Trained Transformer-based Models for the Nepali Language 24 Nov 2024 · 0 repositories · arXiv:2411.15734
-
"All that Glitters": Approaches to Evaluations with Unreliable Model and Human Annotations 23 Nov 2024 · 1 repository · arXiv:2411.15634
-
Improving Next Tokens via Second-Last Predictions with Generate and Refine 23 Nov 2024 · 0 repositories · arXiv:2411.15661
-
Comparative Analysis of Pooling Mechanisms in LLMs: A Sentiment Analysis Perspective 22 Nov 2024 · 0 repositories · arXiv:2411.14654
-
Assessment of LLM Responses to End-user Security Questions 21 Nov 2024 · 0 repositories · arXiv:2411.14571
-
Evaluating the Robustness of Analogical Reasoning in Large Language Models 21 Nov 2024 · 1 repository · arXiv:2411.14215Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
AI-Driven Agents with Prompts Designed for High Agreeableness Increase the Likelihood of Being Mistaken for a Human in the Turing Test 20 Nov 2024 · 0 repositories · arXiv:2411.13749
-
Combining Autoregressive and Autoencoder Language Models for Text Classification 20 Nov 2024 · 1 repository · arXiv:2411.13282
-
Exploring Large Language Models for Climate Forecasting 20 Nov 2024 · 0 repositories · arXiv:2411.13724
-
Leveraging Virtual Reality and AI Tutoring for Language Learning: A Case Study of a Virtual Campus Environment with OpenAI GPT Integration with Unity 3D 19 Nov 2024 · 0 repositories · arXiv:2411.12619
-
Can Open-source LLMs Enhance Data Synthesis for Toxic Detection?: An Experimental Study 18 Nov 2024 · 0 repositories · arXiv:2411.15175
-
CNMBERT: A Model for Converting Hanyu Pinyin Abbreviations to Chinese Characters 18 Nov 2024 · 1 repository · arXiv:2411.11770
-
VersaTune: An Efficient Data Composition Framework for Training Multi-Capability LLMs 18 Nov 2024 · 1 repository · arXiv:2411.11266
-
Does Prompt Formatting Have Any Impact on LLM Performance? 15 Nov 2024 · 0 repositories · arXiv:2411.10541
-
MARS: Unleashing the Power of Variance Reduction for Training Large Models 15 Nov 2024 · 2 repositories · arXiv:2411.10438Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Take Package as Language: Anomaly Detection Using Transformer 15 Nov 2024 · 0 repositories · arXiv:2412.04473
-
BabyLM Challenge: Exploring the Effect of Variation Sets on Language Model Training Efficiency 14 Nov 2024 · 0 repositories · arXiv:2411.09587
-
Beyond Static Tools: Evaluating Large Language Models for Cryptographic Misuse Detection 14 Nov 2024 · 0 repositories · arXiv:2411.09772
-
LLM App Squatting and Cloning 12 Nov 2024 · 0 repositories · arXiv:2411.07518
-
Autonomous Droplet Microfluidic Design Framework with Large Language Models 11 Nov 2024 · 1 repository · arXiv:2411.06691
-
Explore the Reasoning Capability of LLMs in the Chess Testbed 11 Nov 2024 · 0 repositories · arXiv:2411.06655
-
On Active Privacy Auditing in Supervised Fine-tuning for White-Box Language Models 11 Nov 2024 · 0 repositories · arXiv:2411.07070
-
Prompt-Efficient Fine-Tuning for GPT-like Deep Models to Reduce Hallucination and to Improve Reproducibility in Scientific Text Generation Using Stochastic Optimisation Techniques 10 Nov 2024 · 0 repositories · arXiv:2411.06445
-
Detecting Reference Errors in Scientific Literature with Large Language Models 9 Nov 2024 · 1 repository · arXiv:2411.06101
-
Sufficient Context: A New Lens on Retrieval Augmented Generation Systems 9 Nov 2024 · 0 repositories · arXiv:2411.06037
-
Enhancing Visual Classification using Comparative Descriptors 8 Nov 2024 · 1 repository · arXiv:2411.05357
-
GPT Semantic Cache: Reducing LLM Costs and Latency via Semantic Embedding Caching 8 Nov 2024 · 0 repositories · arXiv:2411.05276
-
Learning the rules of peptide self-assembly through data mining with large language models 8 Nov 2024 · 1 repository · arXiv:2411.05421
-
Adversarial Robustness of In-Context Learning in Transformers for Linear Regression 7 Nov 2024 · 0 repositories · arXiv:2411.05189
-
GPT-Guided Monte Carlo Tree Search for Symbolic Regression in Financial Fraud Detection 7 Nov 2024 · 0 repositories · arXiv:2411.04459
-
Selecting Between BERT and GPT for Text Classification in Political Science Research 7 Nov 2024 · 0 repositories · arXiv:2411.05050
-
Can Custom Models Learn In-Context? An Exploration of Hybrid Architecture Performance on In-Context Learning Tasks 6 Nov 2024 · 1 repository · arXiv:2411.03945
-
From Word Vectors to Multimodal Embeddings: Techniques, Applications, and Future Directions For Large Language Models 6 Nov 2024 · 0 repositories · arXiv:2411.05036
-
On-Device Emoji Classifier Trained with GPT-based Data Augmentation for a Mobile Keyboard 6 Nov 2024 · 0 repositories · arXiv:2411.05031
-
Towards Interpreting Language Models: A Case Study in Multi-Hop Reasoning 6 Nov 2024 · 1 repository · arXiv:2411.05037Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Understanding the Effects of Human-written Paraphrases in LLM-generated Text Detection 6 Nov 2024 · 1 repository · arXiv:2411.03806
-
Enhancing Transformer Training Efficiency with Dynamic Dropout 5 Nov 2024 · 0 repositories · arXiv:2411.03236
-
Exploring the Benefits of Domain-Pretraining of Generative Large Language Models for Chemistry 5 Nov 2024 · 0 repositories · arXiv:2411.03542
-
Ask, and it shall be given: On the Turing completeness of prompting 4 Nov 2024 · 1 repository · arXiv:2411.01992
-
Evaluating the Ability of Large Language Models to Generate Verifiable Specifications in VeriFast 4 Nov 2024 · 0 repositories · arXiv:2411.02318
-
MdEval: Massively Multilingual Code Debugging 4 Nov 2024 · 0 repositories · arXiv:2411.02310
-
Enriching Tabular Data with Contextual LLM Embeddings: A Comprehensive Ablation Study for Ensemble Classifiers 3 Nov 2024 · 0 repositories · arXiv:2411.01645
-
Enhancing Neural Network Interpretability with Feature-Aligned Sparse Autoencoders 2 Nov 2024 · 1 repository · arXiv:2411.01220Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Analyzing & Reducing the Need for Learning Rate Warmup in GPT Training 31 Oct 2024 · 0 repositories · arXiv:2410.23922
-
Automating Quantum Software Maintenance: Flakiness Detection and Root Cause Analysis 31 Oct 2024 · 0 repositories · arXiv:2410.23578
-
Semantic Enrichment of the Quantum Cascade Laser Properties in Text- A Knowledge Graph Generation Approach 30 Oct 2024 · 1 repository · arXiv:2410.22996
-
CFSafety: Comprehensive Fine-grained Safety Assessment for LLMs 29 Oct 2024 · 0 repositories · arXiv:2410.21695
-
Coupling quantum-like cognition with the neuronal networks within generalized probability theory 29 Oct 2024 · 0 repositories · arXiv:2411.00036
-
FactBench: A Dynamic Benchmark for In-the-Wild Language Model Factuality Evaluation 29 Oct 2024 · 0 repositories · arXiv:2410.22257
-
BLAST: Block-Level Adaptive Structured Matrices for Efficient Deep Neural Network Inference 28 Oct 2024 · 1 repository · arXiv:2410.21262
-
Causal Interventions on Causal Paths: Mapping GPT-2's Reasoning From Syntax to Semantics 28 Oct 2024 · 0 repositories · arXiv:2410.21353
-
MultiTok: Variable-Length Tokenization for Efficient LLMs Adapted from LZW Compression 28 Oct 2024 · 1 repository · arXiv:2410.21548
-
Semantic Search Evaluation 28 Oct 2024 · 0 repositories · arXiv:2410.21549
-
A Tutorial on Teaching Data Analytics with Generative AI 25 Oct 2024 · 0 repositories · arXiv:2411.07244
-
Integrating Large Language Models with Internet of Things Applications 25 Oct 2024 · 0 repositories · arXiv:2410.19223
-
Little Giants: Synthesizing High-Quality Embedding Data at Scale 24 Oct 2024 · 1 repository · arXiv:2410.18634Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples)
-
No Argument Left Behind: Overlapping Chunks for Faster Processing of Arbitrarily Long Legal Texts 24 Oct 2024 · 0 repositories · arXiv:2410.19184
-
Understanding Ranking LLMs: A Mechanistic Analysis for Information Retrieval 24 Oct 2024 · 0 repositories · arXiv:2410.18527
-
Scaling up Masked Diffusion Models on Text 24 Oct 2024 · 1 repository · arXiv:2410.18514Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Differentially Private Learning Needs Better Model Initialization and Self-Distillation 23 Oct 2024 · 1 repository · arXiv:2410.17566
-
Future Token Prediction -- Causal Language Modelling with Per-Token Semantic State Vector for Multi-Token Prediction 23 Oct 2024 · 0 repositories · arXiv:2410.18160
-
OmniFlatten: An End-to-end GPT Model for Seamless Voice Conversation 23 Oct 2024 · 1 repository · arXiv:2410.17799
-
DNAHLM -- DNA sequence and Human Language mixed large language Model 22 Oct 2024 · 1 repository · arXiv:2410.16917
-
Exploring Possibilities of AI-Powered Legal Assistance in Bangladesh through Large Language Modeling 22 Oct 2024 · 1 repository · arXiv:2410.17210
-
Developing Retrieval Augmented Generation (RAG) based LLM Systems from PDFs: An Experience Report 21 Oct 2024 · 1 repository · arXiv:2410.15944
-
Improving Neuron-level Interpretability with White-box Language Models 21 Oct 2024 · 0 repositories · arXiv:2410.16443
-
Using GPT Models for Qualitative and Quantitative News Analytics in the 2024 US Presidental Election Process 21 Oct 2024 · 0 repositories · arXiv:2410.15884
-
Does ChatGPT Have a Poetic Style? 20 Oct 2024 · 1 repository · arXiv:2410.15299
-
SDP4Bit: Toward 4-bit Communication Quantization in Sharded Data Parallelism for LLM Training 20 Oct 2024 · 0 repositories · arXiv:2410.15526
-
Bias Amplification: Language Models as Increasingly Biased Media 19 Oct 2024 · 0 repositories · arXiv:2410.15234
-
Judgment of Learning: A Human Ability Beyond Generative Artificial Intelligence 17 Oct 2024 · 0 repositories · arXiv:2410.13392
-
Training Compute-Optimal Vision Transformers for Brain Encoding 17 Oct 2024 · 0 repositories · arXiv:2410.19810
-
Context-Scaling versus Task-Scaling in In-Context Learning 16 Oct 2024 · 0 repositories · arXiv:2410.12783
-
FusionLLM: A Decentralized LLM Training System on Geo-distributed GPUs with Adaptive Compression 16 Oct 2024 · 0 repositories · arXiv:2410.12707
-
Kallini et al. (2024) do not compare impossible languages with constituency-based ones 16 Oct 2024 · 0 repositories · arXiv:2410.12271
-
ShapefileGPT: A Multi-Agent Large Language Model Framework for Automated Shapefile Processing 16 Oct 2024 · 0 repositories · arXiv:2410.12376
-
Stabilize the Latent Space for Image Autoregressive Modeling: A Unified Perspective 16 Oct 2024 · 1 repository · arXiv:2410.12490Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)