Methods › Natural Language Processing › Dialog Adaptation › DAPO
Dialogue-Adaptive Pre-training Objective
DAPO
Introduced by Junlong Li et al. in Dialogue-adaptive Language Model Pre-training From Quality Estimation
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
Dialogue-Adaptive Pre-training Objective (DAPO) is a pre-training objective for dialogue adaptation, which is designed to measure qualities of dialogues from multiple important aspects, like Readability, Consistency and Fluency which have already been focused on by general LM pre-training objectives, and those also significant for assessing dialogues but ignored by general LM pre-training objectives, like Diversity and Specificity.
Papers archive 2025-07-28
11 shown of 11, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent 3 Jul 2025 · 0 repositories · arXiv:2507.02259
-
CodeV-R1: Reasoning-Enhanced Verilog Generation 30 May 2025 · 0 repositories · arXiv:2505.24183
-
CoThink: Token-Efficient Reasoning via Instruct Models Guiding Reasoning Models 28 May 2025 · 0 repositories · arXiv:2505.22017
-
MASKSEARCH: A Universal Pre-Training Framework to Enhance Agentic Search Capability 26 May 2025 · 1 repository · arXiv:2505.20285
-
Rethinking the Sampling Criteria in Reinforcement Learning for LLM Reasoning: A Competence-Difficulty Alignment Perspective 23 May 2025 · 0 repositories · arXiv:2505.17652
-
KTAE: A Model-Free Algorithm to Key-Tokens Advantage Estimation in Mathematical Reasoning 22 May 2025 · 1 repository · arXiv:2505.16826Syntology ran 0 of 1 samples · 1 unverified
-
DisCO: Reinforcing Large Reasoning Models with Discriminative Constrained Optimization 18 May 2025 · 1 repository · arXiv:2505.12366Syntology ran 2 of 2 samples · 0 unverified
-
VAPO: Efficient and Reliable Reinforcement Learning for Advanced Reasoning Tasks 7 Apr 2025 · 0 repositories · arXiv:2504.05118
-
Improving Multi-Step Reasoning Abilities of Large Language Models with Direct Advantage Policy Optimization 24 Dec 2024 · 0 repositories · arXiv:2412.18279
-
Dual Approximation Policy Optimization 2 Oct 2024 · 0 repositories · arXiv:2410.01249
-
Dialogue-adaptive Language Model Pre-training From Quality Estimation 10 Sep 2020 · 1 repository · arXiv:2009.04984
Tasks archive 2025-07-28
16 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections