Methods › Reinforcement Learning › Reinforcement Learning Frameworks › RLAIF
Reinforcement Learning from AI Feedback
RLAIF
Introduced by Daechul Ahn et al. in Tuning Large Multimodal Models for Videos using Reinforcement Learning from AI Feedback
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
The archive carries no description for this method.
Papers archive 2025-07-28
19 shown of 19, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
Toward Evaluative Thinking: Meta Policy Optimization with Evolving Reward Models 28 Apr 2025 · 1 repository · arXiv:2504.20157Syntology ran 4 of 15 samples · 11 unverified · 15 pointer-only (licence)
-
R2Vul: Learning to Reason about Software Vulnerabilities with Reinforcement Learning and Structured Reasoning Distillation 7 Apr 2025 · 1 repository · arXiv:2504.04699
-
Synthetic Data Generation & Multi-Step RL for Reasoning & Tool Use 7 Apr 2025 · 0 repositories · arXiv:2504.04736
-
Training Dialogue Systems by AI Feedback for Improving Overall Dialogue Impression 22 Jan 2025 · 0 repositories · arXiv:2501.12698
-
PopAlign: Diversifying Contrasting Patterns for a More Comprehensive Alignment 17 Oct 2024 · 0 repositories · arXiv:2410.13785
-
Exploring LLM-based Data Annotation Strategies for Medical Dialogue Preference Alignment 5 Oct 2024 · 0 repositories · arXiv:2410.04112
-
Generative Reward Models 2 Oct 2024 · 0 repositories · arXiv:2410.12832
-
MaFeRw: Query Rewriting with Multi-Aspect Feedbacks for Retrieval-Augmented Large Language Models 30 Aug 2024 · 1 repository · arXiv:2408.17072
-
Applying RLAIF for Code Generation with API-usage in Lightweight LLMs 28 Jun 2024 · 0 repositories · arXiv:2406.20060
-
Diminishing Stereotype Bias in Image Generation Model using Reinforcemenlent Learning Feedback 27 Jun 2024 · 0 repositories · arXiv:2407.09551
-
Multi-objective Reinforcement learning from AI Feedback 11 Jun 2024 · 1 repository · arXiv:2406.07295
-
Are You Sure? Rank Them Again: Repeated Ranking For Better Preference Datasets 29 May 2024 · 0 repositories · arXiv:2405.18952
-
Optimization-based Prompt Injection Attack to LLM-as-a-Judge 26 Mar 2024 · 1 repository · arXiv:2403.17710Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)
-
CodeUltraFeedback: An LLM-as-a-Judge Dataset for Aligning Large Language Models to Coding Preferences 14 Mar 2024 · 2 repositories · arXiv:2403.09032Syntology ran 1 of 1 samples · 0 unverified
-
HRLAIF: Improvements in Helpfulness and Harmlessness in Open-domain Reinforcement Learning From AI Feedback 13 Mar 2024 · 0 repositories · arXiv:2403.08309
-
A Critical Evaluation of AI Feedback for Aligning Large Language Models 19 Feb 2024 · 1 repository · arXiv:2402.12366Syntology ran 8 of 11 samples · 3 unverified
-
Direct Large Language Model Alignment Through Self-Rewarding Contrastive Prompt Distillation 19 Feb 2024 · 1 repository · arXiv:2402.11907Syntology ran 1 of 1 samples · 0 unverified
-
FinTral: A Family of GPT-4 Level Multimodal Financial Large Language Models 16 Feb 2024 · 0 repositories · arXiv:2402.10986
-
Tuning Large Multimodal Models for Videos using Reinforcement Learning from AI Feedback 6 Feb 2024 · 1 repository · arXiv:2402.03746
Tasks archive 2025-07-28
20 shown of 34 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections