Methods › Reinforcement Learning › Reinforcement Learning Frameworks › RLAIF

Reinforcement Learning from AI Feedback

RLAIF

19 papers tagged archive 2025-07-28

Introduced by Daechul Ahn et al. in Tuning Large Multimodal Models for Videos using Reinforcement Learning from AI Feedback

archive 2025-07-28 Description, source and code snippet are the archive's method entry.

The archive carries no description for this method.

PaperSource

Papers archive 2025-07-28

19 shown of 19, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.

Tasks archive 2025-07-28

20 shown of 34 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.

TaskPapers
Reinforcement Learning5
reinforcement-learning5
Language Modelling4
Large Language Model3
Mathematical Reasoning3
Decision Making2
Language Modeling2
Question Answering2
Reinforcement Learning (RL)2
Retrieval2
Code Generation1
Dataset Generation1
Denoising1
Fairness1
Financial Analysis1
GSM8K1
Gender Classification1
Hallucination1
HumanEval1
Image Generation1

Usage over time archive 2025-07-28

Papers per year tagged with RLAIF: 2024 to 2025, peak 15 15 0 2024: 15 papers 2024 2025: 4 papers 2025
Papers per year the archive tags with this method, by the paper's archive date (19 dated). Bars are counts, not a trend claim.

Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).

Categories archive 2025-07-28

Reinforcement Learning Frameworks

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections