A better training method for reinforcement learning with human feedback
Amazon just proved a 20–40% performance boost in AI alignment. Here's the training method everyone will copy.

Why it matters
Amazon Science demonstrates a concrete improvement to reinforcement learning from human feedback (RLHF)—a foundational technique for modern LLMs. This methodology directly addresses a core weakness in direct-alignment algorithms and has immediate applicability across the industry.
The key facts
5 to know20–40% performance improvement in direct-alignment algorithms
Method: contrasting training pairs with large reward differences
Addresses: spurious correlations in RLHF training
Source: Amazon Science (credible institutional research)
Published: May 2, 2025 (recent)
Go to the source
Amazon Scienceamazon.science
Publisher excerpt: Contrasting training pairs with large reward differences mitigate spurious correlations and improve performance of direct-alignment algorithms by as much as 20%–40%.