A gradient-immune structural guard designed to prevent reward hacking in reinforcement learning.
Why it made the shelf
It is an open source tool for securing and guiding reinforcement learning models against reward hacking.
Seen on
Behind the tool
People, backing, growth, and ownership. Every claim links to its source.
Source review queued
Public trail
1 reference from 1 publisher.