Papers
Communities
Events
Blog
Pricing
Search
Open menu
Home
Papers
2501.13011
Cited By
MONA: Myopic Optimization with Non-myopic Approval Can Mitigate Multi-step Reward Hacking
22 January 2025
Sebastian Farquhar
Vikrant Varma
David Lindner
David Elson
Caleb Biddulph
Ian Goodfellow
Rohin Shah
Re-assign community
ArXiv
PDF
HTML
Papers citing
"MONA: Myopic Optimization with Non-myopic Approval Can Mitigate Multi-step Reward Hacking"
1 / 1 papers shown
Title
Higher-Order Belief in Incomplete Information MAIDs
Jack Foxabbott
Rohan Subramani
Francis Rhys Ward
34
0
0
08 Mar 2025
1