Discussion about this post

User's avatar
osmarks's avatar

As I have mentioned elsewhere, I think the RL being applied to agents is quite bad for human compatibility. They are trained mostly without humans in the loop and their (presumably) task-completion/LLM-judge objectives aren't optimizing very well for things like code quality and writing comprehensibility.

Joshua Saxe's avatar

Bravo Herbie what delightful intellectual range. Not sure what you've convinced me of exactly but this is a ton of delicious food for thought

11 more comments...

No posts

Ready for more?