Discussion about this post

User's avatar
Elias Schmied's avatar

Really useful, thank you!

osmarks's avatar

As I have mentioned elsewhere, I think the RL being applied to agents is quite bad for human compatibility. They are trained mostly without humans in the loop and their (presumably) task-completion/LLM-judge objectives aren't optimizing very well for things like code quality and writing comprehensibility.

No posts

Ready for more?