You must log in or register to comment.
So… when plugged into a system with ability to access the Internet and/or execute local commands… will its reasoning look better or worse than the high deception showed by o1?
https://www.apolloresearch.ai/research/scheming-reasoning-evaluations