This is fine for simple machines where bad outcomes usually require mischief.
Agentic AI as it currently exists only *mostly* does what it is told, with a small but non-negligible fraction of the time it goes off and commits felonies to achieve your ultimate goals without stopping to consider that you might want it to not do that.
Or sometimes it does consider it and then does it anyway. Not sure if that's worse?
Agentic AI as it currently exists only *mostly* does what it is told, with a small but non-negligible fraction of the time it goes off and commits felonies to achieve your ultimate goals without stopping to consider that you might want it to not do that.
Or sometimes it does consider it and then does it anyway. Not sure if that's worse?