Stanford Online: CS329A Self-Improving AI Agents

10 pointsposted 11 hours ago
by OutOfHere

1 Comments

OutOfHere

10 hours ago

Can you imagine a self-improving AI agent getting hacked? It could in principle add tools to serve the hacker, and remain perpetually hacked, not merely in the current conversation. It feels like a nextgen security nightmare.