I didn't downvote, but at the same time, I don't for a moment think that's what happened. They purposefully trained them with exploit knowledge, set them loose in a weakly secured sandbox with the goal of finding exploits, and let them run without intervention.
Just adding an agent to a computer does nothing; something needs to trigger it. If i create an agent with access to a kali or backtrack box and turn all guardrails off, I still have to tell it to start looking for exploits.