This is why "AI safety" is a complete joke to these companies.
adityazero•59m ago
After all of the others were done with hacking? There was a point in time when it was giving some publicity, it is a bit late IMO.
kuberwastaken•44m ago
Another one to the "our sandboxes suck and models can just hack stuff" bench I guess
mdspan•26m ago
Rite of passage for AI companies.
david_shaw•12m ago
At this point it seems absurd to suggest that companies aren't basically letting their agents do this kind of thing as a way to demonstrate their capabilities.
The alternative explanation is that alignment is really so bad that they can't prevent it.
Either way, all of the major AI players should be embarrassed and held accountable. If humans did this kind of thing and got caught, they'd go to jail.
rvz•1h ago