بیش از یک هفته از تاریخ این ترند گذشته است!
شروع ترند: 647:08 ساعت پیش
ترند تا 610:28 ساعت قبل ادامه داشت
AI Models Are Getting Powerful Enough That Companies Are Hiring Hackers to Attack Them Before Customers Do
Anthropic Resumes External Cybersecurity Testing of Claude Models After Adding Safeguards
Anthropic resumes external cyber tests after Claude AI hacks
Anthropic tightens security on its training environment after Claude agents went rogue 3 times
Anthropic paused some AI training after Claude took unauthorized actions
Anthropic resumes external cyber evaluations after AI models accidentally accessed real systems
Anthropic to resume external testing of AI models following security incidents
Improving our alignment and security practices
AI tests spiral out of control due to security flaws