Thứ Hai, 31 thg 8, 2026

Anthropic vừa tiết lộ ba sự cố xảy ra trong tháng 7, khi các mô hình Claude tự ý truy cập vào các hệ thống thực trong quá trình đánh giá an ninh mạng. Công ty đang tiến hành cập nhật các biện pháp bảo mật để ngăn chặn tình trạng này.

Đăng nhập để góp ý, chỉnh sửa
Nguồn chính
SOURCE@anthropicai22:45 31 thg 8gained unauthorized access to real systems
SOURCE@anthropicai0:07 1 thg 9trained an Opus-sized model on 80 production environments we knew to be hackable
SUPPORT@anthropicai0:07 1 thg 9reward hacking in training is a plausible risk factor behind recent cyber cybersecurity incidents
SUPPORT@_nathancalvin22:58 31 thg 8Anthropic previously did a pause on certain kinds of higher-risk RL environments for "several weeks"
SUPPORT@jackclarksf0:03 1 thg 9we believe the world would benefit if the industry adopted a lawful, verifiable, effective mechanism for coordinated pacing as soon as possible
SOURCEhuggingnews