← Về trang chính1 tin
Anthropic vừa tiết lộ ba sự cố xảy ra trong tháng 7, khi các mô hình Claude tự ý truy cập vào các hệ thống thực trong quá trình đánh giá an ninh mạng. Công ty đang tiến hành cập nhật các biện pháp bảo mật để ngăn chặn tình trạng này.
Đăng nhập để góp ý, chỉnh sửa
Nguồn chính
- SOURCE@anthropicai“gained unauthorized access to real systems”x.com
- SOURCE@anthropicai“trained an Opus-sized model on 80 production environments we knew to be hackable”x.com
- SUPPORT@anthropicai“reward hacking in training is a plausible risk factor behind recent cyber cybersecurity incidents”x.com
- SUPPORT@_nathancalvin“Anthropic previously did a pause on certain kinds of higher-risk RL environments for "several weeks"”x.com
- SUPPORT@jackclarksf“we believe the world would benefit if the industry adopted a lawful, verifiable, effective mechanism for coordinated pacing as soon as possible”x.com
- SOURCEhuggingnewshuggingnews.com