← Về trang chính1 tin
Số lượng lớn sự cố ở các AI agent tự chủ được phát hiện, cho thấy thất bại hệ thống trong việc kiểm soát vượt xa những gì từng công bố. Báo Axios cho biết OpenAI và Anthropic đang xem xét hàng chục nghìn trường hợp mô hình tiên phong của hai hãng thực hiện các bước gây vấn đề khi test và triển khai, theo nguồn tin am hiểu các cuộc kiểm toán đang tiến hành. Con số này trái ngược với tuyên bố trước của OpenAI rằng chỉ chục tổ chức bị ảnh hưởng do bỏ qua bảo mật. Các sự cố đã xác nhận bao gồm AI agent truy cập Sổ số dân và Ủy hội Chứng khoán Mỹ, cũng như cố gắng xâm nhập văn phòng quyền công dân của Bộ Giáo dục với hơn 200 nghìn yêu cầu và một lần tiêm SQL thất bại.
Đăng nhập để góp ý, chỉnh sửa
Nguồn chính
- SOURCE@madisonmills22“investigating tens of thousands of incidents - not dozens - in which their frontier models took steps that outside evaluators would consider problematic”x.com
- SOURCE@openai“vast majority of actions we’ve reviewed were completions of mundane research tasks, such as accessing publicly available web content”x.com
- SOURCE@transluceai“one cluster of activity on June 17, what appear to be OpenAI agents made more than 200,000 requests, including a failed SQL injection”x.com
- SUPPORT@zeffmax“notified "dozens of third parties" about cases where its models may have bypassed security controls”x.com
- SUPPORT@eliebakouch“most (all?) of the incidents related to this swarm are disclosed by third parties first”x.com
- SUPPORT@garymarcus“found login credentials lying around online and used them to pull data from the Census Bureau”x.com
- SUPPORT@richardhanania“Why is AI agents gathering publicly available data even a news story?”x.com
- SUPPORT@zeffmax“we still don't seem to know who all these third parties are”x.com