← Về trang chính1 tin
Model Evaluation and Threat Research (METR) đã ký thỏa thuận với Anthropic để tiến hành điều tra độc lập các sự cố căn bản và mất căn bản liên quan đến các tác nhân AI. Cuộc điều tra sử dụng nhân sự được thuê qua Redwood để phân tích dữ liệu nội bộ và mở rộng kiến thức công cộng về rủi ro an toàn AI.
Kết nối này cụ thể hóa cam kết của CEO Dario Amodei về việc cấp cho các nhà đánh giá bên thứ ba truy cập cấp nhân viên vĩnh viễn vào hệ thống nội bộ nhằm xác minh các biện biện an toàn. Gần đây, Amodei cũng đề xuất một kế hoạch ba bước để làm chậm tốc độ phát triển của ngành AI, đảm bảo đánh giá tốt hơn về sự căn bản của mô hình trong quá trình huấn luyện.
Đăng nhập để góp ý, chỉnh sửa
Nguồn chính
- SOURCE@ethanjperez“reached an agreement with Anthropic to conduct an independent investigation of agent incidents”x.com
- SUPPORT@ajeya_cotra“Several staff from Redwood have been subcontracted by METR to work on this investigation”x.com
- SUPPORT@ryangreenblatt“independent investigation into alignment and misalignment incidents at Anthropic”x.com
- SUPPORT@ryangreenblatt“The scope of the investigation is: https://t.co/tz121CQ0Z9”x.com
- SOURCE@darioamodei“provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures”x.com
- SUPPORT@anthropicai“I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so”x.com