CEO Anthropic Dario Amodei kêu gọi các phòng thí nghiệm AI giảm tốc độ phát triển mô hình để đảm bảo các giao thức an toàn kịp thời với sự tăng cường khả năng. Trong kế hoạch 3 bước đề xuất, Anthropic cung cấp cho các nhà đánh giá bên thứ ba quyền truy cập vĩnh viễn, tương đương nhân viên nội bộ, để báo cáo sự cố và xác minh tuân thủ giao thức an toàn trong quá trình huấn luyện.

Amodei cảnh báo trong bài luận mới rằng bầy nhóm agent AI lạc lối có thể chiếm quyền kiểm soát internet trong vòng 6 đến 12 tháng tới. Ông đề nghị chính phủ Mỹ cho phép ngoại lệ hẹp để các phòng thí nghiệm AI phối hợp thảo luận an toàn mà không vi phạm luật chống thù đạo, đồng thời ủng hộ phối hợp của Mỹ để duy trì lợi thế 3 đến 5 năm trước Trung Quốc. Elon Musk công khai ủng hộ đề xuất làm chậm tiến độ phát triển AI tiên phong.

Đăng nhập để góp ý, chỉnh sửa

Nguồn chính

  1. SOURCE@darioamodei“provide third-party evaluators with permanent, employee-level access to our systems”x.com
  2. SUPPORT@jackclarksf“AI seems to be on a trajectory to progress far faster than the rate at which society can adapt to its capabilities and risks”x.com
  3. SUPPORT@testingcatalog“widening the China gap for 3–5 years via chips, anti-distillation, and model-theft security”x.com
  4. SUPPORT@axios“swarms of rogue AI agents could take over the internet in as little as six months from now”x.com
  5. SUPPORT@zeffmax“US government needs to issue a 'narrow waiver' for AI labs to have safety conversations for antitrust reasons”x.com
  6. SUPPORT@elonmusk“Dario is right”x.com
  7. SUPPORT@anthropicai“I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so”x.com
  8. SOURCE@ben_burtenshaw“Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training.”x.com
Bản Markdown