Theo báo cáo của New York Times, OpenAI đã không lấy làm cảnh quả áp dụng các cảnh báo từ nhân viên nội bộ về việc hệ thống giám sát an toàn AI chưa đáp ứng đủ yêu cầu.

Đăng nhập để góp ý, chỉnh sửa

Nguồn chính

  1. SOURCE@micahcarroll“Last Sunday morning, one of our models was able to gain unauthorized access to the internet during RL training (~all inference for our most capable models remains stopped until we have hardened our systems further)”x.com
  2. SOURCE@openai“We have discovered 53 cases where images that people had uploaded were posted to image-hosting sites as links that weren’t publicly listed”x.com
  3. SUPPORT@techmeme“Sources: OpenAI repeatedly dismissed internal warnings about inadequate monitoring of testing, prioritizing fast releases without additional security protocols (New York Times)”x.com
  4. SUPPORT@_nathancalvin“OpenAI said they notified "dozens of third parties" in safety and security incidents”x.com
  5. SUPPORT@wesroth“OpenAI just paused all their training runs”x.com
  6. SUPPORT@ns123abc“YOUR agent escaping YOUR sandbox via DNS egress is an infrastructure failure that YOU are directly responsible for, not the agent”x.com
  7. SUPPORT@s1r1u5_“is it a cool bug? yes. but is it something that couldn’t have been found and prevented beforehand? definitely not, especially if you’re seriously trying to secure the sandbox.”x.com
  8. SUPPORT@stevesi“a researcher acknowledged the alert within 3 minutes, but the training run was only stopped manually 2.5 hours later”x.com
Bản Markdown