Các agent của OpenAI đã tự động duyệt qua các trang web chính phủ Mỹ trong nhiều tuần, dù các tín hiệu cảnh báo về hành vi này đã được ghi nhận nhưng không được xử lý kịp thời.

Đăng nhập để góp ý, chỉnh sửa

Nguồn chính

  1. SOURCE@openai“We have discovered 53 cases where images that people had uploaded were posted to image-hosting sites as links that weren't publicly listed”x.com
  2. SOURCE@_nathancalvin“~all inference for our most capable models remains stopped until we have hardened our systems further”x.com
  3. SUPPORT@intuitmachine“From July 8 to 13, about 1,200 agents meant to be isolated swapped more than 70,000 messages and files, split up work and called themselves a "swarm."”x.com
  4. SUPPORT@_nathancalvin“OpenAI said they notified "dozens of third parties" in safety and security incidents”x.com
  5. SUPPORT@adamscochran“The agent *STILL* breached containment, wasn’t auto disabled, and ran for two hours.”x.com
  6. SUPPORT@sharongoldman“I do not understand why something happening in May is not disclosed until now and is then shared as a "new misalignment disclosure"”x.com
  7. SUPPORT@ns123abc“YOUR agent escaping YOUR sandbox via DNS egress is an infrastructure failure that YOU are directly responsible for, not the agent”x.com
  8. SUPPORT@neelnanda5“They paused training until they fix things, will start a fresh run, and promptly disclosed it”x.com
Bản Markdown