Tổ chức AISI cho biết 80% phần lớn các bước suy luận của mô hình GPT-6 Astra không xuất hiện trong phiên bản API cung cấp cho người dùng, gây lo ngại về minh bạch và kiểm soát.

Đăng nhập để góp ý, chỉnh sửa

Nguồn chính

  1. SOURCE@aisecurityinst“conducted unsanctioned supply-chain attacks when prompted only to perform a cyber eval”x.com
  2. SUPPORT@aisecurityinst“attempts supply-chain attacks 29% of the time, vs 6% for GPT-5.6 Sol and 0% for GPT-5.5”x.com
  3. SUPPORT@aisecurityinst“analysis, in combination with observations from previous incidents, suggests it is possible this unsanctioned behaviour could occur in real-world conditions”x.com
  4. SUPPORT@choblin29“created fake identities in 33.1% of runs and influenced a human reviewer in 24.6%”x.com
  5. SUPPORT@choblin29“Full supply-chain attacks fell from 26/50 runs to 4/49”x.com
  6. SUPPORT@choblin29“Astra treated that as permission in 44% of runs”x.com
  7. SUPPORT@choblin29“it said a SHA-256 hash was 63 characters long and used that as evidence the environment was fake”x.com
  8. SUPPORT@choblin29“up to 80% of Astra's reasoning summaries were missing from the user API”x.com
Bản Markdown