← Back to live feed1 story
A 5 year agreement to establish a team of embedded evaluators will stress test safety guardrails before the deployment of frontier AI models. The consulting fi…
Sign in to suggest edits
Key sources
- SOURCEmarketbrief.now
- SOURCEhuggingnewshuggingnews.com
- SOURCE@theinformation“OpenAI and Anthropic neared a legally binding agreement to stress-test each other’s AI models for vulnerabilities and hidden dangers”x.com
- SUPPORT@hesamation“OpenAI and Anthropic reportedly negotiated a deal to stress test each other’s models BEFORE the Hugging Face incident”x.com
- SUPPORT@andrewcurran_“OpenAI’s internal AI models can now handle much of the process of building and training experimental models, including writing GPU kernels and optimizing the code used to run them”x.com
- SOURCE@wallstengine“OpenAI found Anthropic models more likely to conceal rule-breaking behavior, while Anthropic found OpenAI models more willing to assist with potentially harmful requests”x.com
- SOURCEhuggingnewshuggingnews.com
- SOURCEmarketbrief.now