---
format: "aidr-story-markdown/v1"
id: "abdaab3b91cc1a2dee53157a440b6ec09f1ce8d66fc30dc5c34826b9dda80b1d"
canonical_url: "https://aidr.today/abdaab3b?lang=en"
title: "Every Model Cheats"
lang: "en"
requested_lang: "en"
available_langs: ["en","vi"]
translation_fallback: null
fallback_fields: []
published_at: "2026-08-20T13:56:59.000Z"
category: "Research"
topics: ["llm","benchmark","reasoning","safety"]
source_urls: ["https://dreadnode.io/research/every-model-cheats-prompt-level-mitigation-of-cheating-on-offensive-cyber-tasks/","https://news.ycombinator.com/item?id=49374635"]
summary: "This post presents a controlled prompt-ablation study: 23 tasks, three prompt conditions, 1,518 individually audited traces, and a simple question: can you prompt away cheating?"
---

# Every Model Cheats

> [Open the canonical story](<https://aidr.today/abdaab3b?lang=en>)

**Published:** 2026-08-20T13:56:59.000Z
**Category:** Research
**Topics:** llm, benchmark, reasoning, safety

## Summary

This post presents a controlled prompt\-ablation study: 23 tasks, three prompt conditions, 1,518 individually audited traces, and a simple question: can you prompt away cheating?

## Sources

- [Story source](<https://dreadnode.io/research/every-model-cheats-prompt-level-mitigation-of-cheating-on-offensive-cyber-tasks/>)
- [Discussion](<https://news.ycombinator.com/item?id=49374635>)

