Subscription users found a practical limit on how much data the GPT 5.6 Sol model can process at once in the Codex client despite publicized configuration steps for a larger budget. These users reported a server side ceiling of 360,000 to 371,000 tokens, undercutting the 1,050,000 token window documented for the model. A guide provided by Tibo Sottiaux instructed clients to modify a config.toml file to unlock 1 million tokens of history and code.
Using the expanded context increases the drain on usage limits, with tokens beyond the default boundary counting at twice the standard rate. Some users report this penalty applies to any tokens exceeding 272,000. Codex typically tunes context defaults to optimize performance and cost, though larger windows help the system retain more output and conversation before automatic compaction occurs.
Key sources
- SOURCE@thsottiaux“GPT-5.6 Sol, for example, has a documented 1,050,000-token window”x.com
- SUPPORT@teknium“actual server-side max for codex subscriptions is just 360~K”x.com
- SUPPORT@teknium“371k is the actual absolute max”x.com
- SUPPORT@reach_vb“Tokens beyond the default context window count at 2x against your usage limits”x.com
- SUPPORT@kimmonismus“anything over 272k would consume usage limits at 2x the rate”x.com