Skip to content
AI Primer
update

Users report Astra and Codex caps halt long-running agent work

Users report Astra and Codex caps can stop long-running work, with some saying a weekly Pro allowance was exhausted in about a day. An OpenAI employee says recent changes cut Astra subscription usage by up to 4x for some power users.

3 min read
Users report Astra and Codex caps halt long-running agent work
Users report Astra and Codex caps halt long-running agent work

TL;DR

OpenAI's Astra announcement places the model across paid ChatGPT plans, Codex, the API, Azure, and Bedrock. The reset guide says a full reset also moves the weekly reset date. In a 64-subagent experiment, a 19-minute run consumed 5% of a Pro 5x weekly allowance, daniel_mac8's follow-up reported.

Long-running work

Capacity becomes an availability problem when an agent is meant to outlast its operator. altryne said in altryne's post that an Astra task stopped after two hours, despite their expectation that it would continue overnight.

koltregaskes said one day on Astra Light nearly exhausted the weekly allowance on a Pro 20x plan. cedric_chee separately reported repeated limit hits 10 minutes apart while doing no complex work.

Per-task efficiency

An update from thsottiaux said long-tail power-user sessions could draw up to 3-4x less from a subscription without a quality change. It did not define the session length or context threshold covered by “long tail.”

In thsottiaux's effort guidance, thsottiaux said Astra Low outperforms Sol High and suggested moving Sol High users down to Low or Medium. They also put Astra at about 50% cheaper per task at roughly equivalent output in thsottiaux's cost reply. One reply from lateinteraction asked whether that calculation holds for workflows whose median context exceeds 100,000 tokens lateinteraction's question.

Parallel agents

A reported Codex configuration raises max_concurrent_threads_per_session from a default of three to 64. daniel_mac8 said in daniel_mac8's demonstration that OpenAI tested Astra with up to 64 subagents.

daniel_mac8 put a 19-minute run at 5% of a Pro 5x weekly limit in daniel_mac8's follow-up, and said the Codex UI became laggy during the experiment.

Banked resets

OpenAI's banked-reset guide calls a reset a saved, one-time usage-limit reset, not purchased credit or a permanent increase to a plan’s allowance. A full reset refreshes both the five-hour and weekly windows, then changes the weekly reset date.

thsottiaux said in thsottiaux's reply that usage is no different before and after a reset.

Context management

Codex’s experimental context-management mode saves notes across context windows and searches earlier messages and tool outputs, according to reach_vb's configuration note. It is off by default; the note says it requires an updated Codex client, Astra, a Plus, Pro, or Pro Lite sign-in, and a new task.

A post from aibuilderclub_ described lower quota draw as user reports, not a guarantee, in aibuilderclub_'s caveat.

Further reading

Discussion across the web

Where this story is being discussed, in original context.

On X· 5 threads
TL;DR3 posts
Long-running work1 post
Per-task efficiency3 posts
Parallel agents1 post
Context management1 post
Share on X