upvote
If you run out of sol medium with $100 you're doing something wrong. Astra destroys your usage, I get 1 day of usage with Astra, but 6 sol is almost unlimited and I only use xhigh.
reply
It’s only nearly unlimited if you haven’t just used a banked reset. After a banked reset your weekly usage gets cut by about 80% (not the week you need to wait to get your normal limits back though). ChatGPT has given me a really good reason to cancel.
reply
yeah I use sol constantly and have done maybe $15 of spend in the past week. it's solid and cheaper. this is at least 4-5 investigations, prs, whatever per day.
reply
"You're holding it wrong." Is hardly a retort from a real paying customer having problems with their paid services.

This is why these companies are struggling to make money, they're chastising their customers just like they've been chastising the human race.

reply
> "You're holding it wrong." Is hardly a retort from a real paying customer having problems with their paid services.

It's very appropriate in the cases when you're holding it wrong. The fact that you're paying doesn't mean that you can't make mistakes or waste resources.

reply
If you're compacting every 5 minutes, you have a workflow problem - period.

No LLM will be cost effective if it's compacting this often. You have to find a way around it.

reply
Context window is only 275k or something. And honestly compaction is not that bad in Codex. I often don't even notice I went through 5 compactions in a session.
reply
Sounds like that's the problem then, 275k is a tiny context window. I regularly have sessions that go to 450k or even up to 700k for an unattended overnight Claude Opus session.

Apparently OpenAI makes you manually setup their 1 Million context window, and it seems to be only documented on X:

https://x.com/thsottiaux/status/2089082893804896524

There's at least a forum thread about it here:

https://community.openai.com/t/why-does-codex-report-a-258-4...

reply
Same for me, I started wondering if maybe workflows using compaction instead of clear + markdown memory would be more efficient. Writing a plan or tasks to a file often has the next session repeat part of the exploration, compaction seems to keep most relevant context.
reply
Your tool calls (MCPs?) are very likely too wasteful. Apply some filtering logic on the offending tool’s output. Either a wrapper CLI, or just tell codex how to filter.
reply
you can actually leverage 400k and 1M contexts in codex with very little code changes to the harness. note that excess context past the.. 250k or 400k mark (i don't remember) is charged at 2x the price.
reply
You have a lot of control over compaction, both directly by changing compaction settings, and indirectly by how you structure your codebase/docs so agents use less tokens.
reply
Try Gemini. It’s so cheap I often use my personal AI Pro account for corporate work, and most of the time it doesn’t matter.
reply
Gemini is dumb as hell though, it's not like for like
reply
Gemini 3.8 Flash is actually pretty good.
reply
cheerleader hallucinating agent. That's Gemini.
reply
you can config codex to compact at a higher context limit
reply
As a reference i burn 1% percent for every 40 minutes of sol on average
reply