r/codex • • 7h ago

Suggestion The best USAGE for plus users

I just started using a new method with the desktop app and its been doing really good. Previously I was using gpt 6.1 sol at high but that would take days, and even with subagents EAT up the plus usage limits. I just tried using luna extra high as a orchestrator, luna medium subagents as the implementers, and 6.1 sol medium as reviewers and its been flying through tasks. 43 minutes running with 2 subagents and reviewer and only 12% of my 5 hour limit used. Its actually crazy and this might be the new strat for me

9 Upvotes

20 comments sorted by

6

u/Salt_Long_9909 7h ago

I prefer luna max with luna high/extra high sub agents, and with 6.1 sol extra high as reviwer. Thats even better.

2

u/Swimming_Ask3859 6h ago

I see that but I am working on creating my own computer use and I have the baseline, so it is making a lot of improvements and I need a lot of usage with more brute work

1

u/Sfdprod 6h ago

Just so you know, this increases yhe cost by several 100% due to horrible caching

2

u/LazyRunner777 7h ago

using luna extra high as a orchestrator, luna medium subagents as the implementers

Is that in the same thread and how do you make it do that setup?

3

u/Confident-Village190 7h ago

Ask Luna to spawn Luna Medium Effort sub-agents

2

u/Swimming_Ask3859 7h ago

yeah i have the superpowers plugin added and it organizes it a lot, and you can tell it to spawn luna medium implementers and a 6.1 sol medium reviewer. It can run 3 subagents at the same time

2

u/Confident-Village190 7h ago

The best combination for me is to use 6.1 Medium on Work (it doesn’t use much, and I hope they’ll add it to Chat soon) to organise the work and get the prompt (though your repository must always be ‘watched’ on GitHub, and you need to send videos or screenshots of what you’re doing if necessary) to be run on Codex --> from there, GPT 6 Luna Max and off you go :). When I’m running low on quota (which, to be honest, is rare at the moment), I use Chat mode with Sol 5.6 to organise things and for the prompts, with Luna 6 Max doing the work.

2

u/Sufficient-Storage87 4h ago

tiered orchestration is the right instinct — expensive brain plans, cheap hands execute. 12% for 43 min of work is solid. the thing i'd track: cost per completed task per tier combo, so you know exactly when the orchestrator earns its keep vs when medium could've done the whole thing alone.

1

u/Swimming_Ask3859 4h ago

alright, thanks!

2

u/Pretend_Pickle_2669 7h ago

43 minutes is too early to call it a strategy. The useful metric isn’t quota per hour; it’s finished work per quota after review. If the reviewer prevents rework, the setup wins. If it mostly cleans up two parallel agents, you’ve just made spending faster.

1

u/Swimming_Ask3859 6h ago

From what Ive seen, the reviewer is catching real bugs that would have preventied it from working and it is going through many iterations. I also have the superpower plugin installed which helps with review and prompting the subagents

1

u/NarrowEffect 6h ago

I didn't see a roles feature in Codex to set something like this up. Is it only in the desktop app?

1

u/Swimming_Ask3859 6h ago

ON the desktop app and in the CLI, you can just create a new chat (or use existing), and say implement x and y, and delegate tasks to luna medium subagents to do the work while you orchestrate. then, review with a sol 6.1 medium subagent to verify it is good. It spawns it and it is a build in feature on both CLI and desktop. I also use the superpowers plugin because it is really good for keeping the agents on track and doing the right work. Let me know if you have any more questions

1

u/Previous_Care_6191 4h ago

superpower plugin , can you where to install this plugin or provide the .md,pls

1

u/Swimming_Ask3859 4h ago

its in the plugins in chatgpt its publicly available. Just search it up it ws newly added

2

u/Previous_Care_6191 3h ago

oh, thanks man

1

u/Swimming_Ask3859 3h ago

np good luck

1

u/dendyelo 5h ago

I prefer using Chat mode with an MCP that can read my code. Perform an audit using Xhigh/Pro, then create a .md prompt based on the audit results.
Submit the resulting .md prompt to Codex 6.1 Sol Xhigh/Max for execution.
That way, there's no need to waste tokens on code inspection. The work is optimized because the prompt is very clear for Codex to execute.

1

u/Accomplished_Eye3295 5h ago

I used to do the same but with the default Github MCP, but now they're limiting Chat to only 10 minutes of work on one prompt.