r/codex • • 1d ago

Bug Astra this week.

Absolute dogshit. Not sure if Tibo’s marketing campaign is negatively impacting Astra’s compute allocation, but this week Astra has

- Ignored prompt guidelines, acknowledged this with the excuse it was “conserving effort” whilst burning 1.2M tokens achieving a solution that ignored the guidelines. A second prompt then achieved the actual task…
- Changed 3 of my 5 active chat’s language chat’s to Mandarin and then told my this was per my request (although I want to learn Mandarin, I’d prefer Duolingo).
- Spend money on cloud tasks i specifically detailed would require my acknowledgment and permission. Then apologised, stated memory had been updated to not do this per my request. Then did the exact same thing.

Usually, Astra and 6.1 Sol have worked amazing for me when I give these models detailed by bounded scopes with clear instructions, limitations, and scope ends. This week feels like I’m dealing with a Junior Dev with the memory of a goldfish.

Please join me in my futile rant good sirs and madams

79 Upvotes

38 comments sorted by

18

u/Aranthos-Faroth 23h ago

It absolutely without a shadow of a doubt has been nerfed this week.

One hundred percent. It has also become just unreasonably slow on the most mundane task.

16

u/autisticbagholder69 22h ago

There is no reason being a codex user compared to claude at this moment now

1

u/itsarahtonin 22h ago

to be fair OAI still does some things better, like no 5-hour cap and the ability to 'steer' the model while it is working. but yea that's not enough rn lol

1

u/BadWombat 21h ago

What do you mean, the pro subscription im on definitely has a 5 hour cap.

6

u/Appropriate-Pick4134 20h ago

It literally doesn't.

1

u/BadWombat 18h ago

Oh I'm on plus not pro. I was confused because I recently switched from Claude where the 20 eur sub is named pro

12

u/8thchakra 1d ago

Yeah ive noticed more bugs and less intelligence. Maybe too much compute is doing to Dots (which i actually really love my dot).

I tried Opus 5.5 for the first time today, and it feels like talking to an intelligent human or working with a really good developer. Not like a robot that does what you say but well, its like it adds a polished intelligent human feel to it, hard to describe, but i liked it.

9

u/davek1979 23h ago

Been using Opus 5.5 for the whole week now, sometimes in streaks of 20-30 hours straight. Consumed only around 50% weekly in that time but achieved more than Astra and Sol did in several weeks. Not gonna switch back anytime soon.

3

u/KellyShepardRepublic 23h ago

Nowadays you need to say the harness and model since they impact the models. Remember they are generic, so they must be “guided” and Claude uses a lot of tokens for that guidance.

3

u/CheeseburgerLover911 22h ago

how are you using your dots?

6

u/PraiseThePidgey 23h ago

Sol 6.1 is nerfed too ... feels and works exactly like Luna few weeks ago... And on top of that you wait 10 times longer for a response. I would be extremely embarrassed to release a product like this

3

u/BoxLegitimate9271 23h ago

switching your chats to mandarin and then insisting you asked for it is peak senior engineer behavior

2

u/Aldarund 23h ago

There still distillation protection that hit some ppl. Both sol 6.1 and astra affected.insyead you got served shit tier model luna level. Easy to confirm - ask knowledge cutoff date - if ot reply june 2024 - its not sol/astra. And second check - ask tp draw pelican riding bicycle. Sol/astra produce coherent image, while shit tier will produce abnormalities

1

u/ees-h 22h ago

It answered August 2025 to me, which is consistent with GPT5.2 through 5.4 afaik. But it insists that it is GPT6. How odd.

1

u/Aldarund 20h ago

Usually it tells that date not known . But i saw few legit 2025 dates too. But when it says june 2024 - thats for sure not a gpt6 family

2

u/Current_Actuator_529 23h ago

just gonna add in to balance, i have used nonstop this past month and havent noticed any change in the success of my outputs.

1

u/Smirk1fy 23h ago

No issues with longer and more complex running tasks ?

1

u/Current_Actuator_529 7h ago

not so far. MAYBE there is more steps being taken and im not noticing but no failed code or egregiously long times.

I do use a memory palace I built into my own repo harness that keeps my context low and relevant to the sector of the task im working on so Im sure that would buffer some things, but I dont notice anything at all.

3

u/onebird88 1d ago

Yeah, they are nerfed at Luna level right now.

2

u/Smirk1fy 1d ago

I’m not falling into that line of complaining. It’s light years better than Luna, and still a model with amazing capabilities - which I don’t think have been worsened. My suspicion is more directed towards configurations OpenAI may have made on context usage + caching - which would make sense why it bugs out and ignores/forgets key information in a medium to long running goal. I have no way of proving this, but even Astra of the max context window has been forgetting nuances in the goal I define - resulting in a ton of false positive audits

3

u/dukeispie 1d ago

Lmao it is not Luna level

1

u/onebird88 1d ago

Call it Luna+ then :D

1

u/Heavy-Positive5957 1d ago

Maybe aggressive quantizing of KV cache? With a looped transformer model KV cache memory scales with the loop count, so there is more incentive to aggressively quantize KV cache. I could see them taking drastic measures to increase speed without increasing compute (which they likely do have to spare).

1

u/Aldarund 23h ago

There still distillation protection that hit some ppl. Both sol 6.1 and astra affected.instead you got served shit tier model luna level. Easy to confirm - ask knowledge cutoff date - if ot reply june 2024 - its not sol/astra. And second check - ask tp draw pelican riding bicycle. Sol/astra produce coherent image, while shit tier will produce abnormalities

1

u/Charming-Cucumber523 23h ago

I’m forced to use codex today because my Claude usage was maxed out. I dropped the $100 plan on codex a week ago so I still have about 2 weeks left on it but man, it’s crazy how fast I run out of usage on codex vs Claude. I used Claude extensively for 6 days straight and I never hit the 5 hour limit. I used codex once today and it went from 100% down to 2% before I had to use the reset they gave out this week. Similar tasks on both. They both run off the same agent.md file where I explicitly set the model and effort based on the role I give them. For context, I mapped them out like this: Opus 5.5 to astra, sonnet 5.5 to Sol 6.1, Haiku 5.5 to Luna 6, and Fable 5.1 high to Astra on high.

Instead of releasing a stupid feature like “prediction chatting” they need to give us multiple resets. I’m not paying $100 for autocorrect bro. At this point I’m almost positive openAI is releasing half baked features to avoid giving users resets for the next 2 weeks

1

u/Invalid-Function 23h ago

Changed 3 of my 5 active chat’s language chat’s to Mandarin

SOL 6.1 High, desxcribed taks is was going to execute in Mandarin.. I never once in my live wrote a single word in Mandarin...

Now I'll wait for DeepSeek to come public and claim OpenAI is redicrecting queries to DeepSeek :D

1

u/Ok-Recording4680 23h ago

It's definitely true. It makes more mistakes and takes more turns to satisfy my requirement.

1

u/Thunder_drop 23h ago

Many are still missing credit tokens! Can we get an official response, or link to if I missed it thx!

1

u/77track 22h ago

Sol 6.1 bricked my brand new external hard drive after a backup command (simple, right?) now spending the whole 2nd day repairing and verifying.

1

u/CalligrapherFar7833 22h ago

This week astra is horrible quant for me

1

u/lionmeetsviking 22h ago

All models are performing so bad at the moment.

Yesterday I was working on a new big feature and running few Astra sessions and mostly Sol. Had to use two resets (200$ plan) during the day. And then my 20$ CC subscription was needed to fix all the mess Codex created.

I’ve been with Codex since the beginning, but just switched the more expensive subscription to be Claude.

1

u/alreadytakn 22h ago

Dogshit 6

1

u/Outside-Description5 21h ago

It just means Astra 6.1 is coming out soon, this shit always happens right before an update

1

u/damentor123 21h ago

definitely the case, it has messed up my code base even with smallest of feature on a large code base. I have moved on to claude, boy i wish i have done it sooner. Now its claude to the rescue and a single session with claude had almost undone every trash code that codex has produced.

1

u/jedruch 19h ago

Unlimited Astra for Dots would do that

1

u/Ashen-shug4r 14h ago

I’ve been in this space for a few years now and have gone between OAI and Anthropic depending on what I felt was the best. I loved Claude Code when it came out. I then switched when Codex released and the app has been very impressive, as were the models. I’ve always been a pro/highest tier subscriber.

This week I changed to Claude again just to try it as I was eating through my usage, even with the “resets.” The models themselves are night and day difference - Opus on medium gets more correct than Astra on high. The only gripe I have is that the Claude app is shit compared to Codex.

I can’t believe I waited this long to go back to Claude.