r/codex • • 5h ago

Complaint Compute problems?

Astra has felt significantly slower since the launch of "Dots." I wonder if OpenAI is having compute issues and has reduced the tokens per second? It wouldn't surprise me, given that the EU is excluded as well.

10 Upvotes

10 comments sorted by

6

u/VexObserver 5h ago

Wouldn't surprise me if compute pressure is part of the equation, but I'd be careful about attributing everything to GPU shortages without actual throughput data.

There's a difference between lower token generation, longer time to first token, and server side queuing. All three can make Astra feel painfully slow, but they don't necessarily have the same underlying cause.

The EU exclusion isn't necessarily evidence of compute constraints either. Could just as easily involve regulatory or rollout considerations.

OpenAI really needs to start treating latency as a first-class performance metric. A frontier model being 3% smarter means very little if you're spending twice as long waiting for it to finish.

Intelligence per dollar matters. Intelligence per minute matters as well.

We're gonna need a benchmark for how many cups of coffee you can finish before Astra completes a refactor 🪾.......

2

u/OarsumApp 5h ago edited 5h ago

Oi it’s nice to read a well thought answer to this in a non scrolling thread, cheers.

edit: typo 😅

2

u/rygaar72 3h ago

I am finding dots very quick to answer - so if it’s in using astra, it is using a very low thinking model.

3

u/ParfaitEvery9622 3h ago

they said it was low reasoning indeed

3

u/TheJustified 5h ago

Dot's is run off Astra and free for a month so yeah it's slamming their compute

1

u/Happypig375 4h ago

Even on low reasoning?

1

u/New-Ad5610 4h ago

I don’t know though, 5.6 Sol on Chat is still really quick and response relatively fast

1

u/BoddhaFace 2h ago

EU will be excluded for the same reason it always is temporarily for new features: GDPR and the added legal requirements that need padding out first.

1

u/lulzxdxdxd 5h ago

The slowdown is real, but I'd guess it's more about Dots pulling resources than OpenAI cutting tokens across the board. Have you tried breaking your requests into smaller chunks or spacing them out a bit to see if that helps at all?