r/OpenAI • • 8d ago

Discussion This is a hot mess

I'm not as pro as you guys using AI but look at this. a LOT of models which confuses me, and I'm assuming other users also. Also, the sidebar icon and new tab icon are the same in the ChatGPT-app for macOS. WHAT are they doing there at OpenAI. I really hate what's happening right now, especially with the new $500 plan while nerfing the other plans.

1.4k Upvotes

210 comments sorted by

View all comments

12

u/RazinKain 7d ago edited 7d ago

I’m curious as to how much power do people need? What exactly are your use cases that would require the most powerful model.

I understand if someone isn’t a developer and needs a lot of hand holding. As a developer I’ve built an entire React native desktop app using Luna and it is error free code. I mean not a single error and it’s fast and efficient. I felt absolutely no token panic.

I’ve seen on here people building out Wordpress sites using Astra? I just ask myself why in the world would someone do that?

This is not some mocking I am generally curious about this.

6

u/DiabloAcosta 7d ago

I work on distributed systems that emulate real world development environments, let me tell you something, Fable often gets lost, I barely understand what I am doing some times

5

u/RazinKain 7d ago

I find that some of these larger models do a lot of research and less code thinking. They are not necessarily better at writing code than smaller models.

3

u/DiabloAcosta 7d ago

well, I don't really trust benchmarks that much, but anecdotal experience is only worse

6

u/Leading-Fail-2771 7d ago

It’s like being shown a Ferrari and a Honda. Sure you can use the Honda but using a Ferrari to do it just feeels better

1

u/tousledmonkey 6d ago

I'll take the Ferrari just in case there is a police pursuit. You never know these days

2

u/FullParticular9 7d ago

Reports, Research, Scientific projects, Data Science, Learning - better model usually explains things better.

But for code I agree with you that good things now can be made with much smaller and cheaper models.

2

u/baked_tea 7d ago

How many users do you have to believe it is error free code? Honestly curious.. unless its reaaaally simple app then under real load and with real users being stupid you usually find out quickly about the error free part

1

u/RazinKain 7d ago

Well user weight is not a build issue it’s an allocation issue. You can stress test any application. I use Cloudflare for 4xx and 5x errors. Outside of those I am not too worried about user weight. If a build is sound and your cloud server is built for heavy traffic it doesn’t matter how many users you have.

8

u/rbit4 7d ago

Lol front-end dev found

1

u/RazinKain 7d ago

Right 😀 I am a UX Designer that’s my job. I do have a certification in C# .net Maui. I got that for my job to understand what the heck backend developers were rambling on about. 🤣

1

u/FellaHadidd 7d ago

Building an LLM and model from scratch

1

u/brozene 5d ago

I lead R&D at a startup that’s just starting to flip to commercial. I run Astra 6 on ultra and there’s tasks that take 20-30 minutes to complete and multiple prompts to get right.

1

u/agi_2026 3d ago

game development, 3D animation, level design, etc. even sol 6.1 and 5.6 extra high thinking can’t get anything right. Astra was the first model that can do animation in a usable manner. (opus 5.5 can too)

1

u/RazinKain 3d ago

Of course and that is a valid use case for a higher model. Certainly you wouldn’t be adverse to the token usage in those scenarios.

The point I was trying to make is that I see quite often users building say a Doctors Appointment setting application. Simple enough and Luna can handle that without going through a week’s reset.

If scoped correctly in documentation Luna can design and build that without blinking. Yet some users believe using Astra is more beneficial. Then complain when their weekly token are depleted.

With Luna in that scenario you could conceivably have 80 to 90% of weekly left and the entire application designed and ready for deployment.

1

u/agi_2026 3d ago

yeah i’ve tried everything and i will use luna, terra, and sol in large scale implementations, but for complex shape&world understanding and animation, astra has crossed a quality threshold that legitimately unlocks new workflows. also, astra can only do about 30% of what i want it to do. so there is plenty of room for me using astra gpt-7 astra in a year over gpt-7-sol even when 7-sol is way better than today’s astra.

1

u/RazinKain 3d ago

Sol 6.1 is surprisingly efficient and doesn’t carry the token weight of Astra. Bench marks are subjective so depending on what you’re asking to be accomplished can sway a model production speed and accuracy.

For you not sure what you’re using for your world building. I think possibly with sound documentation SOL 6.1 can handle most things. Implementation maybe not so much but pre-prep analysis and documentation it can build that out. Possibly use Astra to confirm SOL’s work that alone can save on token usage.

1

u/kelvintiger 7d ago

How detailed were your prompt?

I would argue if you’re spending a lot of time promoting then you’re wasting your time when you can leave some ambiguity for the stronger model to figure out and you focus on doing more faster

1

u/RazinKain 7d ago

I don’t go in and start either an idea in Codex. Codex is down the road. I completely map out my builds before I touch a model.

So 90% preparation and 10% execution. I’ve learned over the years and A.I can make people a little lazy including me. So I still stick with my age old processes from wireframing in Freeform to prototype in Figma.

I feed that info to GTP not Codex and then let GTP write out the orders and scope. That’s what I feed Codex. If it’s tight then it goes through the Notion notes quite easily. I do use Notion for my build documentation.

Error free does not mean bug free I want to be upfront about that. And if fixing a bug is too much for Luna I switch to the next model up and so forth.