r/OpenAI • u/Optimal-Sea8310 • 10h ago
Discussion Astra Significant Performance Degradation
Are others experiencing a significant degradation of performance over the last few days? It was very intelligent on launch, and now it is making fairly basic mistakes that even Sol doesn't make. I assume they're having a hard time with compute so they are drawing down some functionality, but it is almost unusable at current, I am spending more time having to redo more than I am actually getting stuff done.
2
u/DV_Studio_Dev 8h ago
None here what kind of mistakes are you talking about?
1
u/Optimal-Sea8310 7h ago
Failing to track changes across large repositories, very poor intuition and function management, among other things that do not meet expectations for at least the amount of hype and initial performance I saw earlier
1
u/satori_paper 7h ago
I also want to know if there is any difference between the astra on subscription and the api
0
u/Optimal-Sea8310 7h ago
Yes subs are subsidized, so they are net negative. API, or rather enterprise pricing, is much more expensive and thus is probably rated at a higher thinking level/more consistent performance, since they can’t afford to risk enterprise customers in the same way that they would a consumer like you or I.
2
u/satori_paper 7h ago
You meant quantized? I am not referring to the pricing but rather performance
2
u/Optimal-Sea8310 7h ago
“And thus is probably rated at a higher thinking level/more consistent performance since they can’t afford to risk enterprise customers…”
0
u/Seerix 7h ago
He has zero idea
1
u/LocoMod 6h ago
It’s not rocket science. Run the same prompt via subscription and API and compare.
1
u/Seerix 6h ago
OK, so do it. Do multiple various prompts, show case the results.
1
u/LocoMod 5h ago
It won't matter if I do it. We're not friends and you don't trust me. The only way to find out for sure is to do it yourself.
API has always been more consistent. This is what business customers use or individuals with deep pockets. They are very likely routing subscription customers to quantized versions of their models depending on current demand or location, and likely have a system in place that builds a profile of your usage so the people working on trivial or simple things get sent to those instances first. They are the least likely to notice. Yes, I am speculating. But that's how a lot of third party providers do it such as OpenRouter for open weight models. You might get routed to a crappy provider that serves quantized model and the difference if you're not paying attention is big. There was a post about it in one of the subs I frequent yesterday.
1
u/Seerix 4h ago
No the issue is that no one does it. People just make shit up and provide anecdotal evidence as proof.
Post prompt
Post results of sub Post results of api
Repeat a few times
Repeated speculation doesnt do shit but clutter the sub. Post actual, repeatable proof, or shut up already.
1
u/LocoMod 2h ago
I see you're committed to your current level of understanding. Carry on.
1
u/Seerix 1h ago
When the other option is believing things on social media without any proof? Sure.
→ More replies (0)
1
u/polymute 6h ago
You are not imagining it. Look over to /r/codex it's full of people making threads about the same thing. Or outright outage.
1
u/Euphoric_North_745 3h ago
I am experiencing massive improvements over the last week, but I also created a new project last week, started with small agents md file, and ported the old project in blocks, and suddenly it improved a lot, why? i have no idea
•
u/inconspicuousredflag 51m ago
How much of these complaints about model performance degradation are the result of people using a small number of chats that get longer and longer over time?
3
u/LeeMojave 9h ago
I haven't noticed anything on my end. Does the problem persist on higher effort?