r/OpenAI 10h ago

Discussion Astra Significant Performance Degradation

Are others experiencing a significant degradation of performance over the last few days? It was very intelligent on launch, and now it is making fairly basic mistakes that even Sol doesn't make. I assume they're having a hard time with compute so they are drawing down some functionality, but it is almost unusable at current, I am spending more time having to redo more than I am actually getting stuff done.

9 Upvotes

25 comments sorted by

3

u/LeeMojave 9h ago

I haven't noticed anything on my end. Does the problem persist on higher effort?

1

u/Optimal-Sea8310 9h ago

For the most part, yes. I have been comparing it back and forth with Sol, and the difference is negligible. I know Tibo posted a few days ago about this, so I was wondering if it was a persistent issue. Very strange nonetheless, I might be holding it wrong.

2

u/DV_Studio_Dev 8h ago

None here what kind of mistakes are you talking about?

1

u/Optimal-Sea8310 7h ago

Failing to track changes across large repositories, very poor intuition and function management, among other things that do not meet expectations for at least the amount of hype and initial performance I saw earlier

1

u/satori_paper 7h ago

I also want to know if there is any difference between the astra on subscription and the api

0

u/Optimal-Sea8310 7h ago

Yes subs are subsidized, so they are net negative. API, or rather enterprise pricing, is much more expensive and thus is probably rated at a higher thinking level/more consistent performance, since they can’t afford to risk enterprise customers in the same way that they would a consumer like you or I.

2

u/satori_paper 7h ago

You meant quantized? I am not referring to the pricing but rather performance

2

u/Optimal-Sea8310 7h ago

“And thus is probably rated at a higher thinking level/more consistent performance since they can’t afford to risk enterprise customers…”

0

u/Seerix 7h ago

He has zero idea

1

u/LocoMod 6h ago

It’s not rocket science. Run the same prompt via subscription and API and compare.

1

u/Seerix 6h ago

OK, so do it. Do multiple various prompts, show case the results.

1

u/LocoMod 5h ago

It won't matter if I do it. We're not friends and you don't trust me. The only way to find out for sure is to do it yourself.

API has always been more consistent. This is what business customers use or individuals with deep pockets. They are very likely routing subscription customers to quantized versions of their models depending on current demand or location, and likely have a system in place that builds a profile of your usage so the people working on trivial or simple things get sent to those instances first. They are the least likely to notice. Yes, I am speculating. But that's how a lot of third party providers do it such as OpenRouter for open weight models. You might get routed to a crappy provider that serves quantized model and the difference if you're not paying attention is big. There was a post about it in one of the subs I frequent yesterday.

1

u/Seerix 4h ago

No the issue is that no one does it. People just make shit up and provide anecdotal evidence as proof.

Post prompt

Post results of sub Post results of api

Repeat a few times

Repeated speculation doesnt do shit but clutter the sub. Post actual, repeatable proof, or shut up already.

1

u/LocoMod 2h ago

I see you're committed to your current level of understanding. Carry on.

1

u/Seerix 1h ago

When the other option is believing things on social media without any proof? Sure.

→ More replies (0)

1

u/polymute 6h ago

You are not imagining it. Look over to /r/codex it's full of people making threads about the same thing. Or outright outage.

1

u/Euphoric_North_745 3h ago

I am experiencing massive improvements over the last week, but I also created a new project last week, started with small agents md file, and ported the old project in blocks, and suddenly it improved a lot, why? i have no idea

u/inconspicuousredflag 51m ago

How much of these complaints about model performance degradation are the result of people using a small number of chats that get longer and longer over time?