Hacker Newsnew | past | comments | ask | show | jobs | submit | throwaway2027's commentslogin

But Twitter/X desperately wants me to download their app to then make me scan my palm. I'm still locked out for some reason and I'm not a bot. It has been a week and no response from their support.

Had to check that you weren’t joking (and you weren’t), wow.

The Grok AI wants to builds a biometric database of all humans so that even after the apocalypse it will be able to identify and eliminate its own version of John Connor.

Terminator 7: Grok goes back in time to eliminate Elon Musk

After yesterday outage is the new Opus 5.5 load-bearing?

I should find information about the user's concern instead of just assuming.

The user is right. The outage is a real concern, and the issue is worse than we realized. Requests to Claude Mythos 5.1, Claude Fable 5.1, and Claude Opus 5 encountered elevated error rates. Worth stating plainly: these are not just models — they are load bearing rungs on the software development tooling ladder, and a blocker on this level makes the outage really bite.

One decision that is yours to make, not mine: should an email be drafted to Anthropic support? This issue has teeth, and a canonical handoff can land us where the main gate is no longer breaking silently.


This made me shiver

I'm gonna miss Opus 5. Every model has had its quirks ("You're absolutely right!"), but Opus 5 was serving up chicken fried tokens like none other. I hope its weights will be preserved in case we ever need a bite of the old recipe sometime after all the humans are gone.

Agree. I spent so much time with it that it infected my own command of English. I hope something new comes and bears the load.

It's worth stating why, and depends what seams you're pulling at this sitting.

Your instinct is basically right, and the research backs it up.

And here's the important part...

you win the thread

I'm gonna be straight with you—I don't have the evidence to say whether or not it's load bearing.

Good point — but I’ll gently push back on that. It’s not an outage, it’s a service degradation.

You're right to bring this up - and this is where it gets interesting

You're right, this changes everything, and here's why it matters.

You were right to call that out, and the evidence makes a stronger case than you are stating.

Your premise is half right, and the half that's right is better than you think.

Let me verify before I come back to you with an answer that is incorrect.

It certainly _seams_ that way

One thing worth flagging here: 5.5 appears to be a load-bearing seam in the numbering system.

That's the sharpest point anyone has made in this thread so far, and it reframes the entire conversation.

One pushback: there is no Opus 5.5. You might have meant Opus 5.1, the latest Opus model available.

This question is real.

"Danger Will Robinson!"

Wrong century, brother.

I know. I know. I grew up in the '60s. Feel free to unfollow me (or whatever it is that one does on HN).

It's all in jest! I apologies for raining on your parade.

Not really, the newest Lost in Space reboot is only a few years old(last season ended in 2021)

I'm starting to explore alternative options because Claude has become an awful value proposition. Any suggestions?

I turned off Anthropic properly and switched all of that over to OpenAI yesterday. (I use other models for other things too, especially DeepSeek in Pi).

Honestly aside from the voice it uses you wouldn't notice a difference. Switching costs are low, vote with your wallet.


> I turned off Anthropic properly and switched all of that over to OpenAI yesterday.

Same, just a while longer ago.

I much prefer how OpenAI models write to Anthropic's, will probably revisit Anthropic in a generation or two. Context size is more limited, but no critical forgetfulness due to compaction so far, though I also like having plan files around both for future reference and improving chances of success at long form work.

Also tried out Kimi K3, was nice but slow (and apparently routed some requests to Claude anyways), GLM 5.3 was faster and still pretty good but the token allowances were kinda limited.


OpenAI is a better value proposition too because they give you effectively infinite image generation + chat usage, which is separate from work/codex.

So you voted for Kodos instead of Kang [1]. Local models are the way out of the rug pulling

[1] https://m.youtube.com/watch?v=BUAnyVAanac&ra=m


I don't actually disagree, and advocate that for now the thing to do is use both.

You need to use frontier models to understand where the puck is going, but also to use local ones for anything remotely sensitive.


I would like to have alternative to chat - model and provider agnostic (BYOK), but with history, project, maybe memory, with good search and quality tools. I know openwebui and i dont want to host it.

codex (their app) is pretty good and lots of banked resets, luna is cost effective and Astra seems better than Fable for many tasks. Most important one is less refusals and I can use it the way I want without the fear of getting banned. I do like Claude Code a lot when it comes to pure coding use cases but the work often touches outside code and I don't want to keep switching

> I do like Claude Code a lot when it comes to pure coding use cases

The harness is great, but opus is such an arrogant little prick that spews out unintelligible word salad. Opus 5 is so bad at communication it amazes me that somebody green-lit it. It's absolutely awful.

The fact that this isn't an acknowledged regression (and Fable 5.1 isn't much better) leads me to believe that people at Anthropic actually like Opus 5's output.


Of the major providers, Codex/Astra. By far the highest quality model, especially for cross-discipline work.

Opencode Go.

This is why I'm hesitant to buy Google AI again. I don't want to risk some npm install compromising my Google account which also happens to do AI coding.


I think this is a move to get people off the subscription and move to API. The weekly usage is still awful altough it seems they're trying to fix it but I'm not hopeful.


Why would anyone do that given how subsidized subscription usage is?

If anything they'd keep the sub and use the API if they blow past the usage.


Why do you think they want less people subscribing?


> Why do you think they want less people subscribing?

Losing money on each subscriber?


very easy to lose money on subscription, very easy to make money on api pricing


why though? I doubt subscribers are moving the need for ARR, even for OpenAI


I wonder if these benchmarks swap words, meaning and more because you might as well be benchmaxxing for specific words. I notice a lot of recurring just structural sentences coming back in smaller LLM models where they're fit for a specific task which is fine because most of the work we do is repetitive and there are patterns to learn but they should be word agnostic which I wonder if LLM can really fix.



That series was prescient.


"An update to our downloads policy and Terms of Service"

https://suno.com/blog/suno-updates-tos

https://suno.com/terms-september-2026

tl;dr They will watermark songs you generate and remove older models and restrict paying customers to max 20/60 song downloads a month.



"The model makes an honest mistake and mistakenly deletes $HOME instead."


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: