Top
Best
New

Posted by corvad 14 hours ago

Claude Status – Elevated errors for multiple models(status.claude.com)
136 points | 104 comments
mariocesar 14 hours ago|
I'm was closing a day of work and it stopped in the middle of a long plan, I started codex and tell it: "I was in the middle of work with Claude could you read the plan and the session "x" in ~/.claude and continue the work" it just completed everything :D
jerpint 13 hours ago||
I’m working on a project that makes switching between coding harnesses essentially unnoticeable. It’s particularly useful when I run out of credits on any given day

The entire platform is skill driven, and based on the premise that state is your local file system. That makes switching harnesses so easy

It’s all open source and has plenty of other features, including inter agent communication, telegram client and much more in the pipeline

https://www.woltspace.com/

calgoo 9 hours ago|||
Not sure what the problem with switching is? When I run out of my kimi session, I just switch model and tell deepseek to continue. Yes I eat the initial cache miss but that's it, no fancy harness required.
ludwik 5 hours ago||
You are talking about switching between models, the parent comment was about switching between harnesses. Not the same thing.
_flux 3 hours ago||
I added a skill to OpenCode that tells it how to read old session logs when asked to. Seems like this would be pretty easy to add to most any harness.
etoxin 10 hours ago||||
I think having a codebase that can use any harness is the optimal setup. I often switch between harnesses and models all the time at work. We have our state/spec stored in git, so we can pack it up on friday and start fresh on monday. Or more commonly, using gemini/opus to create rich specifications and using Luna to implement them. Then switching again for reviews etc.
vasco 8 hours ago|||
No project needed man. There's also a whole ycombinator company for the same thing, skillsync, also useless. These things are either out of the box or they are 2 prompts away.
ammario 10 hours ago|||
Related, have found this handoff skill invaluable for moving tasks between harnesses: https://gist.github.com/ammario/cd57a1d7f6c911e079378c540b38...
adinb 6 hours ago||
My handoff is a handoff_latest.md that points to the latest handoff_datetime.md in a handoff folder.
contentkraft 4 hours ago|||
Both Cursor and Grok build ask if you want to resume a session from another harness. Quite brilliant
RGS1811 14 hours ago|||
I did this with DeepSeek a couple of days ago.
oulu2006 14 hours ago||
yeah same that's my preferred flow now -- sill use Claude and OpenAI a bit, though recent nerfs to subs has really made using codex much harder.

DS 4.1 flash is my main powerhouse and Opus/Astra my auditors (when they're not out of tokens) otherwise K3 or DS4 pro

SOLAR_FIELDS 13 hours ago|||
I started using Pi with Astra on a whim after really enjoying Astra and reading somewhere that you get close to identical results as with the codex harness but for a significant hunk less token usage.

Coming from mostly using Claude models, the terse factual statements coming from Astra via the Pi harness are a breath of fresh air over having to wade through the flowery verbose nonsense that Claude constantly outputs

pdntspa 9 hours ago||
I recently had Astra review a fairly detailed design doc I have for an audio VST fork, that I originally wrote with Opus and/or Fable a few months ago. The doc reaches deep into signal flow and module topology while lifting most of the DSP code from other open-source projects. I had it review for feasibility and architectural soundness.

Astra found a number of flaws that would have come up during implementation and we worked through them. But then I had Fable 5.1 review that document and it found a number of issues with Astra's changes, the least of which had was that Astra duplicated a lot of technical notions that it added rather than using references to an authoritative section. It also flagged some of Astra's designs as technically impossible, pointing out why and I'm actually in the process of digesting its feedback and updating the design spec. (I hand-review each point and we work through a solution together -- I don't trust either model to come up with something that follows my vision on their own)

I'm not promoting one or the other, I just found it interesting how this sort of adversarial review found pretty significant flaws in the other model's work. I am curious as to whether this process will eventually converge on a document that both agree on or if the models are going to perpetually nitpick each other.

I haven't actually started implementation yet, so maybe one or the other is full of shit. Just trying to come up with an architecturally sound design for something I want to write, when I lack the DSP knowledge to be able to write it myself. But the intent is to pass an agent the design doc and list of milestones and let it handle implementation.

cgio 5 hours ago||
In my experience, rather than converging, you end up with a minced up concept. You have to know when to stop the loop. I filter the feedback too, and need to challenge some of the challenges as these models tend to be very conservative. I believe this is intentional, to control AI psychosis, which is indeed quite easy to get. My 2c. If you don’t have good control of what you’re working with you are either searching blind or end up with something basic.
consumer451 13 hours ago|||
Ignorant questions for you and anyone else using DS:

1. who hosts the inference

2. which harness are you using with it, still CC?

RGS1811 23 minutes ago|||
I'm using DeepSeek's own API with opencode. The pricing is absurdly good.
oulu2006 13 hours ago||||
So everyone has their preferred way of doing it.

1. I go direct to source, i.e. DS platform, I find it cheaper than paying the openrouter tax -- I also switch it up a bit

2. I built a local LLM router, that I update with new profiles that have my preferred provider of the week (lowest token costs/speed) with fallbacks, like mimo --> DS4 etc.. if there is overloading,

3. I use 3 diff harnesses, CC/Codex + Opencode -- they all talk to each other through a custom rig system that routes messages between llms using a Rust backed structured JSON system

Not saying this is the best, it's just what I like and works for me^.

I can flow quite naturally between Opus/Astra/K3/GLM/MiMo/DS/etc.. this way and often do...more so these days with subs no longer great as they used to be.

r_lee 3 hours ago||
doesn't DS train on your prompts when you go through their platform?
pavo-etc 13 hours ago||||
Whatever openrouter puts me on, running in pi (though I do all comms over my xmpp wrapper).
logicchains 8 hours ago|||
Deepseek harness is great!
someguyiguess 3 hours ago|||
I don’t understand your point or why people are upvoting this. I’ve done this between many different models. Did you just discover that a frontier model can read context? I genuinely don’t get your point. I’ve had to have Claude agents pick up codex’s context after it hits one of its “cannot connect” issues way more often than the other way around.
mariocesar 3 hours ago||
No point, I just shared something I did
ebbi 13 hours ago||
how do you get the "x" session name?
drewnick 13 hours ago|||
You can even tell the new harness to go find your latest session in the other harness and it works perfectly. You don't even need the name!
consumer451 10 hours ago||||
In CC, "/exit" will tell you how to --resume <session-id> upon exit. This will eat zero tokens, and it actually does not require an active internet connection.
mariocesar 13 hours ago||||
`claude --resume` will list your sessions, you can copy the title. Codex may just list and grep the sessions files in ~/.claude and filter the one with the name.

I also have the habit of naming my sessions with `/rename`

simlevesque 13 hours ago|||
/resume in claude code.
prodigycorp 13 hours ago||
Opus 5.5 and fable 5.2 will release tonight.

gpt-6-sol and Aeon (personal agent) on Thursday. Already preceded by a huge week with step, mimo, grok, and jev releases.

Relentless cycle.

average_r_user 6 hours ago||
I'm still waiting for Anthropic or OpenAI to release their take on OpenClaw.

Meanwhile, Meta's MUSE seems to be gaining traction in the US, while those of us in the EU are once again left watching from the sidelines.

bdcravens 1 hour ago||
OpenClaw isn't the target anymore, Hermes is. It may have been the first, in the same sense that new social networks were often referred to as Facebook-clones, not Myspace (or Friendster) clones.
Razengan 13 hours ago||
In a way it's like the 1980s again, when there were new computers and consoles coming out every week, or the 2000s with all the 3D cards :)
prodigycorp 13 hours ago|||
Whoa, no kidding. Takes me back to the MHz wars we had at school and like you said, the card wars. Excellent perspective.
system2 13 hours ago||
At least after buying 166mmx you didn't get lowered to 133mhz after a month. The progress was legit back then.
AnotherGoodName 12 hours ago||
Hey I bought a ‘cyrix pr166’ so i’m not too sure about that.
kruxigt 8 hours ago||
[dead]
XenophileJKO 5 hours ago||||
I think we feel like the 1980's were fast.. and they were.. but maybe not "this fast"...

https://claude.ai/artifact/6xFW1M4RVPj9rH9Nq8VG7E

Razengan 1 hour ago||
Man that timeline is leaving out a TON of computers and consoles
shepherdjerred 12 hours ago|||
Wow was the 80s really like this? I am interested in AI but it is a bit exhausting to keep up with
Baeocystin 9 hours ago|||
It's hard to overstate how much computers improved with each generation back then. Going from, say, an IBM XT or AT to an Amiga has no modern parallel. It was mind-blowing, exhausting, and exhilarating all at once.
swader999 4 hours ago|||
The Mac M series is more impressive imo. But maybe that's just because it's more recent to me.
Razengan 8 hours ago|||
You could literally count the number of extra colors you would get onscreen with each new computer!
Razengan 8 hours ago|||
In the 1980s and earlier people thought that in 2000 computers would have the kind of AI we finally got in 2025 heh

In 1999 they even made a famous documentary about people in trench coats fighting AI

tombert 13 hours ago||
I don't know if something changed recently, but I have been getting a lot of stuff "flagged by safeguards" in the last two days.

I am quite confident that what I'm doing is well within the law, and I'm not even doing any kind of pen-testing stuff, just some basic reverse engineering, but I can't even use Fable anymore because every time I enable it, it works for about twenty seconds and makes me drop down to Opus 4.8, and often even down to Sonnet.

If anyone here works at Anthropic, did you make the safeguards super sensitive recently?

theophilus76 13 hours ago||
Claude is both on the verge of replacing all software devs and keeping a two 9 SLA
chrisdbanks 13 hours ago||
Your average dev is sick 7 days a year so it's only natural
blitzar 11 hours ago|||
10x dev ... so sick 70 days a year is the target
Hamuko 10 hours ago|||
I wonder if I'm being saved by remote working since I think I have three days in the 365 days, and I've also managed to go years without a single sick day.
Schlagbohrer 4 hours ago||
I hope you are still taking your maximum possible* number of sick days regardless.

*if you live in a real country, which gives all workers unlimited paid sick leave, I am defining "max possible" here as "as many as you can reasonably take without your boss doing something about it"

Hamuko 1 hour ago||
Not really. Although with how soul-crushing working has become lately, maybe I ought to.
nikcub 12 hours ago|||
and the two 9's aren't in order
Schiendelman 13 hours ago|||
Sure, software engineers have less than two nines. :)
prologic 14 hours ago||
Good timing too, Just as xiaomi[1] makes their big announcement :D

[1]: https://news.ycombinator.com/item?id=49792730

ehnto 13 hours ago||
Apologies for the rant, but why do these things constantly hit the front page? It is not interesting, it's not a discussion, and if you are using the models you probably already know.

HN is already a waterfall of AI meta conversations and bike-shedding, now we have to discuss service outages about the AI too?

Can we talk about stuff people are building again, with or without AI, and stop gasping at every minute detail of LLM service providers.

bdcravens 1 hour ago||
Because they are upvoted by readers.

Why are Apple product announcements upvoted? We've known about those changes for months, and they release on a predictable cadence.

Why did we ever care about San Francisco news? Most HNers aren't in SF, California, and many aren't even in the US.

etc

nullc 11 hours ago|||
Gamblers talking about their favorite casino.
system2 13 hours ago||
Everyone is collectively hating on Anthropic, so this news is just adding to the fun.
joegibbs 14 hours ago||
Also for Grok at the same time, must be a problem with Colossus
makeavish 14 hours ago||
There are rumors of Opus 5.5 launching on Tuesday so maybe it due to that? Previously as well there have been incidents just before launch.
Wowfunhappy 14 hours ago||
Do you know where these rumors are coming from? I keep seeing posts like yours on HN and Reddit, which, yes, literally confirms there are rumors, but I can't tell whether they're based on anything.
simlevesque 13 hours ago||
All their API routes for non existing models return 404 and when there's a new model it returns 403.
theGeatZhopa 8 hours ago||
Oh jeez PLEASE opus 5 is the dumbest opus I have ever used. Please make it go away. I lost so much of my hair because of it.
BoxOfRain 7 hours ago||
I wish it were possible to beat Opus 5 over the head with a copy of Orwell's Politics and the English Language sometimes.
camkego 13 hours ago||
When this happens, how often does it cause us to lose our cached pricing? (when we should be getting the cached price)
homo__sapiens 14 hours ago|
Millions of people thought process disrupted.
More comments...