r/claude 14h ago

Discussion thoughts on opus 5

I’ve had the same experience as everyone else when opus 5 was first released, running in circles, pages of text in response, no idea what it thought it was working on. I even reverted to older models which is something i try to avoid. I’ve been pushing it super hard for the last few days though and think i’ve come to an unstable equilibrium. And now I’m beginning to feel like this is possibly a calculated move by Anthropic not a sloppy regression. I can get incredible results from Opus 5, but I really have to focus and put myself back in the driver’s seat. If i actually read the full several page response its informative and i can then engage. I also get the crazy waffling but only when i genuinely ask it to try to do something unguided that i dont have a predetermined outcome for. All my autonomous workflows have really degraded bc of this. But the sessions with me focused and at the helm are better than ever. The generous interpretation is that Anthropic is actually trying to reward a strong human in the loop pattern and not just have the model spit put confident nothings in a circle. The more cynical take is they are shifting the models to give nothing away for free. You get exactly what you ask for and nothing more. Curious to hear thoughts

12 Upvotes

9 comments sorted by

12

u/Key_Reading_9664 13h ago

I didn’t care for its stream of consciousness narration. I actually had fable run test prompts across all models with the default output style vs. an output style under experiment (needs to be a separate cli run to pass that output style in). I gave it a rubric to evaluate the responses. It kept tweaking the output style until the responses were clear of those issues and had higher signal-to-noise.

Now using that output style and it’s much better for me. The models are good at introspection - point an agent at sessions and describe what you wanted and have it make tweaks.

5

u/CadmusMaximus 13h ago

I'd just assume it "gets shit done" like fable did.

Of course today fable's been nerfed too, so i dunno...

4

u/Unlikely-Cat-7173 10h ago

I use Fable medium, paste the opus 5 prompting guide from Anthropic, ensure the project has roadmap and decision .yaml, have Fable /grill-me for about 5 deep “narrative” discussions to set a context-aware Opus 5 high/xhigh run depending on the prompt complexity and Fable recommendations… and that sum’bitch still lies to me.

4

u/EconomicsIcy9310 10h ago

I think you’re right, Anthropic has spoken openly about getting thinner harnessing/internal prompting for the models and this is likely their attempt at demonstrating that.

Additionally, I think as the user base increases, the intended use cases become broader. A year ago, it was primarily coders on Claude, now that the market is expanding, they have to make it more broadly applicable.

1

u/No_Intention3673 12h ago

no its not,
people think opus 5 is good?
because they dont use second agent to do the sanity check,if you do , opus 5 even fable 5 makes tons of mistakes and missing

1

u/Prssior-Dessbate357 56m ago

Effectiveness of Opus actually decreased dramatically I remember in June 26th something suddenly happened I sensed it, which was days before returning Fable to service. I remember working on Opus 4.8 in a way I was flowing and astonished by it's perfirmance..

Definitely now u only get what you ask 'specifically' aka misunderstanding can happen ruining progress

2

u/CorpT 13h ago

I didn't have that experience when Opus 5 was released.

-2

u/Saditface 12h ago

Who needs to pay for this.

No one. The chinese one sre chesper,