Hacker News new | ask | show | jobs
by nomel 18 days ago
I notice there's a whole tech bro speak it gets into anytime it approaches web related tasks around auth, to the point where it stops using complete sentences, making them ambiguous and borderline nonsense. I have to periodically tell it "Speak professionally and use complete sentences." so I can understand what it's even saying. I've even pasted the output to another session to see if it was me, but Claude can't even understand what tech bro web auth Claude writes.
1 comments

I have added a skill to make Opus 4.8 use Opus 4 6 to translate its answers into teadable English, sine it had much higher quality prose. Im seriously considering switching from Claude 20x to Codex because I just cant read more of of Opus or Fable's walls of text
In my case, since "new session" Claude couldn't understand it either, I assume it means it's going to contribute significantly to context rot. In your case, it probably results in a lower quality context.

Hey Anthropic, give agents/skills the ability to modify previous context! This could be used to truly clean reasoning and task definition mistakes, corrected by the user, and strip output to more fundamentals, which should strengthen/purify the context!

CLAUDE_REWRITE_HISTORY=true ;)

Editing agents, and my own, messages is still how I get the best results for high reasoning planning, in chat clients. I remove the mistakes, because I naively assume that if your context has mistakes, now your next token prediction will be biased to a space where the AI makes mistakes! It would be interesting to write some benchmarks around this, but I've seen significant differences when I do this for less capable models.

I haven't been impressed by 5.6 SOL extra. Typically I asked to do a summary of our session, and it went to do a summary of the repo. Tried different things and I just gave up.

Opus 4.8 did the same at first and after it blended better the repo knowledge with the session one.

I have both $20 as I am less a power user these days and I can't justify a 100 let alone 200 but each model as its irks.

I stopped reading Claude's ridiculous distracting nonsense walls of text months ago. Honestly, I prefer Codex being boring and basically ignoring any human utterances I make which are unrelated the work to be done.