Analysis of Opus 5's 'Contrition'
slowmovintarget
17 points
3 comments
July 28, 2026
Related Discussions
Found 5 related stories in 59.9ms across 6,278 title embeddings via pgvector HNSW
- Why does Opus 5 feel worse to work with? numeri · 825 pts · August 14, 2026 · 65% similar
- Elevated Errors for Opus 5 TimCTRL · 93 pts · July 26, 2026 · 55% similar
- Benchmarking Opus 5 on SlopCodeBench dhorthy · 216 pts · July 27, 2026 · 54% similar
- I compared Opus 4.8 vs. Opus 5 on 25 of my tasks to see what the difference was bisonbear · 18 pts · August 26, 2026 · 53% similar
- Opus 5.0 drives incoherence into the stratosphere Bluestein · 176 pts · August 19, 2026 · 52% similar
Discussion Highlights (3 comments)
urbandw311er
Does this work? Appreciate honest testimonials from anybody unrelated to the OP or author.
nightshift1
Interesting. I will try some of of those instructions but I don't mind much about the tics. My actual complaint: end-of-turn summaries that are undecipherable unless you watched the session closely. Both Opus and Fable regularly ends hours of work with status updates phrased in crystal clear terms like: "**V2** debug-helper gate (R4d) | Medium-risk | Gate on `?debug=1` URL param at 3 fixture sites — **not** a localStorage flag (`browser_page` clears it every test). 284 debug-API calls ride the fixture, so R4d goes last+alone"
bomewish
Bit frustrating. any kind of benchmark or even a few examples showing that this produces better outputs would have been nice.