I wouldn’t be surprised if Opus 5 was trained on content written by other LLMs
1,136 karma · joined November 30, 2017
I wouldn’t be surprised if Opus 5 was trained on content written by other LLMs
Traditional error trackers have two failure modes:
1. False positives: They show you thousands of errors, and you can’t tell the impact on the user
2. False negatives: Many user-facing issues don’t throw exceptions, so they go unnoticed.
Opslane combines error tracking and session recording. And there is an agent that acts on both.
Opslane reduces false positives by ranking issues based on how many users are facing a particular issue. It also learns about your product by reading your code and watching your session recordings.
False negatives are harder. Opslane reviews session recordings to spot frustration. They look for rage clicks, dead clicks, and abandoned forms.
Here is a link to the repo: https://github.com/opslane/opslane
hey! thanks for the feedback. Only for web right now - mobile is on the roadmap.
As for privacy, I get where the concerns are coming from. But as somebody who has used session recording before - it is such a helpful tool to understand how your customers are using your product and struggling.
Also, all the data you get from a session recording - is data you already have (or could easily get).
As for the culture - I completely agree with you. This is why companies like Linear stand out (https://linear.app/now/zero-bugs-policy). One of my beliefs for starting Opslane is that we can get more companies to understand the bugs that matter and help improve the papercuts in their product.
We are recording sessions using rrweb (https://github.com/rrweb-io/rrweb)
As for privacy, I get where the concerns are coming from. But as somebody who has used session recording before - it is such a helpful tool to understand how your customers are using your product and struggling.
Also, all the data you get from a session recording - is data you already have (or could easily get).
It was on the higher end of Anthropics range - closer to 30-40% more tokens
https://www.claudecodecamp.com/p/i-measured-claude-4-7-s-new...
"given that Opus 4.7 on Low thinking is strictly better than Opus 4.6 on Medium, etc., etc.”
Opus 4.7 in general is more expensive for similar usage. Now we can argue that is provides better performance all else being equal but I haven’t been able to see that
I find 5 thinking levels to be super confusing - I dont really get why they went from 3 -> 5
https://platform.claude.com/docs/en/about-claude/pricing
So if you are generating more tokens, you are eating up your usage faster
it seems to hallucinate a bit more (anecdotal)
The limits have always been opaque and you never know when they change.
I started building an open-source local proxy that logs every rate-limit header Claude Code sends.
I am using it to track and get a better sense of the 5h and 7d weekly limits.
Some initial data from 11 observed 5h sessions on Max 20x: - 5h budget: roughly $120–$280 per window - 7d budget: roughly $1,300–$1,900 - Separate Sonnet-only 7d budget at ~$150 - 95% of tokens are cache reads. They barely move the meter.
It’s open source so more people can run it and we can figure out the real numbers.
Parallelism on top of bad context just gets you more wrong answers faster
The problem is if you skip that step and ask Claude to write the tests after.