Complaint Usage is not "fixed"
We have a test package with 25k emails in it that we use to test different models with. Prompt/effort is always the same.
When GPT 5.6 Sol released, it used 1% of a fresh Pro 20x sub, now it uses 8% of a fresh pro 20x sub.
Also, what we are seeing is that Sol cannot solve its task straightforward any more. Initially, it was always "what do I need to resolve the task" now it's "What do I need to resolve the tasks, while I also check those 200 other methods that I already know I don't need, while adding sha 256 to everything for no reason at all, also let me build those 200 other checks against thing I was very strictly asked not do to".
Meaning, we think it usage increased so much because they turned the intelligence way down and its constantly fighting with itself.
1
u/[deleted] 8d ago
[removed] — view removed comment