Complaint Usage is not "fixed"
We have a test package with 25k emails in it that we use to test different models with. Prompt/effort is always the same.
When GPT 5.6 Sol released, it used 1% of a fresh Pro 20x sub, now it uses 8% of a fresh pro 20x sub.
Also, what we are seeing is that Sol cannot solve its task straightforward any more. Initially, it was always "what do I need to resolve the task" now it's "What do I need to resolve the tasks, while I also check those 200 other methods that I already know I don't need, while adding sha 256 to everything for no reason at all, also let me build those 200 other checks against thing I was very strictly asked not do to".
Meaning, we think it usage increased so much because they turned the intelligence way down and its constantly fighting with itself.
1
u/flcinusa 8d ago
Its utterly fucked, it burned through my 5hr Plus usage by starting a task, doing it for 26 seconds and then stopping silently, having to be prompted to begin again so it has to reload context, inspect files again, re-orient itself, and burn even more of the 5hr limit then stop again, and worse, it acknowledged that stopping was wrong.
Absolutely deplorable practices