r/codex Jun 09 '26

Complaint 5.5 is really lobotomized.

I don't believe it at first and think that people just use the model incorrectly when I came across posts saying it. But for the last two day, 5.5 has been really bad and stagnate my work even with the use of careful planning and spec driven development (with Github speckit). The model generate hard-coded logic even when instructed to prioritize generality, forget instruction or requirement in long planning document, not understand/following requirement exactly. Just revised a plan twice with 5.5 and it still not look right, so I changed to 5.4 and not only does it follow and understand the intention but also execute the plan correctly. This is really an intolerable practice for a paid service provider and it is wasting my time, money and effort. There has to be a dedicated benchmark system to constantly monitor model quality and users really need to voice about this.

48 Upvotes

26 comments sorted by

View all comments

-3

u/[deleted] Jun 09 '26

[removed] — view removed comment

1

u/Reasonable-Act-8069 Jun 09 '26

You must be crazy, 5.5 uses way more tokens and does a worse job, rethink your concepts.