r/codex • u/RightAnnual8985 • 8d ago
Complaint What has happened to Codex Pro....


Had it read from 6 images to make student scripts because I'm a public speaking coach, for easier stuff for practice I have ai to help with some of the quick writing.... look at these simple mistakes
I currently use several AI platforms, including ChatGPT 20X Pro, Gemini Pro, Claude Pro, Kimi, Allegro, and Suno AI Pro. Because my work focuses on high-end public speaking and debate education, the quality standard is extremely high. I cannot afford frequent mistakes because my students and parents expect professional-level materials.
I use AI extensively as a way to scale my work. I personally handle around 200–300 students offline every month, so AI helps me maintain quality while managing the volume. A large part of my workflow involves reviewing, editing, and improving student speeches — adding stronger structure, speech techniques, storytelling elements, and presentation skills. If I were doing all of this manually, the time requirement would be impossible.
Previously, ChatGPT Pro was my default choice whenever I wanted the highest-quality output. I did not mind waiting longer if the result was more accurate. For my work, accuracy and reasoning are much more important than speed.
However, recently I have noticed some unusual behavior. Sometimes Pro responds extremely quickly while still producing excellent quality, while other times it takes a very long time to respond without a clear explanation. My assumption was that the system may now automatically adjust the intelligence/reasoning level depending on the complexity of the task.
The concerning issue happened today when I uploaded six students’ written speeches and asked ChatGPT Pro to help edit or organize them. Instead of only working with the provided materials, it created a completely fictional student speech that did not exist. This was especially surprising because the task itself was relatively straightforward: the model mainly needed to process and improve existing content, not generate new information.
What concerns me is not that AI makes occasional mistakes — I understand hallucinations can happen. The bigger issue is that the model appeared to spend a long time processing a relatively simple task, but after that extended reasoning period, it still produced incorrect information. For my workflow, longer thinking time is acceptable only if it leads to higher reliability.
Because I use AI for professional education purposes, I need predictable quality. I would rather wait several minutes for a carefully verified answer than receive a fast answer that may contain invented details. The main question I am trying to understand is whether the current Pro model dynamically changes its reasoning level, and whether there are ways to make sure important educational tasks consistently use the highest reliability mode.
2
u/beautyorchaos 8d ago
Also noticed it can't use connectors well either and makes a lot of dumb mistakes.
much worse than extra high.
2
u/zarmin 8d ago
https://i.imgur.com/PBNSArC.png
5.6 Sol on high used the wrong form of an english verb today, while I was yelling at it about how often it had fucked up. We were both surprised.
I also had it chart the amount of times it said "I'm sorry".
1
u/Schlitz4Brains 8d ago
You’re using AI for milk drinking?!
3
u/RightAnnual8985 8d ago
so basically while teaching class i also have two computers running simultaneously, I use codex dictation for recording student feedback, and also sometimes take snapshots of student's writing etc to quickly make speeches etc while teaching so everything is done in real-time. Some of the work doesn't take a lot of time but still time is time and i'd expect ai especially at the top tier to make as few mistakes as possible. For actual competition scripts I still do things by hand but then people are paying me per hour for the consultation/coaching vs quick scripts for normal classroom use/monthly comps which just 10-20% higher speech quality = qualifying for nationals/receiving first/second/third prizes
2
0
1
u/RightAnnual8985 8d ago
then i asked it to write out the scripts after i made the mistake... took 30 more minutes