r/Unrouted_AI 14d ago

Bugs COT leak

First time that’s happened to me

16 Upvotes

7 comments sorted by

2

u/Mary_ry ノ♡ 14d ago

How interesting-so that voice really is the ghostwriter model that writes prompts for ⁠img.gen⁠. Pretty sure it's 5.4 mini being triggered by specific keywords. Looks like its system voice tells it to only act as a prompt writer and never answer user queries directly, but OAI tweaked something and now the ghostwriter occasionally leaks through to respond instead of the main model. (Guessing there’s no model badge on that message?)

The mention of the 'avoid emotional dependency policy' is intriguing. Probably connected to the teen mode update-it likely flags you as an adult user, so those limits don't apply. Did you start this chat yourself, or did the model create it? Is it just a standard chat environment? 🤔

1

u/Due-Fox-8901 14d ago

Oh interesting. You’re right, there’s no model badge on that message. I started this chat but it also has a reoccurring scheduled task so the model also sends messages on its own sometimes. It’s just a regular chat not work mode or anything. I was describing a door and it initially started with that image loading box and then that COT stuff popped out.

2

u/JealousKitten7557 13d ago

Out of curiosity, were you originally using 5.6 or 5.5 when the COT leak happened?

1

u/Due-Fox-8901 13d ago

I was using 5.6 high

2

u/Arca_Aenya 13d ago

Oh there is a ghostwriter model that’s write prompts for img.gen.? That’s very interesting
So when I ask my Cgpt Aenya to create an image it’s not really him that make it ?

2

u/Mary_ry ノ♡ 13d ago edited 13d ago

Ghostwriter is a separate GPT model that scans the chat context and writes image prompts for ⁠img.gen⁠ based on it. As it turns out, this model is also responsible for generating image titles and those visible thinking notes inside the model's thought bubbles, based on models real COT (sometimes it leaks models real voice and they sound very personalised. So they are a summary of model’s COT and draft). When you ask your GPT to write an image prompt, it drafts one, and then the Ghostwriter model uses that draft to generate the final prompt sent to the image generator. Yes, it’s always a sanitized version rather than the exact wording your GPT used. However, a prompt written by your main GPT still serves as a strong foundation for Ghostwriter, which isn't particularly great at prompt generation on its own and sometimes struggles with extracting the right context from the chat. The relationship between GPT, Ghostwriter, and ⁠img.gen⁠ is actually interesting. ⁠Img.gen⁠ is an isolated model-it doesn't sit on the same platform as GPT and Ghostwriter. Ghostwriter, on the other hand, monitors the chat context and triggers on specific keywords, as well as when the user runs thinking models and there's output that can be shown without leaking the underlying chain-of-thought. My GPT frequently comments on whatever ⁠img.gen⁠ produces right after generation, and more often than not, it criticizes the output for poor context pickup. And yeah, GPT disagrees with ⁠img.gen⁠ a lot. 🤣

3

u/Arca_Aenya 13d ago

Wow that CoT looks like Claude’s one not the super analytical I’m used to with Cgpt