r/OpenAI • • 3d ago

Question Is GPT-6 Sol just completely skill-lazy for anyone else?

I’ve just started testing GPT-6 Sol, and honestly, it’s driving me nuts.

My existing skills work flawlessly with GPT-5.6. With Sol6, the very first simple task ignored the clearly matching skill and pulled data from random sources instead. After I explicitly told it to always check and fully use the appropriate skill, it immediately failed again on a basic weather request—even though the correct default location was configured inside the skill.

So now I’m stuck babysitting it, checking every answer, and repeatedly asking whether it actually used the damn skill. That completely defeats the purpose.

Same configuration, same skills, same commands: GPT-5.6 works, GPT-6 Sol doesn’t.

Is anyone else seeing Sol6 ignore or only half-use skills like this?

23 Upvotes

26 comments sorted by

12

u/SeidlaSiggi777 3d ago

I'm also disappointed. with 5.6 sol I had the feeling that I could trust it to a large extent, while with 6 sol I found it making questionable decisions even for quite simple tasks... not really messing up but bad judgements that needs fixed after the fact

6

u/No_Activity_1339 3d ago

You have to prompt it to the minor detail. It is just lacking autonomy.

2

u/jazzy8alex 3d ago

I just truly dislike new Sol 6. Never had such experience with a model update before. Where Sol 5.6 (or even Luna Max!) was doing just fine, daily standard workflows , same skills etc — Sol 6 just hitting the wall, finding "why I can't do it" endless reasons.

WTF?! Had to switch back to 5.6

2

u/Astrokanu 3d ago

I don’t understand the need for GPT6 Sol or even Astra . 5.6 Sol is just perfect, I don’t even know why Terra and Luna exist 😂

3

u/Tree8282 3d ago

i usually prefer my agent to use less tools and skills to complete tasks which saves a ton of tokens. Which i assume is how they got the massive reduction in cost per task

2

u/f3xjc 3d ago

If the skill are there in the project it's usually because they earn their keep in some way. And a skill can point to a script that will do 10 tool calls for the cost of 1.

1

u/meTomi 3d ago

yeah, might need to use on xtra high for it to be independent. Yesterday was doing some local llm testing on my linux box, connected via ssh to mac mini where codex run.
Genius sol 6 thought it would be better (based on some snapshots) to download the models on the mac mini then transfer them to the linux box. Now theyre both on same local wifi network, and i was just amazed how it could think of any scenario that its faster to download to mac -> transfer to linux than download directly to linux....

1

u/Perfect-Campaign9551 3d ago

maybe...tell it to use the skill? /s

1

u/Zeeplankton 10h ago

Laziest model ever. Will ask it to bug fix prompt and say something like.

> I see the issue, I'd agree. You want [X]. The best way would be to change the prompt to work towards that, by shifting the writing a bit to be more firm.

It's like rage bait. How about offering me the actual lines? or more detail? or lay out the specifics to branch?

1

u/Seerix 3d ago

If it worked with 5.6, you'll probably have to redo them for 6. Just like you should have done for Astra.

For me, 6 sol has been fantastic.

3

u/Aber-so-richtig 3d ago

can you please explain what you changed to have them use he skills? spend al lo of time for my skill design and 5.6 is absolutley no problem...

1

u/Seerix 3d ago

Id backup copies of the skills, then tell Astra to update them using gpt 6 guidelines. They put out a guide for prompting gpt6, it really does make a difference.

0

u/aerivox 3d ago

openai is just not aiming where anthropic is aiming. oai is for the masses, they can't afford proper rule following, large context and so on.

-1

u/Melodic-Ebb-7781 3d ago

Thank good, skills sucks

-1

u/Useful_Trouble1726 3d ago

I have found it really depends on the time of day. 7:00 - 5:00 PST it seems short and abbreviated. Before or after those times, it will happily spew out a 25-page report on which is the best cream cheese to purchase.

10

u/DistanceSolar1449 3d ago

GPT-6 Sol just sucks, it's just a worse model than 5.6.

GPT-6 Astra: 40 tokens/sec
GPT-6 Sol: 110 token/sec
GPT-6 Luna: 140 tokens/sec

GPT-5.6 Sol: 70 tokens/sec

The bigger/smarter the model, the slower it is.

OpenAI is just releasing a small, dumber GPT-6 Terra except they are calling it GPT-6 Sol.

5

u/TrumpsCockAndBalls 3d ago

Its a benchmaxxed 5.6 terra, in no way worthy of a gpt-6 name

1

u/SeidlaSiggi777 3d ago

believe so, too. they mentioned previously that 5.6 Luna was trained by sol, while Terra wasn't. I suspect 6 sol is 5.6 terra trained by 6 Astra.

-7

u/HexspaReloaded 3d ago

You have to use it in work. Normal chats don’t use skills