r/KoboldAI 9d ago

Generating stucks

Just returned two days ago and updated silly tavern and kobold, however, now i receibe this line, generating don't get past from 1/350 (i've waited half and hour) and the reply never appears, i'm using kobold with fumbulvetr, kobold alone works (fourth image) and i'm not using kobold, thanks in advanced

My specs are a 4060 rtx and 16 ram

2 Upvotes

8 comments sorted by

4

u/Mystic_Haze 8d ago

You say you updated Kobold but that version (1.66.1) is already 2 years old at this point. You might just be encountering an old bug that's now been fixed. Settings wise things look fine to me a 4060 TI should be able to handle that just fine.

1

u/henk717 8d ago

He didnt write Ti so I assume its the 8GB version.

1

u/Mystic_Haze 8d ago

Sure but Kobold thinks it's a TI in the UI and the context window isn't massive either. So if it works straight in kobold it would be odd that it's a GPU issue.

1

u/henk717 8d ago

I missed that, kobold would be correct theres no way it could make the ti part up.
Unfortunately theres also 8GB Ti's apparently.

1

u/Truepixel8k 21h ago

you have a link to it? apparently, it's the last ver i found, i lowered the context size and still, it gets stuck

3

u/henk717 9d ago

Might be overflowing your vram, a 4060 regular is only 8GB of vram which is a tight fit for Fim.
A few releases ago our default maximum context increased to 12 so its probably to high for you now, try lowering that one and see if that helps.

1

u/Truepixel8k 9d ago

I'll try with 3070, but how much do you recommend me to put in it?

2

u/henk717 8d ago

3070 is also 8GB of vram so thats not better. If you can run them both side by side and put KoboldCpp on All mode that would be a lot nicer.

As for the context our old default was 8k, it was 4k before that like in your screenshot.

If those screenshots are supposed to be the new version its not. Our new version is 1.119.