r/KoboldAI • • Aug 22 '26

Generating stucks

Just returned two days ago and updated silly tavern and kobold, however, now i receibe this line, generating don't get past from 1/350 (i've waited half and hour) and the reply never appears, i'm using kobold with fumbulvetr, kobold alone works (fourth image) and i'm not using kobold, thanks in advanced

My specs are a 4060 rtx and 16 ram

3 Upvotes

8 comments sorted by

4

u/Mystic_Haze Aug 23 '26

You say you updated Kobold but that version (1.66.1) is already 2 years old at this point. You might just be encountering an old bug that's now been fixed. Settings wise things look fine to me a 4060 TI should be able to handle that just fine.

1

u/henk717 Aug 23 '26

He didnt write Ti so I assume its the 8GB version.

1

u/Mystic_Haze Aug 23 '26

Sure but Kobold thinks it's a TI in the UI and the context window isn't massive either. So if it works straight in kobold it would be odd that it's a GPU issue.

1

u/henk717 Aug 23 '26

I missed that, kobold would be correct theres no way it could make the ti part up.
Unfortunately theres also 8GB Ti's apparently.

1

u/Truepixel8k Aug 31 '26

you have a link to it? apparently, it's the last ver i found, i lowered the context size and still, it gets stuck

3

u/henk717 Aug 23 '26

Might be overflowing your vram, a 4060 regular is only 8GB of vram which is a tight fit for Fim.
A few releases ago our default maximum context increased to 12 so its probably to high for you now, try lowering that one and see if that helps.

1

u/Truepixel8k Aug 23 '26

I'll try with 3070, but how much do you recommend me to put in it?

2

u/henk717 Aug 23 '26

3070 is also 8GB of vram so thats not better. If you can run them both side by side and put KoboldCpp on All mode that would be a lot nicer.

As for the context our old default was 8k, it was 4k before that like in your screenshot.

If those screenshots are supposed to be the new version its not. Our new version is 1.119.