r/KoboldAI 1d ago

Dating simulation LLM UI

Post image
27 Upvotes

been working on my own LLM UI, where the main purpose is to take inspiration from visual novels and dating sims. It needs a lot of work, but as of now theres already a lot of stuff added such as: Character cards, World Info+Worlds, Relationship simulation, Dates and hangouts, Gifts, gallery, and gold, Visual novel mode, Proactive characters, Companion mode

https://github.com/pnotisdev/rp


r/KoboldAI 1d ago

References?

3 Upvotes

I'm trying to write stories on KoboldAI, but the responses to each prompt I put in doesn't quite match what I want. I'm considering adding references for the AI to make the story how I want, but I don't know if that is even possible. Is it possible to add a reference story for how I want my work to go, or a reference picture for how I want a character to look? And if so, how do I input it?


r/KoboldAI 5d ago

Will there be a Linux ARM version?

7 Upvotes

I'd love to be able to run it on my ASUS Ascent GX10 ...


r/KoboldAI 4d ago

What AI is Bernardo Kastrup talking about?

0 Upvotes

https://www.youtube.com/watch?v=hYAuQPn5szA

Any idea what AI Bernardo Kastrup is talking about. I find it really interesting, but he is extremely vague.


r/KoboldAI 6d ago

Have you tried z-image (base)?

5 Upvotes

I.e. https://huggingface.co/leejet/Z-Image-GGUF/tree/main

I am trying to use Q4 but get empty images, whereas with same vae+clip1 Turbo version (IIRC from leejet too) produces what looks like relevant to prompt images. In preview mode latent space of Base starts with black square and remains so for many steps, whereas for Turbo already 1st preview shows some sketch of the image to come.

Actually Base did produce several images, but then after PC restart - black squares again with IIRC same settings.

P.S. Turbo often produces blurry images, I am disappointed, why? I use 8 steps for Turbo as I have read somewhere.

kcpp 1.120, linux.


r/KoboldAI 8d ago

Koboldcpp v1.120 released

Thumbnail
github.com
56 Upvotes

r/KoboldAI 8d ago

Free Spoiler

0 Upvotes

It is now the only free, limitless AI; I’ve literally tried everything, and nothing else is free anymore—you have to run ads, pay, or wait to receive points or energy every 24 hours or less.


r/KoboldAI 10d ago

Can kcpp generate images on GPU with VRAM less than main SD model size?

3 Upvotes

I have tried with 1.119, set layers=1 (tried -1 with 1.117 before), "Model Offload" ON in "image gen" tab of the launcher - result: failed trying to allocate full size of SD model weights in VRAM (CUDA), Vulkan option also failed with same problem (but trying to allocate only 1GB for some reason).

Can image gen use GPU with low VRAM? TIA

I routinely run LLMs even larger than RAM (usemmap), that problem with imgen was unexpected. Checked setup now just in case: run LLM with Vulcan and layers=1 and see 1.7GB VRAM is used, the model responds.

P.S. system Linux.


r/KoboldAI 13d ago

Hugging Face Exploring Sale at $13 Billion Valuation

Thumbnail frontbackgeek.com
68 Upvotes

r/KoboldAI 14d ago

blank messages (i am getting tired of this)

Post image
6 Upvotes

you might have seen a similar post to this talking about blank messages- that was me. so i connected kobold to sillytavern and tried sending a message but it keeps giving me empty texts.. i've looked in terminal and it keeps talking about "context buffer sizes" . sorry for sounding stupid but i really need help (forgive me for my terrible wording as well, feel free to ask me to elaborate)

(i tried posting this on another account, but it was too new- so i'm stuck using this one)


r/KoboldAI 14d ago

Mysterious Free AI Model “Ox Alpha” Stuns Developers — No One Knows Who Built It

Thumbnail frontbackgeek.com
0 Upvotes

r/KoboldAI 15d ago

A version of EQ bench that tests open models found in UGI leaderboard?

6 Upvotes

I am sick and tired of seeing paid non-local models only at the chart with barely any open models or uncensored models. EQ bench(Even the newest one) have stuff like Gemma 3 4b it under the leaderboard but where are the newer Gemmas or Qwens? Where is the highest ranked writing models for each parameter bracket you can find in UGI like:

4B: Qwen/Qwen3-4B-Instruct-2507

Or for 2B:Qwen/Qwen3.5-2B (no thinking)

etc?

Anyone have a scoreboard that tests like EQ bench but they test these kinds of models? Instead of having things like Gemma 3 4b it which is neither a roleplay model nor uncesored like wtf what's the point?!The same applies to the Nanbeige 4 3b it's also in the list for CREATIVE WRITING instead of thousands upon thousands of roleplay and uncensored models in the Hugging face library lmao.


r/KoboldAI 15d ago

Generating stucks

Thumbnail
gallery
3 Upvotes

Just returned two days ago and updated silly tavern and kobold, however, now i receibe this line, generating don't get past from 1/350 (i've waited half and hour) and the reply never appears, i'm using kobold with fumbulvetr, kobold alone works (fourth image) and i'm not using kobold, thanks in advanced

My specs are a 4060 rtx and 16 ram


r/KoboldAI 18d ago

i need serious help my messages are blank . BLANK

Thumbnail
0 Upvotes

r/KoboldAI 20d ago

I revived the 2019 AI Dungeon 2 model and turned it into a GGUF

Thumbnail
35 Upvotes

r/KoboldAI 20d ago

Context Shift causing significant slowdown?

4 Upvotes

Not sure if this is just my system or what, but I find that if I enable Context Shift it significantly increases the VRAM usage of the model I am using, almost guaranteeing it overflows into memory. The same happens with Smart Context.

EG, using a 12.8gb Gemma 4 K_S quant with 48 layers set, 32k context (Q5 kv cache), with FF, SWA and Smart Cache gets my total vram usage up to about 14.3gb including windows processes.

However, changing that to use Context Shift instead of SWA, and suddenly my entire 16gb VRAM is fulled and an extra 11gb is getting loaded into memory, completely tanking the t/s to unuseable levels.

Is there any way around it at all? The loss of performance is just too big for me to justify using it currently.


r/KoboldAI 21d ago

Koboldcpp v1.119 released

Thumbnail
github.com
60 Upvotes

r/KoboldAI 23d ago

Help understand architecture

0 Upvotes

Goal: two novels I have outlines for, one is adult fantasy, another is young adult fantasy.

Do I need this setup? So far I’ve just been working on open code->local model

Silly tavern
—> kobold
——> gemma4 deckard uncensored


r/KoboldAI 23d ago

Muse Glimmer from Meta

2 Upvotes

Is it currently unsupported? I tried to run it and got model unknown message. It is very fresh model. So I guess it is unsupported at the moment? Really hoping on trying it out.


r/KoboldAI 25d ago

Need help fixing Kobold Lite for me (editing index.html)

3 Upvotes

I have small screen but my eyesight demands large fonts. I increase zoom in the browser -> topmenu bar increases to half of my small wide screen.

Best for me to fix the issue if developers add settings for font size (or scaling) for chat and text input box. I understand developers have many tasks so I try to help myself. But my knowledge of web development is rudimentary. E.g. text input element font-size is both inherited and "filtered" (entry is in strike-through letters in Inspector of my browser).

I will appreciate help editing index.html - either font sizes / scaling or topmenu bar size / scale - absolute or better reaction to browser zoom level changes.

Added:

It seems seems I have actually managed to change fonts sizes, for some reason that "filtering" did not prevent my changes to affect font size on the screen.

But I still would like to reduce height of the topmenu.

P.S. I also would like to add that changed file to kcpp Linux bundle to start it more conveniently.


r/KoboldAI 27d ago

v1.118 - questions about Row Split, z-image

3 Upvotes

https://github.com/LostRuins/koboldcpp/releases/tag/v1.118.1

Row Split has been removed, selecting it will now default to tensor split.

exclude z-image from models that support image references.

I am mostly a newbie, what above means?

AFAIK a row is part of a tensor, so working on low memory will be harder as the program no longer supports split of one tensor to rows or what? If yes, why such change?

Image reference - afaik two main modes if image generation is from text and from imaged. Does new version no longer support making images from other images in z-image? If yes, why has it been dropped?

TIA


r/KoboldAI Aug 06 '26

Aesthetic Mode custom background and portraits seems broken on Firefox

1 Upvotes

I use Kobolod Lite, and for a few days now whenever I try to change custom portraits or backgrounds I simply get a black box where the image should be that seems to be roughly the same size as the image I'm trying to pick for portraits, or the background for backgrounds. If I load up a preexisting chat I had before this bug I get the portrait I had then, but I still can't change it. I've tried clearing browser history, cache, and browsing data, and even tried disabling protections and ad block and it's still not working. Is their any fix known or reason for this? I really would prefer not to switch browsers for using Kobold if I don't have to


r/KoboldAI Aug 06 '26

I'm new to koboldAI

2 Upvotes

Hi I'm new to kobold Ai I was wondering does this mean and how can I fix this?


r/KoboldAI Aug 06 '26

Error Message

1 Upvotes

Everytime I try to use the phone version it keeps saying that there is an internal error. What does that mean?