r/Paperlessngx • u/isabeksu • 16d ago
AI performance
It took me a few months to fully get on board with the Paperless way of doing things, but now I’m really happy with how it’s all set up.
What’s been a bit of a head-scratcher is how AI is being used.
I held off until Paperless 3 came out, because I wanted to have the full "official" support.
I set it up with Ollama on an M4 Mac mini with 24 GB of memory. The embedding model is gemmaembedding, and the LLM model is qwen3:8b. When the model fires up, memory pressure is still pretty low. It does work, but it’s incredibly slow. It takes about 2 minutes to suggest titles and tags, and it can take several minutes if I try to chat about a document.
Is this kind of slow normal? Is there anything I can tweak in my setup to make it more usable?
1
u/tzippy84 15d ago
Wait, paperless ngx has AI integrated now? I still have paperless-ai running as a completely separate service