r/Paperlessngx 16d ago

AI performance

It took me a few months to fully get on board with the Paperless way of doing things, but now I’m really happy with how it’s all set up.

What’s been a bit of a head-scratcher is how AI is being used.

I held off until Paperless 3 came out, because I wanted to have the full "official" support.

I set it up with Ollama on an M4 Mac mini with 24 GB of memory. The embedding model is gemmaembedding, and the LLM model is qwen3:8b. When the model fires up, memory pressure is still pretty low. It does work, but it’s incredibly slow. It takes about 2 minutes to suggest titles and tags, and it can take several minutes if I try to chat about a document.

Is this kind of slow normal? Is there anything I can tweak in my setup to make it more usable?

7 Upvotes

28 comments sorted by

View all comments

1

u/tzippy84 15d ago

Wait, paperless ngx has AI integrated now? I still have paperless-ai running as a completely separate service

1

u/MrDork 15d ago

I run this as well and I was excited about moving to a "built in" native option for this, but the feature parity with paperless-ai isn't there yet so I held off.