Paperless-ngx is an extensive document management system. 3.0.0 has just been released.
It seems from a user/admin perspective that a lot of obsolete stuff under the hood got removed and they did some performance stuff, but it’s also full of AI features. There’s too many changes to read them all though, I wonder if there’s a blogpost or something somewhere summarizing the main changes.
Personally I’ll stay with 2.X for some days at least to see what people say about the new version. I’m not interested to have AI parse my documents, the alrogirthmic auto tagging etc works quite well.



I created an issue to tell the developers they made crap :
https://github.com/paperless-ngx/paperless-ngx/issues/13209
I do understand and share people’s hatred of corporate-owned, centralised cloud AI.
I understand (though share to a lesser degree) people’s ethical concerns about how these models were trained (copyright has been broken for a while now, this just exacerbates it).
But this level of outrage about using local models on an opt-in basis strikes me as hysterical.
Feel free to change my mind.
It reads as especially hysterical in this context, because Paperless is an automatic document categorization system, and I’m really sure what they think the automatic part of that is if it’s not “AI” of some broad description. Paperless, Papermerge et al are basically wrappers for machine learning tools and have been for as long as they’ve existed. LLMs are a natural and obvious fit for the kind of work these applications exist to do.
This just feels like someone reading “Improved AI pathfinding” in the patch notes for a video game and screaming “OH MY GOD IS NOWHERE SAFE?!”
True, although we must concede there is a difference between the simple machine learning already present in paperless and LLMs where, depending on quantization, CPU-based inference doesn’t get you very far.
Absolutely there’s a difference. LLMs, when it comes to this specific task, are better. That’s why they’re being used here. It’s a job they are uniquely well suited to. They do indeed come with high hardware requirements, which is why you’re not forced to use them, and why they provide the option to offload the work to a cloud service.
Personally I would absolutely not want to ever feed my documents into an off-device model, but the point of self-hosted software is that it does what you tell it to and they absolutely should include letting you make bad decisions.
FTFY (although I share your judgment). No disagreement about the rest.
You should be embarrassed by this.
Guess you won’t be making issues there for a while eh
Lol, you suck
The saddest part is that you clearly generated that message using an LLM. Or are you going to tell me that em-dashes are part of your everyday writing?
I have a dedicated button for em dashes on my keyboard /s
… I do, on my custom macro pad.
I just use the combine key, then tap dash a couple times—boom, em dash. On mobile I just hold the dash key, like it’s genuinely pretty easy to insert them.
You are completely ignorant about where do AIs even get their data from, yes?
A free knowledge cookie for you:
The em-dash is the language standard sign in Español for dialogue tags. Literally.