Hoarder: Self-hostable bookmark-everything app
github.com
github.com
For references there are many options in selfhosted bookmarking apps market. These beside Hoarder are the most known software.
Linkwarden (https://github.com/linkwarden/linkwarden)
Shaarli (https://github.com/shaarli/Shaarli)
LinkAce (https://www.linkace.org/)
Linkding (https://github.com/sissbruecker/linkding)
Wallabag (https://wallabag.org/)
Shiori (https://github.com/go-shiori/shiori)
I've looked into most of these (and instapaper, pocket, etc) and ultimately found Wallabag to be the best. However, their app is quite buggy and site is fairly clunky for my taste. Luckily there's a pretty recent 3rd party client that works offline super well and is on Mac/Linux/android/iOS for free (yay flutter) https://github.com/casimir/frigoligo
Also, I'll note that it's basically a must to use the browser extension with the option to download via what the browser sees if you get content from a lot of sites. That being said the devs are super responsive to reports that sites aren't being scraped appropriately.
My biggest wish is that they supported YouTube (at least titles) and they had a way to indicate when a article needs to be scraped client side.
regarding youtube, my youtube links saved with wallabager browser extension always show the correct title, are you using something else to save them?
Also, I don't love the ads and recommended content in Pocket.
Obviously it archives things for later reading. It works great for that on my e-reader, and I'm super glad it exists and KOReader supports it.
But the API[1] is largely undocumented and undescribed, so I'm kinda at a loss as to what the goals and possibilities are, because the app leaves quite a lot to be desired for e.g. replacing personal cataloguing on something Delicious-like. It seems tailor made for small, barely-customizable offline reading (save it for reading later, maybe with tags, then archive it and don't look at it again), despite the API apparently (maybe?) offering a lot more (but not describing it so it's hard to know how it's intended to be used).
Is there, like... documentation somewhere? Particularly with capabilities / intent / goals? I'm hesitant to sink much time into it without some idea of how it thinks of itself, building a bowl of Hyrum-slaw on top of a shaky foundation is no fun for anyone involved. If it's mostly just what the core app presents, it's probably not what I want.
[1]: https://app.wallabag.it/api/doc/ and https://doc.wallabag.org/developer/api/methods/
Public mode? I'd like people to NOT have to log in.
(Just thinking out loud.. if I were to find all HN threads discussing such apps (maybe searching for raindrop or pocket?) via the search I might find it more quickly.)
It does not have self-hosted, although it does have a free tier. The free tier is pretty limited, allowing just 50 bookmarks. So basically amounts to a free trial more than something you’d use for free for a long time, IMO.
The website is available in English and Chinese, so the developers might be from China.
My wording was a bit confusing in how I said it, but what I meant to say was that like the one you were thinking of this one is also not self-hosted.
For example, I noticed that in the demo access app, there's a note about cooking, and it has 4 tags: - `baking` - `cupcakes` - `oven cooking` - `recipe`
This would get out of hand quickly.
There should be a hierarchy of tags (categories): `cupcakes` in `baking` in `oven cooking` in `recipe`
The only tag needed in this case for the note would be `cupcakes`
I own and operate a "list-taking" app[0] in which every list/kanban-item can itself be a list/kanban.
I currently use it for things I'm the creator of -- tasks, story outlines, etc, but looking to introduce 3rd party content for task management (I want to see GitHub tasks from work next to my own tasks) and, as you say, knowledge management of things like recipes or music.
Items could be part of one or multiple hierarchies. A list of "cake" recipes could be under both "baking" and "party essentials", and music playlists could include other playlists.
As you can tell, this can become convoluted in my mind, and so if that's something that's interesting to you (or anyone reading this), please reach out and let's discuss! hn at nestful.app
"Spontaneous productivity" mirrors some of my own thinking on the subject, especially the JIT and bubbling aspects and how they work together. I haven't seen how it works in the case of Nestful, but I'm keen to try it out. It may adjust the design principles guiding development.
IMO it would be interesting to try to combine the two approaches (curation + auto tagging).
It starts out with the user scaffolding an initial hierarchy, then (after enough usage to provide meaningful data for ML predictions) the ML model predicts on subsequent entries, and asks the user for approval (which feeds a reinforcement learning model)
Hierarchies get out of hand quickly too. You will soon find that different people (or the same person at different times) create different hierarchies for the same thing and that the same thing belongs in multiple places.
I have been experimenting with different representations of data in Neo4j, Markdown, and Orgmode. I even tried cludging the polyhierarchies into different file systems using symlinks and tagging,
I'm still researching for better storage techniques.
I want a good mix between hand editing, but robust machine readable formats. Orgmode works pretty good, but it's fairly complicated to parse, and I think it could be improved.
The retrieval and search part could be improved with RAG, but I don't have the hardware or time at the moment to hacking around with the compute intensive AI stuff.
I guess one small request - could the chrome/firefox extension include a way to transfer the page data from the browser, as it's being displayed to the user? (as in, transfer the entire page/html instead of the page's link). This would likely result in much better support for nasty sites like twitter and such that require credentials, etc..!
As for your request, we're tracking this in (https://github.com/hoarder-app/hoarder/issues/172), which aims to do exactly what you're asking for.
In line with with that, I would expect the login to not be case sensitive when it accepts an email.
Wikipedia has a good summary on what is valid. https://en.wikipedia.org/wiki/Email_address
From the second paragraph:
Although the standard requires the local-part to be case-sensitive,[1] it also urges that receiving hosts deliver messages in a case-independent manner,[2] e.g., that the mail system in the domain example.com treat John.Smith as equivalent to john.smith; some mail systems even treat them as equivalent to johnsmith.[3]
You'll find the footnote links at the Wikipedia article, I'm not going to paste the here. So yes, if my email address is my username, then I would expect it to work the same in uppercase or lowercase. If my username is "like an email address, but not an email address" then you make the rules for your site.
“Data Not Collected — The developer does not collect any data from this app.”
*bookmarks it in huge Trello list where cool bookmarks go to die*
I use Evernote since what, 2005, 2008? Yet I hate every time I start it up. Such ugly bloatware it has become. And the silly “AI powered” features tacked on when it became fashionable… Man, replacing it would feel so good.
It's very good. Some points that can be improved:
- Search inside the description of the bookmark, it doesn't. - Update to a new version of hoarder. Since the software isn't stable, it's a real problem. - Related to the previous point => More archive formats.
Otherwise, it's a very good software. Easy to use, nice front-end, good UX.
Also for searches, Hoarder indexes all the content of the websites it crawls. If it doesn't for you then that would be a bug!
Is it just bookmarks or does it download full pages?
Bookmark applications are generally a failure for long term storage because links always change over time.. so i'm not sure what lense to look at this app through.
Combine them with an automatic-upload-to-archive-dot-org-(if-not-there-already?) function, and save the link to that also? Dunno if any bookmarking app has that already.
[EDIT:] Heh, look – someone is apparently doing precisely that: https://news.ycombinator.com/item?id=42502175 [/EDIT]
I really like hoarder, but kind of surprised how many tools either totally skip over backup functionality or treat it as an afterthought (like this Hoarder issue here: https://github.com/hoarder-app/hoarder/issues/75). Feels like this should be a no-brainer feature, right?
Relatedly (and I think the authors are working on it) anyone using local AI for tags and know good ways to tweak (I'm using Ollama and would love to constrain the the tags a bit?)
It's self-hosted and all packed into SQLite so, IMO, very portable.
Recently added a trick to snapshot all the public links I save - my copy and on Archive.is - link rot is real.
It would host webapps like yours that use in browser sqlite to store data, then the service provide a sync their sqlite data across different devices. The user not the app would pay for the storage of the data, so they would own their data. And you can use CSP to lock down the app from sharing with other domains, meaning an app can't leak your data.
The service would handle identity (only you can access your sqlite data - the app just ) and could provide an app store like experience with different apps of this type.
Sort of like a firebase style backend as a service, but the user would own the data instead of the app.
The concept I'm thinking about is different - it doesn't run any app code on the server, the apps are SPAs that run in the browser only (no backend supporting code), and then the server just syncs the data from the apps. This means the apps can focus on building ux/business logic and not worry about database, how data gets to/from clients, identity, etc. Somewhat like firebase but where the users pays for the server, not the app. That should hopefully be simpler for developers, and theres a lot less likelihood of issues with server config/etc (although presumably pikapod will handle that). It should also be cheaper since you're not constantly running a container, just storing data.
I'm not sure if it's a useful concept or not yet :)
Someone here needs exactly that: https://news.ycombinator.com/item?id=42502576 (And possibly me too.) So yeah, please do continue and then publish it.
I can still distinctly remember reading a site which had Barry Hughart's typewritten notes for his novels (_Bridge of Birds_, _The Story of the Stone_, and _Eight Skilled Gentlemen_), and I considered saving the files for the image scans, but didn't figuring they would always be available.
Since then, he has passed away and the pages in question vanished (no idea on what order that occurred), and I haven't been able to figure out who is in charge of his literary estate, nor where his papers are stored.
I'd give a lot for someone to write a book examining his writings in a scholarly context.
If I think something might be interesting enough to read later, I print it to a PDF --- as a bonus, that means I can send it to my Kindle Scribe to read at my convenience.
We’ve got a shed full of boxes and bags of stuff. Want an easy way to take pics of the contents of a box and the “box number” and be able to browse for the box or specific contents later. Eg a home archive solution.
Anyone know of tools for that?
I think Evernote had something like this when I was using it.
I'm not sure on the support for PDFs with hoarder-app, the github README doesn't seem to mention anything about it.
Thinking in particular of browser bookmark exports and Notion, I'm sure other people would like something else. Being able to parse any kind of text file for links would be great + full-text search in text files / markdown.
It's fascinating even to myself how fast I closed the tab. The annoyance and oversaturation with "ai" is on a level I didn't think was possible
Otherwise, yeah, me too.
For me, the sentiment is mostly, if I'm looking for a tool for something, I want it to do exactly that job, and I want to do it precisely. The presence of AI features gives me a gut feeling both of feature bloat, and that "precision" wasn't a major focus behind the application's design.