Show HN: Never lose a website again
fetching.io
fetching.io
1. major - Breaks back button
2. minor(nitpick) - If I click on the green "notify me" button and then "cancel" on the prompt dialog instead of "OK", it still shows me the green "Success - you will be notified" message
Good luck though!
But, recent nuanced trends in web design/navigation aren't my top skill, so I'm asking the question honestly.
The rest? Maybe meteor is the best choice, cannot really say.
To the other comments, the service is considerably more complex (and yes, built on Meteor, MongoDB and ElasticSearch) as it serves up search results and updates in real time. Though the UI is simple, it does take a fair bit of effort to organize all the browsing information and index such that search results are relevant.
Why not just elasticsearch ?
If you store texts in mongodb, look later for tokumx that compresses data.
(1) Saves an exact copy of the page also.
(2) Indexes the text.
(3) Encrypted (search occurs on your computer).
Been in beta for a while. Thanks for feedback.
Like the Idea of fulltext search on history, but I'm not going to send my full history to some random dude on the internet. No offence intended:)
[1] ”Simple full-text search in your browser” http://lunrjs.com/
I'm curious, would you be more comfortable if all your content were encrypted such that not even the app developer (me) could read it? Or would only a local index do?
My main problem is: As I understand it, the plugin sees what I see and just sends away everything. Including payments and balance in my bank account, business messages in Basecamp, code in private Github repos and content on sites that are not yet public I've signed an nda for.
The benefits don't justify the risks for me. Using incognito mode to avoid a plugin I've installed isn't really feasible.
As a second reason: I have mediocre internet connections most of the time since I'm traveling. Therefore I try to avoid as much requests as possible.
But this is HN, and we wouldn't use even an AES-512-encrypted service if it has to leave our computers. Our password could still be cracked with enough computing power. So, we'd all be extremely happy if you could make a version for paranoid people like us.
Also, I don't know how your indexing works, but please make it easy to backup (just allowing us to define the index folder would do). Possible data loss is the tradeoff with local content.
Unfortunately, logon via Twitter fails with 500 error bar flashing at the site's top and logon via FB fails with "App Not Setup: The developers of this app have not set up this app properly for Facebook Login.", so I can't try it.
Nonetheless, my biggest worry is why this is a service instead of a standalone package. (Actually, I've considered trying with the hope plugin may be possibly FOSS and if so researching whenever it could be hacked to be used with locally-installed Solr/Lucene server) I'm not really comfortable with directly or indirectly sharing my browsing history with most third parties.
I totally hear you on some people not wanting to share browsing history externally. This first version was easiest to do as a hosted service. Next up I intend to package it up as an installable app.
I tried address the privacy angle from day 1. Data is encrypted and only ciphertext is sent to the cloud. Index is stored only on your own machine. Searching occurs on your own machine.
A couple of cool features are that it also saves an exact HTML copy of the page including images and stuff if you read a page "long enough" (currently 90 seconds).
Been in beta for a while. Thanks for feedback.
The goal is to make my computer a search engine that can also recommend articles based on the ones I have visited. It can also check websites to see if there is any new content.
The project is still in its very early stages and any tool I can use will be very helpful. This looks just like what I need.
Basically we should be able to find _any_ piece of information that we have already encountered with anytime in our lifetime. This service is indexing this layer of information.
I would add another layer for information with which we engage more: e.g. liking, linking, sharing etc.
The downside compared to something like this is that it only works if I have the foresight to clip a page. But the upside is that I don't end up indexing all the crap I definitely won't want again, so it's easier to find things in the remainder.
Where's your privacy policy?
It's an interesting idea. Personally I have a script that wgets all the pages I bookmark and I very rarely use that content. What use cases are you anticipating?
Thats a use case I have hit a few times. I even started backing up useful posts just in case they died.
http://gigablast.com/rants.html is an example. Its a really good insight into the creation of a search engine which really should be preserved.
One that did dissapear but has since come back is this post http://widgetsandshit.com/teddziuba/2010/10/taco-bell-progra... which I went looking for a few years ago but had dissapeared from all search indexes.
It would also be important to rank searches well, since otherwise an entire history may have too much. Though this will be a hard area to take on Google...
Part of what I'm trying to validate is exactly the point you raised: will people generally be too freaked out to store their browsing history "in the cloud"?
The next thing I'm hoping to determine is if it'll be enough to encrypt this content is such a way that people are OK with it being stored externally or if only a locally installed version will do. Security aside, there's a great deal of advantage to offer this service hosted.
Thanks for your feedback!
I look forward to seeing what you come up with.
...and by default more and more web sites lose me as a potential client/user/whatev because they require cookies just to display static welcoming information...
...I lose nothing by this, as far as I can tell. Perhaps ignorance really is bliss. :->
True, but in my experience, the major frameworks don't automatically lock out users with cookies disabled. For example, on a Rails app with no before_filter on the homepage, you can start the server and do this:
echo "GET / HTTP/1.1" | nc localhost 3000
You should get back the homepage HTML.But for pages containing information intended to encourage a user to spend time on a site, to get convinced they need to sign-up, it's counter-productive.
Ironic, no?
:->
- who are both anal-retentive enough to want to index the content of every page they visit, and yet...
- are not anal-retentive enough to want to index the content of any page visited on a mobile device, and also...
- don't care about their privacy by sending the address and index of every single page they ever visit to someone.
Good luck.