Remix.run Logo
asciimoo 7 hours ago

Ohi, author here! Thanks for posting Hister. Feel free to A.M.A. My first free software search project was Searx, a privacy respecting metasearch engine, but because of the limitations of the metasearch concept, I've decided to take a different approach.

Hister builds a personal search index from pages you visit, bookmarks, browser history, local files, and crawled websites. It stores extracted content with offline result previews, so information remains searchable even when the original page changes or disappears. It supports full text and semantic search, can run entirely on your own machine, and includes a web interface, command line tools, and an MCP endpoint for assistant integrations.

Website: https://hister.org/

Tiny read-only demo: https://demo.hister.org/

Ps.: It looks like our name conflicts with a registered trademark in the US. The owner of the other project has asked us to change it, so we’ll probably need to comply sooner or later.

Name suggestions are welcome! Ideally, the new name should be relatively short, sound good, and have an available .org domain.

Thanks!

usernomdeguerre an hour ago | parent | next [-]

Been using hister for a number of weeks now; i'm coming across sites whose content would be better handled with a custom extractor; but it looks like extractors need to be bundled into the build in order to work? Is that correct?

Put another way, i can't write an extractor for Reuters and then point a config to it from my current hister binary?

elektor 5 hours ago | parent | prev | next [-]

Hi asciimoo, seems like this is the second time Hister is hitting the HN front page in a month, so congrats on the success!

Question for you: For the less tech savvy of us on here, is there any chance Hister can be can hosted on something like Pikapods? https://www.pikapods.com/

asciimoo 5 hours ago | parent [-]

Yes, that's something I'd like to support. The main missing piece for a user-friendly hosting option such as PikaPods is a configuration UI. At the moment, customizing Hister requires editing a configuration file, which isn't practical for this kind of hosted service.

elektor 5 hours ago | parent [-]

Appreciate the response! I'll be eagerly following Hister's progress. For now, I've settled on a mix of Instapaper and using SingleFile uploads to Dropbox.

cyanlimetea 44 minutes ago | parent [-]

[dead]

culi 3 hours ago | parent | prev | next [-]

The thing I most struggle with in this domain is recalling information from videos. I watch/listen to a lot of hour+ lectures and I rely on this website deeply:

https://filmot.com/

It lets you search YouTube transcripts. If you could somehow integrate video transcripts into this tool, I would be extremely interested in trying it out

mxuribe 5 hours ago | parent | prev | next [-]

Hi @asciimoo , related to a name suggestion, how about something like...

* chronilog.org ...as in, a log of one's chronicles.

* And if you will include this into KDE, then can use a 'k' instead, such as kronilog.org :-)

Both seem to be available. ;-)

asciimoo 5 hours ago | parent [-]

This is a great suggestion, thanks! I'll definitely add it to the list of candidates. My plan is to do a vote on our social platforms if we have a few decent candidates.

consumer451 4 hours ago | parent [-]

Genuinely curious, how could one fight pre-emptive domain squatters once any candidate is publicly suggested?

When I have suggested names in other situations like this in the past, I spent the ~$10 to get the domain, and offered the transfer the free. Of course, not everyone would do this.

mircea 5 hours ago | parent | prev | next [-]

On HN does it capture both the HN post page and the target page?

E.g.: for this submission I would want both https://news.ycombinator.com/item?id=49743097 and https://github.com/asciimoo/hister captured.

asciimoo 5 hours ago | parent [-]

The extension captures the content of the opened tabs, it does not create new requests. If you open both, it captures both.

iririririr 4 hours ago | parent [-]

You may be the first "search engine" capable of indexing instagram and other closed sites.

asciimoo 4 hours ago | parent [-]

Exactly, this is the biggest advantage of the extension. It is fully invisible for the websites, so no captcha, anti-bot protection, no authentication issues, every common bottleneck of a classic crawler is solved by the browser/user.

asdfqwertzxcv 6 hours ago | parent | prev | next [-]

Thanks so much for creating this. Installed last time it was posted and have been loving it. The MCP server and extensions and userscripts are great QOL additions, as well. Always wondered if something was out there like this and you answered my prayers! New name suggestion: MisterHistory

jamienk 6 hours ago | parent | prev | next [-]

Would like: * Local web page interface or even browser UI element (since extension needed anyway) * Ability to add notes to history * Flag if bookmarked, allow filtering "bookmarks only" * Keep old versions of pages * Human-readable text diff vs current live page

Beijinger 6 hours ago | parent | prev | next [-]

"No mandatory cloud - A complete personal setup can run on one local machine."

How does it sync via several computers?

al_hag 5 hours ago | parent [-]

Tailscale is one option

5 hours ago | parent | prev | next [-]
[deleted]
dwedge 7 hours ago | parent | prev | next [-]

It feels like I'm the only person using this but I'd like to throw another potential bookmark manager integration into the ring, cherry https://github.com/haishanh/cherry

pava0 6 hours ago | parent [-]

Not even a readme?

dwedge 5 hours ago | parent [-]

Never even noticed that was missing and it's probably why nobody else uses it. I took it from https://www.reddit.com/r/selfhosted/comments/xyepiu/cherry_a... and https://cherry.haishan.me/ and just worked from the Dockerfile

Anonymous106 6 hours ago | parent | prev | next [-]

Histerekishi / Histereki - れきし/歴史 means history in Japanese. reki れき/歴 is a suffix which means (history of)

Histeri

MyHister(i)

Hyster(y)

Also, I have been using your app for two months now. I have only had to rely on it a few times, but each time I did it worked beautifully. Thank you.

dbliss 7 hours ago | parent | prev | next [-]

Does it work accross multiple computers? Ideally the service runs on a linux box on my tailnet, and my windows and mac systems share the same server.

Edit: I RTFD - and it seems yes.

asciimoo 7 hours ago | parent [-]

Sure, as long as you (and the browser extension) can reach the server, it can be used from as many machines as you want even in a multi-user setup.

nottorp 6 hours ago | parent [-]

"Optional global or personal access token used to authenticate extension requests."

Looks like you can even set authentication up so you can run it at home but connect while you're away too...

corndoge 7 hours ago | parent | prev | next [-]

Are you aware of ArchiveBox?

https://archivebox.io/

What does Hister do differently? Search seems like a major differentiator, I'm wondering if leveraging the existing archivebox project for archival and implementing good search on top would be more efficient

asciimoo 6 hours ago | parent [-]

The main difference I see is Hister focuses on creating an active knowledge base and finding information quickly, while ArchiveBox focuses on preserving web content for the long term.

hefner1456 3 hours ago | parent | prev | next [-]

May I suggest Yahoox!

Cider9986 6 hours ago | parent | prev | next [-]

What about "searchy.me"?

swyx 6 hours ago | parent | prev | next [-]

is this like a pihole? is there a design difference you are going for here?

danielrmay 6 hours ago | parent | prev | next [-]

Thanks for making Hister, I've been using it for a few days (~7k docs) and I'm impressed so far.

smellf 3 hours ago | parent | prev | next [-]

Srchr

Zizizizz 5 hours ago | parent | prev | next [-]

hisect (history and bisect)

Seekfold (seek and manifold)

Seekdex (seek and index)

fooqux 4 hours ago | parent | prev | next [-]

Histeria

6510 5 hours ago | parent | prev | next [-]

I gather rss feeds from the websites I visit and it's hard to express how interesting they are. The gut says it borderlines some random collection but that couldn't be more wrong. I also enjoyed YaCy, that project should have a good amount of ideas for you. I kinda end up assigning more and more bandwidth until it gets in the way and I forget to enable it again. The turtle button on some torrent clients is a good invention.

Beijinger 6 hours ago | parent | prev | next [-]

"It looks like our name conflicts with a registered trademark in the US. "

So? Where are you based? For what class was the trademark filed? When was it filed?

I doubt that he has any leverage, but I don't know the background.

Capricorn2481 7 hours ago | parent | prev | next [-]

Thanks so much for this, I'm using it all the time. I self host a few things, but I'm using this the most.

yeahdef 6 hours ago | parent | prev | next [-]

"hister 2: histlectric histerloo"

mxuribe 5 hours ago | parent [-]

OMG, you win!!! :-D

adfm 7 hours ago | parent | prev | next [-]

histro.org is available.

asciimoo 7 hours ago | parent [-]

Unfortunately, it is still considered too similar from a legal standpoint.

adfm 7 hours ago | parent [-]

In that case, clipshot.org is also available.

aaron695 2 hours ago | parent | prev | next [-]

[dead]

cyanlimetea 44 minutes ago | parent | prev [-]

[dead]