Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Fujitsu launches made-in-Japan next-generation CPU FUJITSU-MONAKA(global.fujitsu ↗)
    141comments
  2. Hister: A private search engine for the pages you visit and the files you keep(github.com/asciimoo ↗)
    31comments
  3. CrowdSec Source Code Leak(crowdsec.net ↗)
    22comments
  4. Rate limits on GitLab.com are changing(about.gitlab.com ↗)
    79comments
  5. Whoisinspace.com/(whoisinspace.com ↗)
    37comments
  6. Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data(arxiv.org ↗)
    3comments
  7. Vinix – A modern operating system written in V(vinix-os.org ↗)
    32comments
  8. Show HN: Aclif – Agent CLI framework: one grammar, canonical names across SaaS(aclif.ai ↗)
    7comments
  9. Zettascale (YC S24) Is Hiring ASIC/FPGA Engineers to Build Chips for ASI(zscc.ai ↗)
    discuss
  10. Launch HN: Skillsync (YC W26) – AI chat sessions made portable across agents
    12comments
  11. How GLM built its own inference infrastructure(z.ai ↗)
    221comments
  12. Why I didn’t sign the Fields medallists’ letter(gowers.wordpress.com ↗)
    142comments
  13. One Year of Sponsored Servo Development(servo.org ↗)
    125comments
  14. Grand MS-DOS Gaming General MIDI Showdown(johnnovak.net ↗)
    3comments
  15. Ask HN: How to recover Google auth after phone stolen?
    44comments
  16. Show HN: Share your AI Setup, Learn from others(mysetup.ai ↗)
    60comments
  17. The American Religion of Self-Storage Facilities(newyorker.com ↗)
    145comments
  18. CCC invites all model citizens to 40C3(ccc.de ↗)
    108comments
  19. Towards Self-Driving Codebases(detail.dev ↗)
    3comments
  20. The Return of Sail Power: Cargo Ships Are Turning Back to the Wind(gcaptain.com ↗)
    98comments
  21. Mastering Layout Engines in Graphviz: Dot vs. Neato vs. Twopi vs. Circo(visual-paradigm.com ↗)
    5comments
  22. LLM Classification Is Feature Engineering(minimallysufficient.com ↗)
    11comments
  23. Running Ubuntu on the Lenovo IdeaPad Duet(vhaudiquet.fr ↗)
    1comments
  24. Artificial intelligence now beats some of the best human forecasters(economist.com ↗)
    69comments
  25. My temporary PHP fix from 2014 has nearly 20M installs. Today I'm deprecating it(jakeasmith.com ↗)
    82comments
  26. Economic policy for AGI(deepmind.com ↗)
    3comments
  27. Show HN: AutoBot – live voice control for long-running AI work(github.com/demeyer1 ↗)
    2comments
  28. Show HN: I built a new version of my fun spatial 3D online meeting app(flat.social ↗)
    50comments
  29. The Relation Between Mathematics and Physics by Paul Dirac (1939)(cam.ac.uk ↗)
    46comments
  30. Keys Not Included: recovering the signing keys for US driver's license barcodes(ryan.science ↗)
    137comments

Hister: A private search engine for the pages you visit and the files you keep

95 pointsby 1h agogithub.com
31 comments
1h agoHN ↗

Integrate it with linkwarden so it searches the bookmarked pages.

43m agoHN ↗

Thanks, that looks interesting, I will test it out and see if it continuously sync or must be manually imported from time to time, or if it can replaces linkwarden entirely. My linkwarden instance also saves as a pdf not just html and bookmarks are in GB in size, if hister does it more efficiently it’s even better.

41m agoHN ↗

I had the same problem for a very long time but it is largely solved now. I started to simply ask chatgpt "hey I read something about x, y month ago but can't find it now". There is a surprisingly high chance chatbot can just give the exact answer back to me, usually with extra interesting reading materials as a plus.

37m agoHN ↗

Is all your browsing history already with chatgpt or something?

This has an MCP server specifically so a workflow like that would work for you. This is just made to gold the data, and I'm a human accessible way should your AI fail you

34m agoHN ↗

No I don't share anything with chatgpt. But I do have a $20 subscription, if it matters. I feel It's just capable enough to find what I want from my usually vague and inaccurate description.

35m agoHN ↗

And now you are even more dependent on OpenAI...

You didn't solve the problem, you are just trading pain points.

33m agoHN ↗

Nothing about this depends on the provider, you could spin up a local Qwen and give it a search tool like SearXNG or something. At this point local models are more than good enough for simple tasks like that. Using ChatGPT is just (usually) faster and simpler

31m agoHN ↗

There should be no real vendor lock in in my opinion. You can ask the same question with pi + qwen (or any harness + good enough model) with internet access and it will work. Chatgpt is just one option came in handy.

6m agoHN ↗

"Goddammit, I'm wholly dependent on Merriam-Webster, what ever am I going to do".

36m agoHN ↗

Google Chrome did this in 2008. Full-text search over all visited pages, stored offline. It was very useful and I miss it.

Nobody seems to remember it, even though it was a headline feature. Was removed in 2013, I think due to technical constraints.

Will definitely try this.

5m agoHN ↗

I think due to technical constraints.

I think due to shareholders wanting new sportcars. The offline pages don't show Google Ads.

35m agoHN ↗

This is really cool. I like the idea of combining it with a offline Wikipedia cache.

30m agoHN ↗

I've been using this since the last time it came up on here. I don't have it index every page I visit, I use the browser plugin to tell it to index specific ones. It is useful for sure, but I think it will really shine once I've been using it long enough for it to build up a bigger index of things that are old enough that I've actually forgotten about them.

28m agoHN ↗

Ohi, author here! Thanks for posting Hister. Feel free to A.M.A. My first free software search project was Searx, a privacy respecting metasearch engine, but because of the limitations of the metasearch concept, I've decided to take a different approach.

Hister builds a personal search index from pages you visit, bookmarks, browser history, local files, and crawled websites. It stores extracted content with offline result previews, so information remains searchable even when the original page changes or disappears. It supports full text and semantic search, can run entirely on your own machine, and includes a web interface, command line tools, and an MCP endpoint for assistant integrations.

Website: https://hister.org/

Tiny read-only demo: https://demo.hister.org/

Ps.: It looks like our name conflicts with a registered trademark in the US. The owner of the other project has asked us to change it, so we’ll probably need to comply sooner or later.

Name suggestions are welcome! Ideally, the new name should be relatively short, sound good, and have an available .org domain.

Thanks!

15m agoHN ↗

Thanks so much for this, I'm using it all the time. I self host a few things, but I'm using this the most.

12m agoHN ↗

Unfortunately, it is still considered too similar from a legal standpoint.

15m agoHN ↗

Are you aware of ArchiveBox?

https://archivebox.io/

What does Hister do differently? Search seems like a major differentiator, I'm wondering if leveraging the existing archivebox project for archival and implementing good search on top would be more efficient

13m agoHN ↗

Does it work accross multiple computers? Ideally the service runs on a linux box on my tailnet, and my windows and mac systems share the same server.

Edit: I RTFD - and it seems yes.

11m agoHN ↗

Sure, as long as you (and the browser extension) can reach the server, it can be used from as many machines as you want even in a multi-user setup.

20m agoHN ↗

I am happy to see the idea of history search more. I am on my 3rd version of my own. the use of local LLMs has made it easier to support features like weekly summarized and recipe extraction.

I like the search ui. my projects become functional but never polished. https://github.com/sbeckeriv/memoir

17m agoHN ↗

I've been trying this out the past few weeks. I was literally just in there searching for a link 5 minutes ago.

It's badly needed, and so far it's working well for me.

7m agoHN ↗

Your own search engine — Hister is a private search engine for the pages you visit and the files you keep.

Is there any site/project that works as a fully customizable personal front-end to all other SERPs?

When I search for something, I always want a link to the best Wikipedia result. This should always be in the same place and have a giant icon/picture.

Then there could be easily clickable links to the SERP pages for Google, DDG, etc. for that query.

A big link to route it to your favorite LLM.

Seems like you could have a really useful "homepage" for all searches that sat in front of all the other sites. It could be local only and would not require indexing the web. Also wouldn't be a files search thing, as Hister appears to be.

5m agoHN ↗

Kind of related to this in that I built it to hoard knowledge from web pages I've visited along with implementing a Karpathy-style LLM Wiki, but the knowledge is collected automatically from sources I browse.

I have it up on GitHub, but I don't think anyone should use my implementation.

Loosely, what I built:

* On each of my machines I have a cron job running that looks at all my web browser history (usualy it's inspecting the brower's SQLlite across firefox and chrome). If it matches my rule list: hacker news stories, certain reddits, etc. it'll grab the page, convert to markdown and drop in my Obsidian Vault incoming.

* It has a whole de-duping architecture since I might open the same page on multiple machines. Uses the CloudFlare SQLITE D1 storage for tracking the processed links.

* it'll then trigger the LLM to do some Karpathy wiki style taxonomy assignment to the articles, organize them, create an index etc.

It's then available for my "bot" stuff to do writings for me.... I will probably write more about it at some point. I'm not certain it's totally useful and not just a yak-shave on hoarding knowledge.

Ai-drafted article on this [1]

Example AI-Drafted article based on some discussions the other day on Ollma vs LLama.cpp [2]

[1] https://taude.xyz/posts/how-archivore-turns-browsing-into-a-...

[2] https://taude.xyz/posts/skip-ollama-run-llama-cpp-directly-o...