‹ BackHN Continuity

Thread

Hister: A private search engine for the pages you visit and the files you keep

741 points · 202 comments · bookofjoe

  1. taude · · focus · HN ↗
    Kind of related to this in that I built it to hoard knowledge from web pages I've visited along with implementing a Karpathy-style LLM Wiki, but the knowledge is collected automatically from sources I browse.

    I have it up on GitHub, but I don't think anyone should use my implementation.

    Loosely, what I built:

    * On each of my machines I have a cron job running that looks at all my web browser history (usualy it's inspecting the brower's SQLlite across firefox and chrome). If it matches my rule list: hacker news stories, certain reddits, etc. it'll grab the page, convert to markdown and drop in my Obsidian Vault incoming.

    * It has a whole de-duping architecture since I might open the same page on multiple machines. Uses the CloudFlare SQLITE D1 storage for tracking the processed links.

    * it'll then trigger the LLM to do some Karpathy wiki style taxonomy assignment to the articles, organize them, create an index etc.

    It's then available for my "bot" stuff to do writings for me.... I will probably write more about it at some point. I'm not certain it's totally useful and not just a yak-shave on hoarding knowledge.

    Ai-drafted article on this [1]

    Example AI-Drafted article based on some discussions the other day on Ollma vs LLama.cpp [2]

    [1] <a href="https:&#x2F;&#x2F;taude.xyz&#x2F;posts&#x2F;how-archivore-turns-browsing-into-a-wiki&#x2F;" rel="nofollow">https:&#x2F;&#x2F;taude.xyz&#x2F;posts&#x2F;how-archivore-turns-browsing-into-a-...

    [2] <a href="https:&#x2F;&#x2F;taude.xyz&#x2F;posts&#x2F;skip-ollama-run-llama-cpp-directly-on-a-mac&#x2F;" rel="nofollow">https:&#x2F;&#x2F;taude.xyz&#x2F;posts&#x2F;skip-ollama-run-llama-cpp-directly-o...

    1. skinfaxi · · focus · HN ↗
      Why don&#x27;t you think people should use your implementation? Just curious
      1. taude · · focus · HN ↗
        I vibed most of it (and though I like the code that was output, I coached it with some custom skills). If I were to share it out, I&#x27;d probably clean up some of the tools to make their interfaces smaller and cleaner. And make the docs a lot simpler. I&#x27;d probably also create some form of dashboard so you can see what&#x27;s happening (all the docs scraped and saved), etc.

        EDIT: it&#x27;s also the type of thing that feels very personally customized for my needs. I encourage you to build something similar on the idea. Much like how Karpathy Wiki was suggestive and not a runtime to just use...

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.