‹ BackHN Continuity

Thread

Show HN: AI search for every photo and every frame of video on macOS

153 points · 69 comments · allenleee

  1. postalcoder · · focus · HN ↗
    Since this is for the mac you really should be using apple's vision framework for OCR. It smokes tesseract in both speed and accuracy.

    Edit: I'm curious which LLM was used to generate the code. I fed the title of your post to claude/deepseek/qwen/codex asking to recommend a stack for this project, expecting to frown thinking that they still recommend tesseract. However, I found that they all recommend apple's vision framework. In fact the latest model to recommend Tesseract is gpt-4.1.

    1. vavkamil · · focus · HN ↗
      I recently tried to recover text from 8 frames of an office-shot YouTube video, where only a small, blurry portion of a computer screen was visible. After spending half a day with Astra on it, the conclusion was that it’s not possible to read.

      Later that evening, I just paused the YouTube video on my phone, circled the part of the display with Google Lens, and it read the whole thing with pretty good accuracy. It was mindblowing :)

      1. sheept · · focus · HN ↗
        All those years of captchas must've made Google's model bulletproof
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.