Using Opus 5.5 to discover a new eyewitness record of the dodo
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Using Opus 5.5 to discover a new eyewitness record of the dodo
Unofficial Hacker News client; not affiliated with Y Combinator.
jamienk · · focus · HN ↗
komali2 · · focus · HN ↗
I would say it's about 80% accurate, which means it's missing enough key words to make a lot of it uselessly unintelligible. I can easily compare the images against text I turn up in a grep which is nice if I'm looking for something.
Allegedly Claude set up a system for retraining for my handwriting, but it would require me to manually revise several hundred pages by hand so I don't think I'll ever do it.
<a href="https://github.com/508-dev/journal-ocr" rel="nofollow">https://github.com/508-dev/journal-ocr
jiggawatts · · focus · HN ↗
GPT 6.1 and Gemini Flash 3.8 both do pretty well, their OCR of your sample image is only "wrong" in the sense that the original has typos and they corrected some inadvertently and/or filled in gaps where you had "unintelligible" in the canonical text.
If you have the budget and want the best possible results, you need to run each image through multiple models and then combine the outputs into a final "merge these" prompt. Better scanning helps too, your sample image is rotated and you used a phone in low light. Try a DSLR or a flatbed scanner and process only one page at a time instead of two at once.
jamienk · · focus · HN ↗
komali2 · · focus · HN ↗