Not Total Recall (1990)

ordellrb@lemmy.world · edit-2 27 days ago

Not Total Recall (1990)

Aux@lemmy.world · 27 days ago

You can start by running sudo apt install tesseract-ocr and then reading its docs.

Morphit @feddit.uk · 27 days ago

Fulfills the AI quota 👍

MacN'Cheezus@lemmy.today · edit-2 26 days ago

It appears to be as simple as tesseract <infile> <outfile>. Possibly could even pipe (or tee) the screenshot straight into that and save both an image and a text file in a single command line.

So something like this should do the trick:

gnome-screenshot -f - | tee /Microsoft/yourPrivacy/$(date +%s).png | tesseract - /Microsoft/yourPrivacy/$(date +%s).txt

Skip the database, just use grep to search that directory if you need to find anything. Voilà, homemade Recall.

Aux@lemmy.world · 23 days ago

It is much better to search using ElasticSearch or Sphinx. Grep is super slow, non indexed and can’t do natural language full text searches. It’s pretty much useless for any real world text search you’d want from OCRed content. And all these better tools are free and open source, so really a no brainer.