How to Check Whether Text Inside a Scanned Diary Attachment Is Searchable
If you want to find an old diary entry by a word that appears inside its scanned attachment, search for a word you know is there, then confirm the matching result points to the attachment rather than the typed note or filename. A scan may be only an image, a PDF with an existing text layer, or an image that the app must process with OCR. Those are different routes to searchable text, and support varies by app, file type, plan, and device.
What does “searchable” mean for an attachment?
A successful app search can match several different things: the typed body of the diary note, the attachment’s filename, selectable text already embedded in a PDF, or words recognized from an image through optical character recognition (OCR). A result alone does not show which route produced the match. For example, searching “Station” proves little if the note itself says “Station ticket” or the file is named `Station.pdf`.
A scanned PDF can look like a normal document while containing only page images. If you cannot select a printed word in a PDF viewer, that is a clue that the PDF may lack a text layer; it is not proof about what your note app can OCR. Conversely, selectable PDF text may be searchable without image recognition. Check the file in a PDF viewer and test the app separately.
Make a controlled test with a known word
Use a copy of the attachment or a disposable note if possible. Pick a distinctive word clearly visible inside the scan, and make sure that word does not occur in the note title, typed body, or filename. Then:
This is a procedure, not a report of testing any particular account. A failed query is not conclusive by itself: OCR may be disabled, processing may still be underway, a format may be unsupported, or the current device may not have the relevant index.
Check the format and text layer first
Record whether the attachment is JPEG/PNG, a scanned PDF, or a PDF with selectable text. Then check the app’s current support documentation for each format separately. Joplin’s [OCR documentation](https://joplinapp.org/help/apps/ocr/) says its OCR scans PNG and JPEG images and PDF files, and that image and PDF processing is available in the desktop app. It also describes an option to view OCR text for a PDF link or image. This makes it possible to check both the extracted text and whether the expected word appears in it.
Microsoft’s [OneNote OCR instructions](https://support.microsoft.com/en-us/office/copy-text-from-pictures-and-file-printouts-using-ocr-in-onenote-93a70a2f-ebcd-42dc-9f0b-19b09fd775b4) describe copying text from pictures and file printouts. For a multi-page printout, the instructions offer copying from one page or all pages and warn that some results can take 24–48 hours. That describes recognition and text extraction from OneNote pictures or printouts; do not assume every kind of attached file is processed the same way.
Separate app support from account and device conditions
An app may support OCR in general while your account, plan, settings, or current device affects whether it works for a particular attachment. Evernote’s [search feature page](https://evernote.com/features/search) currently says its Personal, Professional, and Enterprise plans can search text in PDFs, Office documents, images, presentations, and scanned documents. Check the updated plan details and your own plan before relying on that capability.
Joplin documents another important distinction: OCR scanning runs in its desktop app, while mobile apps can access OCR data through sync. So a missing match on a phone does not establish that the file cannot be indexed; first check desktop processing and whether synced data is available. More generally, verify the app version, applicable plan, OCR setting, file type, processing status, and sync state against current documentation. Do not assume that uploading a file automatically extracts its text or that every device builds an index locally.
Read the result carefully and diagnose misses
If the known word appears in the app’s OCR text but search does not return the note, the problem is likely further along the search or sync path; check the app’s search documentation and repeat after processing or sync completes. If OCR text omits the word, inspect scan clarity, orientation, language settings, and whether the writing is printed or handwritten. Joplin says its current system is reliable for printed text and clear images, but does not support handwriting; Microsoft also cautions that OCR effectiveness depends on image quality. Recognition errors are therefore a plausible cause of a miss.
If the word is selectable in the PDF but not found through the note app, verify that the app indexes PDFs with embedded text as well as image-only scans. If only the filename or typed note triggers results, that confirms those fields are searchable, but says nothing about words in the attachment. Keep these checks separate when recording the outcome.
Keep a short capability record
For each app you evaluate, note the app and device, account plan if applicable, attachment format, whether the PDF word is selectable, OCR setting and processing status, sync state, and exact test outcome. Record whether the matching text came from the note, filename, existing PDF text layer, or OCR text. This turns “search seems to work” into a check you can repeat after changing device, plan, settings, or file format.
