wardethan2000-eng makes scanned evidence searchable
This fork adds the document-processing layer needed to turn scans, photos and office files into material the assistant can actually use.
For legal teams, the important change is not just broader upload support. It is that text trapped inside a scanned bundle can be recognised after upload, then used in previews, citations and assistant responses while retaining page references and clearly identifying OCR-derived text.
- Scanned PDFs and photographs can be converted into searchable text.
- Plain-text and several common office and image formats can enter the same document workflow.
- PDF renditions support more reliable viewing and page-level reference.
The trade-off is operational: processing happens after upload and may be slow or memory-hungry, so a deployment needs enough capacity and the required document-conversion tools in place.
Spotted something wrong? Or know the PR text has fresher detail than the writeup above?