- Scripts read
- Devanagari and Latin, including pages that mix the two.
- Accepted inputs
- Scanned documents and photographs of documents, including pages captured on a phone.
- How documents arrive
- Direct upload, bulk upload driven from a spreadsheet, connected Google Drive, Dropbox or OneDrive, and indexed email.
- Fields extracted
- Dates, reference numbers, and the parties named in the document, pulled out on arrival.
- Duplicate handling
- Repeat copies of the same document are detected on the way in rather than accumulating.
- Access control
- Extracted text inherits the document’s permissions. A record stays private to whoever added it until it is deliberately shared, and search will not surface it to anyone else in the meantime.
- Tenancy
- Each organisation’s records are stored fully separated from every other organisation’s.
- Not yet available
- Handwriting recognition, a searchable text layer written back onto the original scan, and bulk import of historical registers sorted into categories on the way in.