Full-text OCR and Aadhaar masking.
Available as part of our on-premise deployment, not as a cloud API. If you need searchable-PDF OCR at volume, or UIDAI-compliant Aadhaar redaction inside your own environment, our team will walk you through the options.
Two capabilities, deployed where your documents live
Searchable text, JSON and searchable PDFs
- ✓ PDFs, TIFFs, scans and photos in, full text out
- ✓ Searchable PDF output, text overlaid on the original image
- ✓ 9 languages including Indian and African scripts
- ✓ Auto-deskew, orientation detection, bounding boxes
- ✓ Built for bulk digitization and archive search
UIDAI-compliant redaction in your environment
- ✓ Detects 12-digit Aadhaar patterns, PAN and other sensitive fields
- ✓ Redacts Aadhaar QR code and barcode regions too
- ✓ Layout-preserving output, audit-ready redaction counts
- ✓ Works on scanned cards, eKYC PDFs and KYC bundles
- ✓ Documents never leave your infrastructure
For the workloads that can't leave the building
Plain OCR at archive scale and Aadhaar redaction over KYC files are exactly the workloads most teams want inside their own perimeter — regulated data, high volume, and retention policies that a cloud round-trip complicates. So that's where we deliver them: deployed on your hardware, integrated with your pipeline, sized for your volume.
Full-text OCR and Aadhaar masking
Available as part of our on-premise deployment, not as a cloud API. If you need searchable-PDF OCR at volume, or UIDAI-compliant Aadhaar redaction inside your own environment, our team will walk you through the options. Talk to sales →