- Test extracted fields using Boolean, logic, numeric, array, string, and other operations.
- If Sensible extracted a field from OCR’d text, test the confidence score for the field’s anchor and value as a measure of the quality of the text images. For example, test that text in a scanned document isn’t blurry or illegible.
- pass a document extraction automatically through your pipeline if there are no errors and 10% of warning validations fail
- flag a document extraction for human review if 5% of error validations fail
Create validations
Sensible app To create validations in the Sensible app:- Click the document type.
- Click Create validation.
- Enter the parameters for the validation.
- Click Create.
Parameters
A validation has the following parameters:Examples
Say that you have a document type for scanned sales quotes, called “sales_quotes”, with configs for- company_A
- company_B
- company_C
Validation 1
- Description: If OCR’d, the source text for quoted rate value is a high-quality, unblurred image.
- Severity: warning
- Condition:
JSON
company_A are scanned documents, check if the field came from OCR’d text. If it was OCR’d (confidence score is not null), then test that it has a high OCR confidence score for both the anchor text and the extracted value text. This validation requires that you set a high verbosity setting in the SenseML configuration.
Validation 2
- Description: The quoted rate value isn’t null
- Severity: error
- Condition:
zip_code is a 5-digit number if the country field equals USA, or 6 alphanumeric characters if the country field equals Canada. Uses a Sensible operation (match) to test regular expressions.
Validations output
For example output of the preceding conditions, see the following extraction excerpt and validation output: Extraction excerptJSON
- Validation 4: Sensible skips the broker email because the prerequisite field
broker.emailis null - Validation 5: fails because
zip_codeis 17 digits
JSON

