Confidence tells you where to look first. A High-confidence value can still be wrong, and a Low-confidence value is often right. Use confidence to decide the order and depth of review, not to skip review entirely on regulated or high-value documents.
The three levels
In the interface, the levels are colour-coded consistently:
- High — green
- Medium — amber
- Low — red, with a pulsing dot to draw attention
Where you see confidence
1
Table view
Each cell shows a small coloured dot next to the value. Scan a column for red dots to find the fields that need attention.
2
Block view
Each field card shows the same dot. Hover it to see the level, and open the field to see ALK’s reasoning for the value.
3
Review modal
Opening a field shows a High / Medium / Low Confidence badge, ALK’s reasoning, the quoted source text, and a highlight over the region of the document the value was read from.
Review status replaces confidence
Once a person has acted on a field, the reviewer’s judgement takes over from ALK’s:- Verified (check icon), Rejected (X icon), and Edited (pencil icon) replace the confidence dot entirely.
- The confidence dot only appears on fields that no one has reviewed yet.
When the source text was hard to read
Separately from the field’s own level, opening a field shows an amber warning when the part of the page a value was read from was itself poorly scanned, asking you to check the value against the document. It usually means a low-resolution scan, a skewed page, a stamp or handwriting over the text, or a fax. The value may still be correct, but the underlying characters were uncertain, so it is worth reading the highlighted region yourself before verifying the field.How the level is calculated
ALK does not take the model’s word for it. The level is a composite of five independent checks. Text recognition quality, format validity, and the model’s own self-assessment carry the most weight; cross-page consistency and source verification refine the result.- Text recognition quality — the recognition score of the specific regions the value came from. A crisp native PDF scores high; a poor scan scores low.
- Format validity — whether the value fits the field type: a real date for a Date field, a parseable number for a Number field, a match for any validation pattern you set, and no signs of garbled text.
- Model self-assessment — how confident the extraction model reported itself to be. Counted, but never on its own.
- Cross-page consistency — whether the value appears on more than one page. Found on two or more pages counts in its favour; found on exactly one page is neutral; conflicting variants elsewhere in the document count against it.
- Source verification — ALK searches the document text for the quote the model claims to have read the value from, independently of what the model reported, and checks that the quote is genuinely there.
The exact weighting and cut-off points are tuned over time as extraction improves, so treat the levels as a review-prioritisation aid rather than a fixed score you can calculate yourself.
Long documents and collections
Long documents are processed in sections, and a collection is processed as a bundle of related documents. In both cases ALK produces several candidate answers for a field and merges them:- If only one section returned a value, that value and its confidence are used as-is.
- If several sections returned the same value, the strongest candidate’s confidence is kept and every supporting source is recorded.
- If sections returned different values, the highest-confidence candidate wins.
- If two candidates disagree at the same confidence level, ALK treats it as a genuine conflict and resolves it explicitly rather than picking arbitrarily.
- If no section returned a value at all, the field is marked Low.
Confidence in Reconciliation
Reconciliation uses a different scale. Exceptions on the Exceptions page carry a numeric match confidence expressed as a percentage rather than a High/Medium/Low level, shown next to each suggested bank-to-General-Ledger pairing along with the reasoning behind it. Percentages of 80% and above are shown in green; anything lower is shown in amber for review. See the Reconciliation Guide.Suggested review policy
A practical starting point for most teams:- Low — review every field, every time. Verify, edit, or reject each one.
- Medium — review every field on documents that carry financial, legal, or compliance consequences; sample the rest.
- High — sample. Review a fixed percentage of rows, and always review fields you know are hard to read in your document set (handwriting, stamps, signatures, totals in tables).
- Re-run extraction on individual fields with Retry before hand-correcting them — a targeted re-run often resolves a Low-confidence field on its own.