Skip to main content
Formatters help you clean up and standardize extracted values — so your data is consistent, structured, and ready for use in downstream systems. They run after extraction, allowing your model to focus on finding the right values while the formatter ensures they’re in the right shape.

Available Formatters

  • Numeric
    Turns numbers written any way into a plain number, dropping currency symbols, spaces and thousand separators.
  • Date
    Converts dates into ISO 8601 (YYYY-MM-DD), including dates written in other languages.
  • RegExp
    Rewrites a value with a regular expression. Use it to strip prefixes, remove separators, or keep only part of a value.
  • Classification
    Restricts a field to a set of categories, such as document type or a signed/not signed state.
  • Match & Map
    Matches the value against a list of known values and optionally replaces it with a mapped value. The list can be edited in a table, imported from CSV, or fetched from your own URL.
Some formatters are added for you rather than picked from the list. When your agent is built, fields that hold a currency, a country or a gender get a Currency, ISO Alpha-2 country code or Gender formatter, which enforces the valid values for that field and falls back to Unknown.

How formatters run

A field can have several formatters. They run in the order they are listed, each one receiving the value the previous one produced — so a RegExp formatter that strips a currency symbol can be followed by a Numeric formatter that parses what is left. If a formatter cannot produce a value — an unparseable number, a date it does not recognize, a value missing from a classification — it clears the field and reports an error, and the remaining formatters are skipped. The document then goes to human review instead of continuing with a wrong value. A field with no formatters keeps the value exactly as it was extracted.