Skip to main content

Document Data Extraction

The Document Data Extraction Workflow Builder block sends one business document to Lendflow’s extraction service and stores the categorized, extracted result. The current block represents arsen_ai_document_data_extraction.

Requirements

API flow

Use the dedicated document-extraction endpoint, not the general enrichment endpoint:
  1. Send POST /api/applications/{application_id}/document/extract_data.
  2. Submit either multipart file or selected_file_id.
  3. A successful request returns HTTP 201 with an empty body and queues extraction.
  4. Poll Get Commercial Data with services[]=arsen_ai_document_data_extraction.

Data Orchestration availability and flow

The block is available in the Data Extraction group for business entities. Configure it with the path of an attached or uploaded PDF. Each execution processes one file and appends a stored extraction record; it does not merge multiple documents into one response.

What the service returns

Representative response

The extraction schema depends on the uploaded document. This example shows only fields confirmed by the storage flow.

Field meanings

Do not assume a tax-return, bank-statement, or other document schema unless those fields are present in the returned response.

Errors and statuses

FAQ

No. The dedicated extraction request currently validates mimes:pdf.
Yes. Send its valid selected_file_id instead of uploading file.
The provider response varies by document type, and the current Lendflow code does not define one stable public extraction schema. Inspect the returned response instead of assuming fields.