Extract
Create Extract Job
List Extract Jobs
Get Extract Job
Delete Extract Job
Validate Extraction Schema
Generate Extraction Schema
ModelsExpand Collapse
class ExtractConfiguration:
Extract configuration combining parse and extract settings.
string? ParseConfigID
Saved parse configuration ID to control how the document is parsed before extraction
string? ParseTier
Parse tier to use before extraction. Defaults to the extract tier if not specified.
string? TargetPages
Comma-separated page numbers or ranges to process (1-based). Omit to process all pages.
string Version
Use ‘latest’ for the latest release for the selected tier or a date string (YYYY-MM-DD format) to pin to the nearest release at or before that date. Job responses always report the concrete resolved version the job runs, fixed at job creation; saved configurations keep the value as provided.
class ExtractJobMetadata:
Extraction metadata.
ExtractedFieldMetadata? FieldMetadata
Metadata for extracted fields including document, page, and row level info.
IReadOnlyDictionary<string, DocumentMetadata?>? DocumentMetadata
Per-field metadata keyed by field name from your schema. Scalar fields (e.g. vendor) map to a FieldMetadataEntry with citation and confidence. Array fields (e.g. items) map to a list where each element contains per-sub-field FieldMetadataEntry objects, indexed by array position. Nested objects contain sub-field entries recursively.
class ExtractV2Job:
An extraction job.
required string Status
Current job status.
PENDING— queued, not yet startedRUNNING— actively processingCOMPLETED— finished successfullyFAILED— terminated with an errorCANCELLED— cancelled by user
ExtractConfiguration? Configuration
Extract configuration combining parse and extract settings.
string? ParseConfigID
Saved parse configuration ID to control how the document is parsed before extraction
string? ParseTier
Parse tier to use before extraction. Defaults to the extract tier if not specified.
string? TargetPages
Comma-separated page numbers or ranges to process (1-based). Omit to process all pages.
string Version
Use ‘latest’ for the latest release for the selected tier or a date string (YYYY-MM-DD format) to pin to the nearest release at or before that date. Job responses always report the concrete resolved version the job runs, fixed at job creation; saved configurations keep the value as provided.
ExtractJobMetadata? ExtractMetadata
Extraction metadata.
ExtractedFieldMetadata? FieldMetadata
Metadata for extracted fields including document, page, and row level info.
IReadOnlyDictionary<string, DocumentMetadata?>? DocumentMetadata
Per-field metadata keyed by field name from your schema. Scalar fields (e.g. vendor) map to a FieldMetadataEntry with citation and confidence. Array fields (e.g. items) map to a list where each element contains per-sub-field FieldMetadataEntry objects, indexed by array position. Nested objects contain sub-field entries recursively.
ExtractResult? ExtractResult
Metadata? Metadata
Job-level metadata.
ExtractJobUsage? Usage
class ExtractV2JobCreate:
Request to create an extraction job. Provide configuration_id or inline configuration.
ExtractConfiguration? Configuration
Extract configuration combining parse and extract settings.
string? ParseConfigID
Saved parse configuration ID to control how the document is parsed before extraction
string? ParseTier
Parse tier to use before extraction. Defaults to the extract tier if not specified.
string? TargetPages
Comma-separated page numbers or ranges to process (1-based). Omit to process all pages.
string Version
Use ‘latest’ for the latest release for the selected tier or a date string (YYYY-MM-DD format) to pin to the nearest release at or before that date. Job responses always report the concrete resolved version the job runs, fixed at job creation; saved configurations keep the value as provided.
IReadOnlyList<WebhookConfiguration>? WebhookConfigurations
Outbound webhook endpoints to notify on job status changes
IReadOnlyList<WebhookEvent>? WebhookEvents
Events to subscribe to (e.g. ‘parse.success’, ‘extract.error’). If null, all events are delivered.
IReadOnlyDictionary<string, string>? WebhookHeaders
Custom HTTP headers sent with each webhook request (e.g. auth tokens)
string? WebhookSigningSecret
Shared signing secret used to sign webhook deliveries. When set, each request includes an HMAC-SHA256 signature of the request body in the ‘LC-Signature’ header (value ‘sha256=
class ExtractV2JobQueryResponse:
Paginated list of extraction jobs.
The list of items.
required string Status
Current job status.
PENDING— queued, not yet startedRUNNING— actively processingCOMPLETED— finished successfullyFAILED— terminated with an errorCANCELLED— cancelled by user
ExtractConfiguration? Configuration
Extract configuration combining parse and extract settings.
string? ParseConfigID
Saved parse configuration ID to control how the document is parsed before extraction
string? ParseTier
Parse tier to use before extraction. Defaults to the extract tier if not specified.
string? TargetPages
Comma-separated page numbers or ranges to process (1-based). Omit to process all pages.
string Version
Use ‘latest’ for the latest release for the selected tier or a date string (YYYY-MM-DD format) to pin to the nearest release at or before that date. Job responses always report the concrete resolved version the job runs, fixed at job creation; saved configurations keep the value as provided.
ExtractJobMetadata? ExtractMetadata
Extraction metadata.
ExtractedFieldMetadata? FieldMetadata
Metadata for extracted fields including document, page, and row level info.
IReadOnlyDictionary<string, DocumentMetadata?>? DocumentMetadata
Per-field metadata keyed by field name from your schema. Scalar fields (e.g. vendor) map to a FieldMetadataEntry with citation and confidence. Array fields (e.g. items) map to a list where each element contains per-sub-field FieldMetadataEntry objects, indexed by array position. Nested objects contain sub-field entries recursively.
ExtractResult? ExtractResult
Metadata? Metadata
Job-level metadata.
ExtractJobUsage? Usage
class ExtractedFieldMetadata:
Metadata for extracted fields including document, page, and row level info.
IReadOnlyDictionary<string, DocumentMetadata?>? DocumentMetadata
Per-field metadata keyed by field name from your schema. Scalar fields (e.g. vendor) map to a FieldMetadataEntry with citation and confidence. Array fields (e.g. items) map to a list where each element contains per-sub-field FieldMetadataEntry objects, indexed by array position. Nested objects contain sub-field entries recursively.