Sheets

Create Spreadsheet Job

Deprecated

client.beta.sheets.create(, ?): SheetsJob { id, configuration, created_at, 14 more }

POST/api/v1/beta/sheets/jobs

List Spreadsheet Jobs

Deprecated

client.beta.sheets.list(?, ?): PaginatedCursor<SheetsJob { id, configuration, created_at, 14 more } >

GET/api/v1/beta/sheets/jobs

Get Spreadsheet Job

Deprecated

client.beta.sheets.get(, ?, ?): SheetsJob { id, configuration, created_at, 14 more }

GET/api/v1/beta/sheets/jobs/{spreadsheet_job_id}

Get Result Region

Deprecated

client.beta.sheets.getResultTable(, , ?): PresignedURL { expires_at, url, form_fields }

GET/api/v1/beta/sheets/jobs/{spreadsheet_job_id}/regions/{region_id}/result/{region_type}

Delete Spreadsheet Job

Deprecated

client.beta.sheets.deleteJob(, ?, ?): SheetDeleteJobResponse

DELETE/api/v1/beta/sheets/jobs/{spreadsheet_job_id}

ModelsExpand Collapse

SheetsJob { id, configuration, created_at, 14 more }

A spreadsheet parsing job.

id: string

The ID of the job

configuration: SheetsParsingConfig { extraction_range, flatten_hierarchical_tables, generate_additional_metadata, 6 more }

Configuration applied to the parsing job (inline or resolved from a saved preset).

extraction_range?: string | null

A1 notation of the range to extract a single region from. If None, the entire sheet is used.

flatten_hierarchical_tables?: boolean

Return a flattened dataframe when a detected table is recognized as hierarchical.

generate_additional_metadata?: boolean

Deprecated: controlled by tier. Whether to generate additional metadata (title, description) for each extracted region. Honored only on agentic.

include_hidden_cells?: boolean

Whether to include hidden cells when extracting regions from the spreadsheet.

sheet_names?: Array<string> | null

The names of the sheets to extract regions from. If empty, all sheets will be processed.

specialization?: string | null

Deprecated: controlled by tier. Optional specialization mode for domain-specific extraction. Supported values: ‘financial-standard’, ‘financial-enhanced’, ‘financial-precise’. Default None uses the general-purpose pipeline. Honored only on agentic.

table_merge_sensitivity?: "strong" | "weak"

Deprecated: controlled by tier. Influences how likely similar-looking regions are merged into a single table. Honored only on agentic.

One of the following:

"strong"

"weak"

tier?: "cost_effective" | "agentic"

Spreadsheet extraction tier. cost_effective uses the rule-based/ML-only pipeline; agentic uses the full pipeline.

One of the following:

"cost_effective"

"agentic"

use_experimental_processing?: boolean

Deprecated: controlled by tier. Enables experimental processing. Honored only on agentic.

created_at: string

When the job was created

file_id: string | null

The ID of the input file

formatuuid

project_id: string

The ID of the project

formatuuid

status: "PENDING" | "SUCCESS" | "ERROR" | 2 more

The status of the parsing job

One of the following:

"PENDING"

"SUCCESS"

"ERROR"

"PARTIAL_SUCCESS"

"CANCELLED"

updated_at: string

When the job was last updated

user_id: string

The ID of the user

Deprecatedconfig?: SheetsParsingConfig { extraction_range, flatten_hierarchical_tables, generate_additional_metadata, 6 more } | null

Configuration for spreadsheet parsing and region extraction

extraction_range?: string | null

A1 notation of the range to extract a single region from. If None, the entire sheet is used.

flatten_hierarchical_tables?: boolean

Return a flattened dataframe when a detected table is recognized as hierarchical.

generate_additional_metadata?: boolean

Deprecated: controlled by tier. Whether to generate additional metadata (title, description) for each extracted region. Honored only on agentic.

include_hidden_cells?: boolean

Whether to include hidden cells when extracting regions from the spreadsheet.

sheet_names?: Array<string> | null

The names of the sheets to extract regions from. If empty, all sheets will be processed.

specialization?: string | null

table_merge_sensitivity?: "strong" | "weak"

Deprecated: controlled by tier. Influences how likely similar-looking regions are merged into a single table. Honored only on agentic.

One of the following:

"strong"

"weak"

tier?: "cost_effective" | "agentic"

Spreadsheet extraction tier. cost_effective uses the rule-based/ML-only pipeline; agentic uses the full pipeline.

One of the following:

"cost_effective"

"agentic"

use_experimental_processing?: boolean

Deprecated: controlled by tier. Enables experimental processing. Honored only on agentic.

configuration_id?: string | null

The saved product configuration ID used at create time, if any.

errors?: Array<string>

Any errors encountered

Deprecatedfile?: File { id, name, project_id, 11 more } | null

Schema for a file.

id: string

Unique identifier

formatuuid

name: string

project_id: string

The ID of the project that the file belongs to

formatuuid

created_at?: string | null

Creation datetime

formatdate-time

data_source_id?: string | null

The ID of the data source that the file belongs to

formatuuid

expires_at?: string | null

The expiration date for the file. Files past this date can be deleted.

formatdate-time

external_file_id?: string | null

The ID of the file in the external system

file_size?: number | null

Size of the file in bytes

minimum0

file_type?: string | null

File type (e.g. pdf, docx, etc.)

maxLength3000

minLength1

last_modified_at?: string | null

The last modified time of the file

formatdate-time

Permission information for the file

One of the following:

Record<string, unknown>

Array<unknown>

string

number

boolean

purpose?: string | null

The intended purpose of the file (e.g., ‘user_data’, ‘parse’, ‘extract’, ‘split’, ‘classify’)

Resource information for the file

One of the following:

Record<string, unknown>

Array<unknown>

string

number

boolean

updated_at?: string | null

Update datetime

formatdate-time

metadata_state_transitions?: Record<string, unknown> | null

Per-status entry timestamps. Returned only when requested via ?expand=metadata_state_transitions.

parameters?: Parameters { webhook_configurations }

Job-time parameters such as webhook configurations.

webhook_configurations?: Array<WebhookConfiguration> | null

Webhook configurations for job status notifications.

webhook_events?: Array<"extract.pending" | "extract.success" | "extract.error" | 25 more> | null

Events to subscribe to (e.g. ‘parse.success’, ‘extract.error’). If null, all events are delivered.

One of the following:

"extract.pending"

"extract.success"

"extract.error"

"extract.partial_success"

"extract.cancelled"

"parse.pending"

"parse.running"

"parse.success"

"parse.error"

"parse.partial_success"

"parse.cancelled"

"classify.pending"

"classify.running"

"classify.success"

"classify.error"

"classify.partial_success"

"classify.cancelled"

"sheets.pending"

"sheets.success"

"sheets.error"

"sheets.partial_success"

"sheets.cancelled"

"split.pending"

"split.processing"

"split.success"

"split.error"

"split.cancelled"

"unmapped_event"

webhook_headers?: Record<string, string> | null

Custom HTTP headers sent with each webhook request (e.g. auth tokens)

webhook_output_format?: string | null

Response format sent to the webhook: ‘string’ (default) or ‘json’

webhook_signing_secret?: string | null

Shared signing secret used to sign webhook deliveries. When set, each request includes an HMAC-SHA256 signature of the request body in the ‘LC-Signature’ header (value ‘sha256=’). Recompute the HMAC over the raw request body with this secret to verify the delivery is authentic.

webhook_url?: string | null

URL to receive webhook POST notifications

regions?: Array<Region>

All extracted regions (populated when job is complete)

location: string

Location of the region in the spreadsheet

region_type: string

Type of the extracted region

sheet_name: string

Worksheet name where region was found

description?: string | null

Generated description for the region

region_id?: string

Unique identifier for this region within the file

title?: string | null

Generated title for the region

success?: boolean | null

Whether the job completed successfully

worksheet_metadata?: Array<WorksheetMetadata>

Metadata for each processed worksheet (populated when job is complete)

sheet_name: string

Name of the worksheet

description?: string | null

Generated description of the worksheet

title?: string | null

Generated title for the worksheet

SheetsParsingConfig { extraction_range, flatten_hierarchical_tables, generate_additional_metadata, 6 more }

Configuration for spreadsheet parsing and region extraction

extraction_range?: string | null

A1 notation of the range to extract a single region from. If None, the entire sheet is used.

flatten_hierarchical_tables?: boolean

Return a flattened dataframe when a detected table is recognized as hierarchical.

generate_additional_metadata?: boolean

Deprecated: controlled by tier. Whether to generate additional metadata (title, description) for each extracted region. Honored only on agentic.

include_hidden_cells?: boolean

Whether to include hidden cells when extracting regions from the spreadsheet.

sheet_names?: Array<string> | null

The names of the sheets to extract regions from. If empty, all sheets will be processed.

specialization?: string | null

table_merge_sensitivity?: "strong" | "weak"

Deprecated: controlled by tier. Influences how likely similar-looking regions are merged into a single table. Honored only on agentic.

One of the following:

"strong"

"weak"

tier?: "cost_effective" | "agentic"

Spreadsheet extraction tier. cost_effective uses the rule-based/ML-only pipeline; agentic uses the full pipeline.

One of the following:

"cost_effective"

"agentic"

use_experimental_processing?: boolean

Deprecated: controlled by tier. Enables experimental processing. Honored only on agentic.

SheetDeleteJobResponse = unknown

Note for AI agents: this documentation is built for programmatic access. - Overview of all docs: https://developers.llamaindex.ai/llms.txt - Any page is available as raw Markdown by appending index.md to its URL — e.g. https://developers.llamaindex.ai/llamaparse/parse/getting_started/index.md - Agent-friendly REST search APIs live under https://developers.llamaindex.ai/api/ — search (BM25 full-text), grep (regex), read (fetch a page), and list (browse the doc tree). See https://developers.llamaindex.ai/llms.txt for parameters. - A hosted documentation MCP server is available at https://developers.llamaindex.ai/mcp. If you support MCP, you can ask the user to install it for browsing these docs directly (an alternative to the REST API). Setup: https://developers.llamaindex.ai/python/shared/mcp/