Controls handling of pages that don’t match any category. ‘include’: pages can be grouped as ‘uncategorized’ and included in results. ‘forbid’: all pages must be assigned to a defined category. ‘omit’: pages can be classified as ‘uncategorized’ but are excluded from results.
One of the following:
"forbid"
"include"
"omit"
custom_instructions: Optional[str]
Free-form guidance for where segment boundaries are placed.
maxLength5000
min_pages_per_split: Optional[int]
Minimum pages per segment. Shorter segments are merged into an adjacent segment; 1 disables merging.
minimum1
configuration_id: Optional[str]
Saved configuration ID
transaction_id: Optional[str]
Idempotency key scoped to the project. Reusing a key returns the original job; the new request body is ignored.
Events to subscribe to (e.g. ‘parse.success’, ‘extract.error’). If null, all events are delivered.
One of the following:
"batch.cancelled"
"batch.error"
"batch.pending"
"batch.running"
"batch.success"
"classify.cancelled"
"classify.error"
"classify.partial_success"
"classify.pending"
"classify.running"
"classify.success"
"extract.cancelled"
"extract.error"
"extract.partial_success"
"extract.pending"
"extract.success"
"parse.cancelled"
"parse.error"
"parse.partial_success"
"parse.pending"
"parse.running"
"parse.success"
"sheets.cancelled"
"sheets.error"
"sheets.partial_success"
"sheets.pending"
"sheets.success"
"split.cancelled"
"split.error"
"split.pending"
"split.processing"
"split.success"
"unmapped_event"
webhook_headers: Optional[Dict[str, str]]
Custom HTTP headers sent with each webhook request (e.g. auth tokens)
webhook_output_format: Optional[str]
Response format sent to the webhook: ‘string’ (default) or ‘json’
webhook_signing_secret: Optional[str]
Shared signing secret used to sign webhook deliveries. When set, each request includes an HMAC-SHA256 signature of the request body in the ‘LC-Signature’ header (value ‘sha256=’). Recompute the HMAC over the raw request body with this secret to verify the delivery is authentic.
Controls handling of pages that don’t match any category. ‘include’: pages can be grouped as ‘uncategorized’ and included in results. ‘forbid’: all pages must be assigned to a defined category. ‘omit’: pages can be classified as ‘uncategorized’ but are excluded from results.
One of the following:
"forbid"
"include"
"omit"
custom_instructions: Optional[str]
Free-form guidance for where segment boundaries are placed.
maxLength5000
min_pages_per_split: Optional[int]
Minimum pages per segment. Shorter segments are merged into an adjacent segment; 1 disables merging.
minimum1
transaction_id: Optional[str]
Idempotency key scoped to the project, if one was provided.
updated_at: Optional[datetime]
Update datetime
formatdate-time
Create Split Job
import osfrom llama_cloud import LlamaCloudclient = LlamaCloud( api_key=os.environ.get("LLAMA_CLOUD_API_KEY"), # This is the default and can be omitted)split = client.split.create( file_input="dfl-aaaaaaaa-bbbb-cccc-dddd-eeeeeeeeeeee",)print(split.id)
Note for AI agents: this documentation is built for programmatic access.
- Overview of all docs: https://developers.llamaindex.ai/llms.txt
- Any page is available as raw Markdown by appending index.md to its URL — e.g. https://developers.llamaindex.ai/llamaparse/parse/getting_started/index.md
- Agent-friendly REST search APIs live under https://developers.llamaindex.ai/api/ — search (BM25 full-text), grep (regex), read (fetch a page), and list (browse the doc tree). See https://developers.llamaindex.ai/llms.txt for parameters.
- A hosted documentation MCP server is available at https://developers.llamaindex.ai/mcp. If you support MCP, you can ask the user to install it for browsing these docs directly (an alternative to the REST API). Setup: https://developers.llamaindex.ai/for-agents/mcp/
- Other LlamaIndex tooling for agents — the LlamaParse Platform MCP server, agent skills and plugins, and the n8n node — is mapped at https://developers.llamaindex.ai/for-agents/