Overview
A dataset is a curated collection of pipeline execution results — structured fields, extracted text, and metadata — ready for LLM training, RAG, or analytics. Building and exporting datasets does not consume credits.Build a dataset
After your documents are processed (job statuscompleted), build a dataset:
Export formats
Export a dataset
RAG export with quality filtering
Passmin_quality to exclude low-quality chunks before export:
Dataset profile
Get aggregate statistics for a dataset:Chunks API
Available on Pro and Enterprise plans.
Response:
Semantic indexing
Available on Pro and Enterprise plans.