Search Agentik

CtrlK

Glossary · Retrieval & memory

Document parsing

Document parsing is extracting text, tables, headings and layout from files such as PDFs, slides or web pages into clean, structured content that can be chunked and indexed for retrieval.

Document parsing sits in the Retrieval & memory part of the Agentik {OS} glossary, which defines the words used to build and run AI agent systems.

Also called document extraction.