From files to connected knowledge.
The pipeline proceeds in resumable phases: inventory and metadata recovery, canonical naming, text extraction, summaries, offline topic classification, graph export, embeddings, cross-book relationships, and an assistant surface.