openclaw/extensions/document-extract
Patrick Erichsen 5fdb5d942b
feat: show declared plugin capabilities and setup guides (#157956)
* feat(plugins): show declared capabilities in plugin details

* docs(plugins): add concise setup guides for empty detail pages

* chore(workboard): refresh generated manifest formatting

* fix(ui): inline customization settings and prioritize capabilities

* fix(plugins): keep invalid display metadata nonfatal

* fix(ui): retain plugin context in inline alerts

* chore(workboard): refresh generated settings bundle reference

* fix(ui): load customization controls from plugin settings

* chore(workboard): regenerate UI bundle with pinned dependencies

* test(plugins): complete staged overview metadata fixture

* fix(tooling): include plugin capabilities in extracted PR wrapper

* test(gateway): include selected plugin capabilities in discovery fixture

* fix(plugins): keep remote details available with invalid UI metadata
2026-09-25 16:43:53 -07:00
..
assets improve(plugins): give bundled logos consistent white icon tiles (#155259) 2026-09-23 19:09:26 -07:00
document-extractor-worker-entrypoint.ts improve: keep image and PDF processing responsive (#146094) 2026-09-12 10:58:14 -07:00
document-extractor.runtime.test.ts fix(pdf): surface partial document extraction (#131922) 2026-09-23 17:52:04 -07:00
document-extractor.runtime.ts fix(pdf): surface partial document extraction (#131922) 2026-09-23 17:52:04 -07:00
document-extractor.test-support.ts fix(workers): prevent PDF cancellation from aborting Node (#150752) 2026-09-17 03:38:18 -07:00
document-extractor.test.ts fix(pdf): surface partial document extraction (#131922) 2026-09-23 17:52:04 -07:00
document-extractor.ts improve: keep image and PDF processing responsive (#146094) 2026-09-12 10:58:14 -07:00
document-extractor.worker.ts improve(workers): reduce worker startup time and memory (#154293) 2026-09-20 19:45:53 -07:00
index.ts
openclaw.plugin.json feat(plugins): assign one purpose category to every bundled plugin (#142760) 2026-09-10 20:44:20 -07:00
package.json chore(release): close out 2026.9.6 on main (#156869) 2026-09-23 23:02:54 -07:00
README.md feat: show declared plugin capabilities and setup guides (#157956) 2026-09-25 16:43:53 -07:00

Document Extraction

Extract text from PDF attachments locally. When a selected page has too little text, the plugin can render it as an image for a vision-capable model. PDF processing runs in a worker through the bundled PDFium-based extractor.

Get started

The plugin is enabled by default and needs no extraction API key. Configure a model for PDF analysis, then attach a PDF or ask your agent to analyze a local PDF file.

Local extraction and model analysis are separate: the selected model still needs its normal credentials. Models with native PDF support can receive the document directly instead. This extractor handles PDFs, not every document format.

See the PDF guide for model selection, page limits, and encrypted documents.