sourcestring · requiredextract_document_text
Extract searchable text from a PDF or DOCX file. `source` may be a file under /workspace, a /data/telegram path returned by Telegram MCP, or a public HTTP(S) URL. Results are cached privately; when text_truncated is true, continue with get_document_text and next_offset. Native PDF text is supported; scanned PDFs report that OCR is required.
Parameters
max_charsinteger · default 30000include_structureboolean · default false