Input
ExecutionExecution trigger to start AI-powered document extraction.
AI/Processing
Extracts text and content from documents using AI for enhanced image descriptions and OCR.
Scores range from 0 to 10. Higher values mean more impact, exposure, or operational weight.
Execution trigger to start AI-powered document extraction.
Document file to extract (PDF, DOCX, XLSX, images, etc.).
Vision-capable AI model for image analysis and OCR.
Whether to extract and embed images from the document.
Number of images to batch per LLM request (higher = faster but may hit token limits).
Number of PDF pages to process in parallel (higher = faster but uses more memory).
LLM temperature (0.0 = deterministic, 1.0 = creative). Lower is better for extraction.
Maximum output tokens per LLM call. Leave at 0 for model default. Set lower for unreliable models.
Prompt contract of the selected model. Document-parsing models only answer to their own trained prompt: • Default — a tuned general prompt for vision models (GPT-4o, Claude, Gemini, Qwen-VL) • Unlimited-OCR — baidu/Unlimited-OCR, self-hosted via vLLM • DeepSeek-OCR — deepseek-ai/DeepSeek-OCR and -OCR-2 • olmOCR — allenai/olmOCR-2, emits YAML front matter • Nanonets-OCR — nanonets/Nanonets-OCR-s and -OCR2 • dots.ocr — plain text extraction; use Page Prompt for its JSON layout mode • Granite-Docling — IBM Granite-Docling and SmolDocling, emits DocTags • PaddleOCR-VL — PaddlePaddle/PaddleOCR-VL Every preset except Default forces full-page OCR and one image per request. Recommended temperature is 0.0 for all of them except olmOCR (0.1).
Prompt for converting a rendered document page to text. Switching the preset rewrites this unless you have typed your own. Leave empty to fall back to the preset.
Prompt for describing a standalone or embedded image. Switching the preset rewrites this unless you have typed your own. Leave empty to fall back to the preset.
Prompt used when Images Per Message is greater than 1. Only the Default preset fills this in — the OCR presets never batch. Leave empty to fall back to the preset.
Run every PDF page through the model instead of only pages whose extracted text looks poor. Presets other than Default turn this on regardless.
Execution output after extraction completes.
Extracted document pages with AI-generated descriptions and images.