Understanding the trade-offs between Fast, Balanced, and High Accuracy extraction modes.
Supabind doesn't use static OCR. Instead, we use state-of-the-art Large Language Models (LLMs) with vision capabilities. Depending on the complexity of your PDF—such as scanned images, nested headers, or multi-page tables—you can choose between three distinct processing modes.
Powered by
Gemini 2.0 Flash
Fast Mode is optimized for near-instant results on digital-first PDFs. It is ideal for high-volume processing where the document layout is standardized and the text is already selectable.
Best for:
Avoid when:
Powered by
Gemini 2.5 Flash
The default setting for most Supabind users. This model offers a significant upgrade in spatial reasoning, allowing it to correctly identify table boundaries that span across multiple pages without losing row alignment.
Why we recommend it: It provides 90% of the accuracy of the Pro model at a fraction of the processing time and token cost.
Powered by
Gemini 2.5 Pro
Our most advanced document processing engine. High Accuracy mode utilizes deep reasoning to handle document layouts that defeat standard OCR engines.
Deciphers text even from handheld photos of spec sheets.
Handles multi-level headers and merged cells with precision.
Every table extracted by our AI includes a **Confidence Score**. This is visible in the extraction results panel (e.g., "Confidence: 94%").
Processing documents via AI consumes "tokens." The total token count is a combination of the input (PDF size and page count) and the output (number of rows and columns extracted).
You can view the exact Input Tokens, Output Tokens, and Total Tokens used for any job by clicking the "View Logs" button in the Processing History tab.
If your PDFs are highly unique or very low resolution, our data team can help you calibrate the AI models for your specific use case.
Contact Engineering Support