KognitCapture
Capture, extract and structure data from any document — on your own infrastructure, with no page meter running.
An API-first intelligent document capture platform. Documents arrive from thirteen channels, pass through a sixteen-step image pipeline, are read by local OCR, local models or twelve remote providers, are mapped onto versioned templates, reviewed by people where it matters, and delivered to thirteen kinds of downstream system. Built as a modern replacement for legacy capture suites.
Document capture is priced by the page and delivered by the quarter.
The incumbent platforms in this category share two properties. They meter you per page, so the more value you extract the more you pay. And they take eight to twelve weeks of specialist consulting to configure a single document type. Meanwhile the cloud alternatives require your invoices, claim forms and identity documents to leave your building — which for a growing number of buyers ends the conversation.
One platform, on your hardware, with the economics inverted.
Capabilities
- Thirteen import channels: upload, REST, SFTP, FTP, hot folder, e-mail, Kafka, RabbitMQ, ZeroMQ, webhook, script, plugin, dataset.
- Driverless network scanning via server-side eSCL discovery, or direct-attach through the browser — nothing to install on the scanning desk.
- Scan Hub sessions survive a closed laptop: reorder, rotate, delete, rescan, separate, then submit as one batch.
- Eight recognition engine modes across Tesseract 5, native PDF text, hybrid and open data loading.
- 129 bundled OCR language packs, warm pooled engines, no external OCR service.
- Local models in-process plus twelve remote providers behind one interface with failover.
- Barcode and QR reading on a dedicated decode path — never guessed from OCR text.
- 31 field types and 10 extraction methods — zone, label search, regex, keyword proximity, model, manual, anchor text, companion file, classification, XML mapping.
- Versioned templates that integrations pin to, so a design change never breaks a live interface.
- Table and line-item extraction with column zones, row grouping, header detection and continuation across pages.
- Template generation from a single sample page, and training from operator corrections.
- Four machine-learning frameworks and 29 model types running in-process — no external inference service.
- Document splitting on fixed page count, blank page, barcode match, model detection, header/footer or physical separator sheets.
- Reference datasets for lookup, autofill and validation, plus template alternatives for per-supplier layout variants.
- 16 workflow step types, from import and OCR through verification hubs to export and custom scripts.
- A visual designer where the drawing is the workflow, with decision gateways on document properties and field values.
- Three operator hubs — Verification, Processing and Scan — with the page image beside the fields and click-to-fill.
- Per-field trail: raw recognised value, original value, who verified it and when.
- 14 validation rule types, from not-empty and ranges through checksum, expression, script, REST call and database lookup.
- Custom scripts in JavaScript, Groovy or Java, sandboxed, versioned and testable against sample documents.
- Check-out locking so two operators never touch the same work; public single-document review links for outside parties.
- Thirteen export destinations and eight formats, including searchable PDF and PDF/A-2b built from persisted word geometry.
- Field mapping per export goal — unmapped fields are dropped by design.
- Eleven external connection types, each with a setup wizard and a test action.
- Encrypted per-project credential vault.
- Plugin SDK for compiled workflow steps, importers and exporters, with their own migrations and endpoints.
- A native Model Context Protocol server with around 109 curated tools across projects, documents, verification, recognition, templates, workflows and administration.
- OAuth 2.1 with dynamic client registration and an explicit consent screen.
- Scoped tools with risk tiers, and a separate audit channel that records agent reads, not only writes.
- Embeddable Verification and Scan Hubs inside your own product, in an iframe, behind an API key, a short-lived session and a named origin.
- Duty- and capability-aware cluster nodes that claim work without contention — drain a node, resume it, or dedicate it to OCR, export or import.
What it does
- 01No licence gate, no page counter — the marginal cost of a page is the electricity to process it.
- 02Six industry blueprints provision a full working project in one transaction, against an industry norm of eight to twelve weeks.
- 03Local OCR and local models keep regulated documents on your own infrastructure, with remote AI providers as an option, never a requirement.