Skip to content
Docuoria

Roadmap

What is shipped, what is next, what comes later.

Version 1.1.3 is the current release of Docuoria, in public preview. Everything on this page is stated without dates.

Status

What public preview means

Live todayThe engine, the skill, both installers, licence acquisition, the Free and Pro plans, and the self-hostable template store host
VerifiedThe six use cases on the use cases page, run end to end on real invoices and receipts and checked against hand-verified values
Not yet verifiedThe document types listed as roadmap use cases
May change on noticePlan prices and limits (Terms of Service, section 8)
SupportGitHub issues and inquiries@sidub.net. The software is provided as is (Terms of Service, section 10)

Shipped in 1.1.3

The engine and everything around it

  • A stateless, deterministic, in-process PDF extraction engine
  • Seven match rules for recognising document types
  • Five extraction sources and a fallback source
  • Five pipeline steps: extraction, transformation, retrieval, Python, publish
  • CSV and JSON output, including one row per line item
  • Collection extraction from tables and repeating patterns
  • Running ledgers that append month over month without duplicates
  • Authoring tools: inspect, pattern tests, dry-run, diagnostics, regression checks
  • Nineteen CLI scripts with a fixed JSON contract
  • The Docuoria skill for seven AI tools, installed by one command from npm or NuGet
  • A self-hostable template store API host for teams
  • Free and Pro plans with licence acquisition from the CLI or the Monaiq marketplace

Next

What is being built

InitiativeWhat it adds
Hosted template store serviceA Sidub-hosted store for teams to share, version, and discover templates, on top of the self-hostable host that ships today
Template improvement suggestionsThe engine proposes refinements from extraction success rates and failure patterns
Processing offloadExtraction in the cloud for large batch workloads, by choice
Email attachment automationInbound email, PDF attachment, extraction, delivery downstream
ERP and accounting integrationsExtracted data pushed into QuickBooks, Xero, SAP, and other systems
XML outputA third output generator
Additional match rulesFont fingerprint, structural, image profile, embedded content, and a model-assisted rule

Later

The platform around the engine

Service layer

Accounts, subscription management, and template resolution

REST API and client SDKs

A hosted API with typed .NET and Node.js clients

Web portal

Template management and monitoring in a browser

Template generation from a sample

A model drafts the template from one example PDF; the engine still runs it

The engine you install today is the layer everything above it builds on.

Tell us what would change your month-end.

Tell us which document type or integration would change your month-end, and how you work today.

Which requests are proven today: use cases.