Roadmap
What is shipped, what is next, what comes later.
Version 1.1.3 is the current release of Docuoria, in public preview. Everything on this page is stated without dates.
Status
What public preview means
| Live today | The engine, the skill, both installers, licence acquisition, the Free and Pro plans, and the self-hostable template store host |
| Verified | The six use cases on the use cases page, run end to end on real invoices and receipts and checked against hand-verified values |
| Not yet verified | The document types listed as roadmap use cases |
| May change on notice | Plan prices and limits (Terms of Service, section 8) |
| Support | GitHub issues and inquiries@sidub.net. The software is provided as is (Terms of Service, section 10) |
Shipped in 1.1.3
The engine and everything around it
- A stateless, deterministic, in-process PDF extraction engine
- Seven match rules for recognising document types
- Five extraction sources and a fallback source
- Five pipeline steps: extraction, transformation, retrieval, Python, publish
- CSV and JSON output, including one row per line item
- Collection extraction from tables and repeating patterns
- Running ledgers that append month over month without duplicates
- Authoring tools: inspect, pattern tests, dry-run, diagnostics, regression checks
- Nineteen CLI scripts with a fixed JSON contract
- The Docuoria skill for seven AI tools, installed by one command from npm or NuGet
- A self-hostable template store API host for teams
- Free and Pro plans with licence acquisition from the CLI or the Monaiq marketplace
Next
What is being built
| Initiative | What it adds |
|---|---|
| Hosted template store service | A Sidub-hosted store for teams to share, version, and discover templates, on top of the self-hostable host that ships today |
| Template improvement suggestions | The engine proposes refinements from extraction success rates and failure patterns |
| Processing offload | Extraction in the cloud for large batch workloads, by choice |
| Email attachment automation | Inbound email, PDF attachment, extraction, delivery downstream |
| ERP and accounting integrations | Extracted data pushed into QuickBooks, Xero, SAP, and other systems |
| XML output | A third output generator |
| Additional match rules | Font fingerprint, structural, image profile, embedded content, and a model-assisted rule |
Later
The platform around the engine
Service layer
Accounts, subscription management, and template resolution
REST API and client SDKs
A hosted API with typed .NET and Node.js clients
Web portal
Template management and monitoring in a browser
Template generation from a sample
A model drafts the template from one example PDF; the engine still runs it
The engine you install today is the layer everything above it builds on.
Tell us what would change your month-end.
Tell us which document type or integration would change your month-end, and how you work today.
Which requests are proven today: use cases.