Comparison
This is the option that wins most of these meetings, and it deserves a straight answer rather than fear. Open-weight document models are genuinely good and genuinely free, they run on your own GPUs, and your platform team is right that the extraction problem is largely solved. The disagreement is about which part of the work is hard.
| building it in-house | Densery | |
|---|---|---|
| Extraction | Open-weight models read complex documents locally for cents per thousand pages. Your team can stand this up in weeks. They are right about this. | We buy the same layer, including the same open-weight models. We do not claim an advantage here and do not charge for one. |
| The ontology | Has to be elicited from your specialists and written down — what makes a file complete, which mismatch is material, which exception gets asked about. This is 12–24 months of work and it is nobody's day job. | Already built and running for banking, insurance, manufacturing and telecom. Extended with your specifics during implementation rather than started from zero. |
| Write-back into the system of record | Achievable, and where scope quietly triples. Every core, LOS, claims and quality system has its own permissions, service accounts and failure modes. | Through your existing APIs under your existing service identities. Where a legacy system has no API, we operate it the way your staff do, under a named identity with every action logged. |
| The evidence layer | Almost always deferred to phase two, and phase two is where the project dies. Per-decision provenance is harder to retrofit than to build in. | The reason the other three layers are deployable at all. Built first, not last. |
| Cost | Free at the margin and politically favoured. The real cost is your best ML engineers not doing something differentiating. | A line item, visible, and easy to cancel. That visibility cuts both ways and we know it. |
| Odds | MIT NANDA: purchased solutions reach production roughly twice as often as internal builds. 88% of agent pilots never make it. | Six deployments in production. Not a large number — but they are in production, in supervised institutions, and you can call them. |
| When it clearly wins | You have 30+ ML engineers and document work is core to your product rather than a cost of running it. | We disqualify those accounts ourselves. If that is you, we are not going to win and should not. |
Statements about building it in-house are drawn from their public website and published materials as at 15 August 2026 and are our reading of them, not their words. Product capability changes; check anything here that matters to your decision, and tell us if we have it wrong.
Written straight, because you will find this out anyway and it costs us less to say it now.
Narrower than the list on the left, deliberately.
We win this comparison less often than the other three, and when we lose it we usually deserve to. A team that has already shipped a governed agent into production does not need us.
But the honest version includes the base rate. Internal builds in this category fail at a high rate, and they fail late — typically after twelve to eighteen months, at the point where the evidence and governance layer turns out to be the whole problem rather than a finishing task. The extraction demo works in week three, which is exactly what makes the timeline feel achievable.
The question worth putting to your own team is not "can we build this" — they can, and telling them otherwise is insulting. It is: who writes down the ontology, when does the audit record get built, and what else stops while that happens? If those three have clean answers, build it. We mean that.
Both are ungated and neither asks for an email. Run your own volume through the calculator, then open a real file in the audit-trail explorer and click the fields that failed.
The next step
A scoping session is not a demo. Bring twenty real files, redacted if you need to. We take three numbers off you — annual volume, fully loaded cost per file today, and what happens when the output is wrong — and hand back a one-page value case in your own KPIs.
If the arithmetic says we are not a fit, we will tell you in the room rather than six weeks later.
If your data can go anywhere and your documents are already clean and digital, you do not need us. Use a hyperscaler document API and spend the money on something harder.