Turn statements into structured data, including the custodians that publish no feed.
Every automated data flow assumes the custodian publishes a feed. A feed is a commercial and technical decision somebody else makes, and for a meaningful part of any multi-custody book that decision has gone the other way.
An API, SWIFT, or FIX feed from the custodian. Positions, transactions, and valuations arrive on a schedule, reconciled, and nobody touches them. This is what WealthArc Data Feeds already does across a network of more than 170 institutions.
For any of these reasons a live feed may simply not exist, and waiting for one is not a plan.
What you can always get is a statement. WealthArc Data Intelligence reads that statement and produces the same structured data a feed would have delivered, so a custodian that will not connect stops being a hole in the book.
WADI is the first of WealthArc's released AI agents. It reads financial documents, extracts and standardizes every field, validates the result against the reference data already in your WealthArc Data Box, and delivers structured output to whatever runs downstream. It is a module of the infrastructure, not a tool beside it.
The engine classifies the document, finds the regions that carry data, extracts every field into a predefined schema, and checks the result against your own portfolio and reference data before anything is released.
| Holding | ISIN | Quantity | Market value |
|---|---|---|---|
| Global Equity Fund | XX0000000101 | 12,400 | 3,118,880 |
| Corporate Bond 2.25% 2031 | XX0000000208 | 1,500,000 | 1,442,250 |
| Money Market Fund | XX0000000315 | 9,880 | 988,120 |
| Emerging Markets Equity | XX0000000422 | 4,260 | 1,704,000 |
| and 30 further positions | |||
DisclaimerThe interface shown is an illustration, not a live product. All names and figures shown are sample data. ISINs are placeholders and do not identify real securities.
A document arrives and comes out the other end as data in your WealthArc Data Box, on the same path every custodian feed takes. There is no configuration to write and no template to maintain per custodian.
WealthArc clients read it directly in Portfolio Management and Portfolio Viewer. Everything else is served through Prism, the API and MCP distribution layer, so every system you run reads the same validated data rather than its own copy.
Once a channel is configured it runs in the background, so documents arrive without anyone sending them. Manual upload stays available for the one-off that turns up by email.
WealthArc Data Intelligence collects documents from your own tenant-side storage on a schedule, so the operations team keeps working the way it already works. Nothing has to be forwarded.
Custodians deliver documents straight into dedicated, secure storage inside your WealthArc environment, which is the closest thing to a feed a custodian without one can offer.
Drag and drop, either in the Data Intelligence interface or embedded inside WealthArc Portfolio Management, for the document that arrives by email the day before a committee meeting.
Extraction is against a predefined schema per document type, covering the identifiers and dates that place the document, the headline metrics, and every row of every table it contains.
documentTypecustodianaccountNumberportfolioName
portfolioReferenceclientReferencereferenceCurrency
valuationDatestatementDateperiodStartperiodEnd
bookingDatevalueDate
totalPortfolioValueassetsUnderManagementnetContribution
performanceTWRperformanceMWRaccruedInterest
isininstrumentNameassetClassquantity
pricepriceCurrencymarketValuemarketValueRefCcy
costPriceunrealisedPnLweight
transactionTypeinstrumentquantityprice
grossAmountfeestaxesnetAmountcounterparty
cashBalancecurrencycashMovementTypeamount
fxRateincomeType
NoteField names are illustrative of the schema shape. The exact schema per document type is set during implementation, and the output maps into WISDOM investment objects either way.
The document types that carry most of the data in a multi-custody book, delivered as PDF.
Private markets paperwork arrives as PDF almost everywhere, and is still read by hand in most places. Parsed into structured positions, the illiquid side of the balance sheet lands in the same record as everything else.
Reading a document is the easy half. The half that decides whether the output is usable is what happens next: three layers of checking, and an optional human sign-off before anything reaches a downstream system.
Is every field the schema requires actually there? A statement missing a valuation date or a reference currency is caught before it is parsed further, not after it has been loaded.
Are the values possible? Prices outside a sensible range for the date, quantities that do not match a known instrument, and totals that do not add up are flagged rather than passed on.
Does it agree with what you already hold? Extracted values are checked against the reference and portfolio data in your WealthArc Data Box, which is the check a standalone tool cannot make.
Flagged at layer 2: a price outside the plausible range for the valuation date. Nothing on this document is released until it is resolved, and the resolution is recorded against the field.
DisclaimerThe interface shown is an illustration, not a live product. All names and figures shown are sample data.
An agent is only as good as what it reads. Point one at a folder of PDFs and it will answer confidently and inconsistently, because there is nothing underneath the answer. Normalization and validation are not housekeeping steps ahead of the interesting part; they are what makes the interesting part trustworthy.
The same holds for your own models. Data stays inside your WealthArc environment and is readable only by the models and agents you authorize, over the API or the MCP server.
Any modern model will read a statement and return something plausible. The question is what happens when it is wrong, and whether you would know.
Almost nobody delivers this well, which is why almost everybody has the problem. Because the same engine can parse a statement from 2015 as easily as one from last month, history stops being a box of PDFs and becomes data in your WealthArc Data Box, at the same quality as anything arriving today.
DisclaimerThe chart shown is an illustration, not a live product. All names and figures shown are sample data.
An asset manager wants a new client. Instead of reading years of statements by hand, they parse the prospect's existing statements and get structured historic data straight away: enough to run real analytics and build a data-driven proposal on the spot. What took days becomes a same-meeting turnaround.
Clients need historic performance and holdings that predate their relationship with their current provider, either to meet record-keeping requirements or to give a new client continuous reporting that includes the history from before they onboarded. Today that data is trapped in old statements. This makes it usable again.
Because WealthArc Data Box can absorb historic data regardless of source, nobody is stuck with a data provider purely because that provider is the only one holding their history. That removes a switching barrier which most providers create simply by being unable to backfill.
“Your current provider is the reason you feel locked in. We remove that.”
Digitize custodian statements across a multi-custody book without re-keying, and close the accounts where no feed will ever exist.
Turn capital calls, distributions, and NAV letters into structured positions, so the illiquid side of the balance sheet stops living in a spreadsheet.
Automate extraction for regulatory and investor reporting, for managers and administrators alike, on the document types that arrive as PDF by default.
One consolidated data layer regardless of which custodian produced the statement, so consolidation is not limited to the custodians that happen to publish an interface.
Consolidate trust and foundation holdings across custodians into one clean, auditable record, with the supporting statements and deeds stored alongside the data they produced.
Take structured output through Prism and ship a product on top of it, rather than building document parsing and a validation layer of your own.
WealthArc Data Intelligence plugs in on top of your existing Connector and WealthArc Data Box setup. There is no new integration to build, no new store to secure, and the output lands in the same place your feeds already land.
Begin with a scoped pilot on your highest-volume document type, with commercial terms to match. If the accounts with no feed are the pain, start there; if it is eleven years of statements, start there instead.
Close the gap where a custodian offers no API, SWIFT, or FIX feed at all, and never will. WealthArc Data Intelligence ingests the PDF statements instead and delivers them as a document based data feed, the same structured positions, transactions, and valuations a live feed would have produced, so held-away portfolios and smaller or newly onboarded custodians stop being keyed in by hand.
Parse years of legacy custodian bank statements into clean, structured historic data with WealthArc Data Intelligence, instead of a manual read-through. Turn a prospect pitch into a same-meeting, data-driven proposal, satisfy regulatory record-keeping requirements, and let clients switch providers without losing their history.
Regularly check the underlying portfolios held within every trust, keep the supporting statements, deeds, and legal documents safely stored alongside the reconciled data, and generate beneficiary and regulatory reports on demand - one trusted record per trust, not a filing cabinet.
Connect every custodian, bank, and exchange, and get one clean feed out.
One system for portfolios, orders, clients, compliance, and fees, on data that is already reconciled.
One reconciled view of everything, across custodians, entities, and currencies.
Send us one statement from a custodian that will not connect, and we will send back the structured data.
Get started →