What we do
01  Advanced Infrastructure 02  Applied AI & Data 03  AI Cybersecurity 04  AI Assurance
Engagements
AI estate inventory Assurance review
Industries
Financial Services Government & Public Sector Energy & Utilities Telecommunications Healthcare & Life Sciences Transport & Logistics Industrial & Manufacturing Retail, Hospitality & Real Estate
Research
The Trust Maturity Model The GCC Assurance Index Readiness self-assessment Case studies Perspectives Sector briefings Technology evaluations
Company
About us Partners Events Careers Contact العربية Talk to our team
Applied AI & Data  ·  Data Organization

Where did the data come from?

AI amplifies whatever the data already was. Before a system is allowed to act on its own, somebody has to be able to say where the data came from, how current it is, and what a join quietly excluded.

6–12 WEEKSINVENTORY, LINEAGE, QUALITY THRESHOLDSFIXED, NOT JUST DOCUMENTED
The decisions underneath

Three questions decide the design.

Data work is treated as hygiene until an autonomous system makes it a governance problem. These are the three questions that decide how high the ladder can safely go.

01

Where did this actually come from

Lineage matters most at the moment somebody challenges an output. Reconstructing it afterwards is expensive and rarely convincing.

02

What did the join leave out

The most common silent failure is not bad data. It is a whole segment dropped by a null, in a pipeline nobody has read since it was written.

03

How stale is too stale

Every field feeding a limit or a decision needs a freshness threshold and an alarm. Without one, the system keeps acting correctly on a picture that is no longer true.

The ladder

Every rung is standing on the same data.

Autonomy amplifies whatever the data already was. Choose a rung to see what the data underneath has to be able to prove at each level.

AUTONOMY— OF 5 RUNGS EARNED
Select a rung

Five levels of autonomy. Each one is a different system with different consequences, and each one has a control that has to exist before you stand on it.

Data Organization is not a tidying exercise. It is the part of the programme that decides how high the ladder can safely go.

What you receive

An inventory you can stand behind.

Four stages, every engagement. Hover a stage to see what happens in it.

DURATION
6–12 weeks
DELIVERABLE
Dataset inventory, lineage, quality thresholds
DELIVERED
Remotely; in-country where residency requires
INDICATIVE FEE
[FEE BAND — pending sign-off]
The boundary

We build it. We do not grade our own work.

Three moves. Two of them are ours, and the one in the middle is deliberately somebody else’s.

MOVE 01 — OURS

We build and run it

Design, integration and operation of the system, with a named human accountable at every gate that matters. This is delivery work and we do not pretend otherwise.

MOVE 02 — NOT OURS

Somebody else forms the opinion

The assurance opinion on anything we built is not written by the team that built it, and where you need it to carry weight externally, not by Orvix at all. A firm that audits its own delivery is offering you a marketing document.

MOVE 03 — OURS

We remediate what the review found

We fix what the independent review says is wrong, and the fix is re-examined by the same reviewer rather than signed off by us.

This is the same independence test we apply to other people’s vendors. It would be difficult to argue for it and then exempt ourselves.
Applied AI & Data

The rest of this pillar.

Three engagements inside this pillar. Start with the question you can name, or take the whole estate at once.

Questions we are asked

Before you ask us.

Is this a data platform project?
Not usually. It is an assessment and a remediation of what the AI programme is standing on. Where a platform is genuinely required we say so, and you contract it directly.
Our data is fine. The model is the problem.
It may well be. The test is cheap either way: reconcile the training or retrieval set against source and measure what fraction of the records the model never saw. That number is usually the fastest explanation available for a model behaving oddly.
How does this connect to residency?
Every dataset rests somewhere, and so does the catalogue describing it. Where regulated data is in scope, jurisdiction is recorded per dataset as part of the count, before any recommendation is made.
Does this have to happen before the AI work?
Not entirely, but the inventory does. Building an autonomous system on an uninventoried data estate means the first serious question about an output will take weeks to answer.

Start with the extract nobody owns.

Thirty minutes. Name the dataset that feeds a decision and has no named owner, and we will start there.

Book a 30-minute scoping call

ORVIX · INDEPENDENT AI & TECHNOLOGY ASSURANCE · WE DISCLOSE EVERY COMMERCIAL RELATIONSHIP ON THE PAGE FOR THE SERVICE IT BELONGS TO. WHERE LICENSING IS REQUIRED, DELIVERY IS PERFORMED BY NAMED PARTNERS UNDER THEIR OWN LICENCE.