Nairobi-based · Serving businesses across Kenya & East Africa +254 711 362 029 [email protected]

Service · Foundations

Fix the data, and the AI becomes easy

Nobody puts “data plumbing” in a board paper. But almost every stalled AI project we are called in to rescue stalled here — duplicated customers, numbers that disagree, and a critical spreadsheet on one person’s laptop.

Data auditPipelinesSingle source of truthDashboardsGovernance

What is AI readiness?

AI readiness is the degree to which an organisation’s data, systems, processes and governance can support useful artificial intelligence. In practice it means: the data exists, it is accessible through something better than a manual export, it is accurate and de-duplicated enough to be trusted, definitions are agreed across departments, access is controlled, and there is a lawful basis and retention rule for the personal data involved.

There is a version of this work that costs a fortune and takes two years, and we are not proposing it. Enterprise data transformation programmes have a poor record everywhere, and a worse one in mid-market Kenyan businesses where nobody has a spare team to run them.

Instead we fix the data needed for the next thing you want to build, and only that. If the first automation needs clean supplier records, we fix supplier records. Customer records can wait until they block something. This sequencing is the difference between a data project that delivers and one that becomes a permanent line item.

Signs you need this

  • Finance and sales report different revenue for the same month.
  • The same customer appears three times, spelled three ways.
  • A critical calculation lives in one spreadsheet on one laptop.
  • Getting a report means asking someone to export something.
  • You cannot say how many active customers you have without a debate.
  • An AI pilot stalled because the data was not there.

In practice

The work, in order of how often it is needed

We recommend starting with the audit — it is cheap, fast, and it usually changes what you thought you needed.

Data audit & quality assessment

What data you hold, where it lives, how good it is, who owns it and what is blocking you. Delivered as a map plus a prioritised fix list with effort estimates, not a lecture on best practice.

Pipelines & integration

Scheduled, monitored, retrying flows that move data between your systems reliably — replacing the manual export-and-import ritual that fails whenever someone is on leave.

Single source of truth

Agreed definitions, identity resolution across systems so one customer is one customer, and one place everyone queries. Half technical, half diplomatic.

Dashboards & reporting

The daily, weekly and monthly numbers assembled automatically and delivered where people already look — including WhatsApp, which in Kenya beats any portal for actually being read.

Governance & access

Who can see what, retention schedules, audit logging and the documentation the Office of the Data Protection Commissioner would expect you to produce.

Legacy & paper migration

Getting historical records out of filing cabinets, old systems and dead spreadsheets into something queryable — often using document AI, which makes this far cheaper than it used to be.

The same answer, two ways

What “fixing the data” actually means

Getting your business to agree on the facts, and keeping them in one place.

Most companies do not have one set of numbers. They have several, and each department trusts its own. Sales counts a sale when the order is signed; finance counts it when the money lands; the store counts it when the goods leave. None of them is wrong, but nobody has written the differences down, so every management meeting starts with an argument about whose figure is right.

The first job is agreement. We get the relevant people in a room and write down what each number means and when it counts. This is unglamorous and occasionally tense, and it is the highest-value hour in the whole project.

The second job is joining things up. Right now your customer might be “Kamau Enterprises” in one system, “Kamau Ent Ltd” in another and a phone number in a third. We match them so that one customer is one customer, and you can finally see what any of them is actually worth to you.

The third job is making it automatic. No more Friday exports. The numbers assemble themselves, on schedule, and arrive where you already look — a dashboard, an email, or a WhatsApp message at 07:00.

Once that exists, AI becomes straightforward. Most of what looks like an AI problem was a “we could not see it” problem all along.

What you get

What you end up with

Foundations that make every later AI project cheaper, faster and less risky.

  • A data map: what you hold, where it lives, how good it is, who owns it.
  • A prioritised fix list with effort and impact against each item.
  • Automated, monitored pipelines replacing manual exports.
  • A modelled data store with tested transformations in version control.
  • Agreed metric definitions, documented and referenced everywhere.
  • Dashboards answering the questions people actually ask.
  • A personal-data inventory and retention schedule for DPA 2019 compliance.

Typical engagement

Timeline2 weeks for an audit; 4–10 weeks to build
Indicative investmentFrom KES 75,000 for a data audit
Who we work withAny organisation whose numbers disagree with themselves
You ownCode, prompts, docs and accounts — outright

Every quote is fixed before work starts. If we scoped it wrong, that is our risk, not yours. See how pricing works.

Typical stack

PostgresBigQuerydbtAirbyteAirflowPower BILooker StudioPythonMetabase

We choose tools your team can maintain, not the ones that make us look clever.

Questions

Data & AI Readiness — your questions answered

Why do we need data work before AI?

Because an AI system is only as good as what it can see. If your customer exists three times under slightly different names, no model can tell you how much that customer is worth. If your stock figures are wrong, a demand forecast built on them will be confidently wrong too. Most AI projects that quietly fail did not fail on the model; they failed because the data underneath was fragmented, duplicated or simply absent — and nobody checked before the budget was approved.

How bad is our data, really?

Probably about the same as everyone else’s, which is worse than you think and less catastrophic than you fear. The common pattern we find in Kenyan organisations: a reliable transactional core, a duplicated and inconsistent customer master, spreadsheets holding critical logic that exists nowhere else, and one system whose data nobody trusts but everyone uses. A two-week audit tells you exactly where you stand, with a prioritised fix list.

Do we need a data warehouse?

Often not, and we will say so. If you have three systems and a few million rows, a well-designed Postgres database with scheduled syncs will serve you for years at a fraction of the cost and complexity. A warehouse earns its keep when you have many sources, real volume, or several teams querying independently. Selling infrastructure a Kenyan mid-market business does not need is one of the more expensive things a consultancy can do to you.

What is a single source of truth?

An agreed answer to questions like “how many customers do we have?” or “what did we sell last month?” that everyone in the business uses. Today those numbers usually differ by department because each pulls from a different system with different rules and different cut-off times. Establishing a single source of truth is partly technical — pipelines, definitions, identity matching — and partly political, because someone has to accept that their familiar number was wrong.

Does this help with data protection compliance?

Substantially. You cannot comply with the Kenya Data Protection Act, 2019 if you do not know what personal data you hold, where it lives, who can see it and how long you keep it. The data mapping we produce for AI readiness is the same artefact your compliance obligations require, so the two pieces of work should be done together rather than twice.

Can you build our dashboards too?

Yes, and we usually do, because a dashboard is how the data work becomes visible to the people who paid for it. We build in Power BI, Looker Studio or a lightweight custom front end depending on what your team can maintain. The rule we hold to: a dashboard must answer a specific recurring question that someone actually asks, or it will be opened twice and abandoned.

Keep exploring

Other ways we help

Most clients combine two or three of these.

Start here

Let’s find the hours your business is losing

Book a free 30-minute call. We will map one process end to end, tell you honestly whether AI is the right answer, and put a number on what fixing it is worth. No slide deck, no jargon.

Nairobi-based · We reply the same working day · English & Kiswahili

Talk to a human

Phone & WhatsApp +254 711 362 029 Email [email protected]
Hours Mon–Fri 8:30–18:00 · Sat 9:00–13:00 EAT
Chat with us