Qualimetry
Enterprise
The two questions every engineering leader is being asked about AI. Qualimetry answers both from your own estate: a disposition for every system based on pressure and feasibility, and an adoption measure built from signals your organisation controls.
Readiness scoring is not an abstract exercise. A system is only safe to hand over once the rules that govern it are somewhere an agent can read, and once what comes back can be measured against those same rules. Every input to the score below depends on the rest of the platform.
Architecture standards, language standards, principles and policies exist as governed content rather than as knowledge held by the people who built the system.
Those standards reach the agent's context automatically, so it works to your architecture rather than to a generic idea of good.
Five perspectives check the change against the same standards, so scale does not mean unreviewed volume.
A compliance score per file and per project, gating the merge, is what turns a rebuild programme into something you can report on.
Without those four, an agentic readiness score is a guess about a system nobody can govern. With them, it is a decision you can defend.
Two axes decide it. Replacement pressure is what the business is paying to keep the system as it is. Replacement feasibility is how safely agents could take it on. Where a system lands gives you the verdict, and the evidence behind it.
Every system is placed on two axes. Replacement pressure is how much the business is paying to keep it as it is. Replacement feasibility is how safely agents could take it on, given test coverage, interface clarity, specification stability and the risk of cutting over. The four resulting verdicts are Rebuild, Contain, AI-maintain and Maintain.
Readiness is a weighted average of the criteria that actually determine whether agents can work safely on a system. A criterion with nothing to measure scores neutral, never zero, so an unmeasured system is not falsely condemned.
Agents can start here now.
Fix the blockers first, usually coverage or interface clarity.
Too much would have to change before this is safe.
Agentic work is not the right answer for this system.
Start from one of four profiles, then adjust the weights directly. Nothing is hidden in a black box, and the ranking recalculates against the criteria below.
What gets scored
Codebase size, test coverage, active churn, contributor concentration and bus risk, tech-stack obsolescence, stack standardness, maintenance burden, interface clarity, specification stability, verification readiness, business risk during cutover, dependency exposure, compliance, lifecycle, and end-of-life or known-exploited exposure.
A project starts an agentic programme on its first agent-attributed pull request, and the programme completes after thirty days without another. You do not have to declare a programme, and nobody has to remember to close one.
Qualimetry does not infer authorship from how code is written. Stylometry is a deliberate non-goal, because a number produced that way cannot be defended when someone senior challenges it. Adoption is measured from two independent signals your organisation controls.
Two independent signal sources are used and are never blended: work observed to come from a recognised agent account, and work that follows a convention your organisation agreed, such as a branch prefix, a title prefix or a pull request label. Dependency and CI bots are removed from both the numerator and the denominator. What remains is reported as a floor, because work that follows no convention and comes from no recognised account cannot be attributed even if an agent produced it. Writing style is never used to guess.
The two lanes are counted separately and never blended.
An observed agent is a pull request authored by an account you recognise as an agent. Claude, Cursor, Codex, GitHub Copilot and Blitzy are recognised out of the box, and you can add your own accounts and markers.
A convention is a marker your teams agreed to use: a branch prefix, a title prefix, or a pull request label. It catches agentic work that runs under a human account.
How attributed agentic work has moved over time, split by the signal that identified it.
How much of your estate is even capable of being attributed, which is the honest health check on the number itself.
Where agentic work concentrates across your own business structure, whether that is domain, journey or owning team.
Which agents are actually being used, with convention-only work shown as untracked rather than misattributed.
Volume is the easy question. The one that matters to a risk committee is whether agentic changes are being reviewed at all.
Work that comes from no recognised account and follows no agreed convention cannot be attributed, even when an agent produced it. Qualimetry says so, in the product, next to the number. A measure you can defend when challenged is worth more than a measure that flatters the programme.
None of this works as a separate product. Each stage exists because of the one before it and is only worth doing because of the one after it.
Book a demo and see your own systems ranked by disposition, readiness and the weights that matter to you.