Performance
Time, quality, cost and scope of the product obtained with assistance, discounting review, integration and error correction.
What does it really improve and for what tasks?Case studies · Organizations
An organization does not become dependent because it uses a lot of AI, but when it stops knowing how to decide, verify, learn or continue without it. The dependency can be hidden by good productivity indicators until a plausible error, a crash, a change of supplier or a situation that requires undocumented judgment appears.
It does not stop adoption nor does it propose to keep everything manually. It helps decide where to accelerate, where to introduce control and what capabilities should be maintained because they support security, learning or adaptation.
Indicative overview
Small teams · concentrated knowledge
Illustrative values: show relative priorities, not results of an evaluation.
Double control panel
Responsible adoption measures both the value created and the capacity that remains within the organization. What seems like efficiency today may turn into dependence on a tool, a supplier, or a few people tomorrow. The double scorecard avoids optimizing the product while degrading supervision, collective learning or continuity.
Time, quality, cost and scope of the product obtained with assistance, discounting review, integration and error correction.
What does it really improve and for what tasks?What people and teams understand, remember, and can reconstruct, plus the tacit knowledge they need to interpret exceptions.
What do we know to do if we withdraw aid?Ability to detect errors, absorb a fall, change suppliers and continue operating safely without improvising responsibilities.
What happens when the AI fails or is not there?Who receives the benefits, who endures the change, who gets training, and who retains voice, mobility, and learning opportunities.
Does the improvement reduce or widen inequalities?Field evidence · customer service
In a company with 5,172 agents, assistance increased average productivity by 15% and that of less experienced profiles by 30%. The study also observed changes in quality and retention, and signs of learning in periods without access. It is a relevant result, but it comes from a specific company, tool and task; it does not by itself describe other sectors or the long-term effect.
Brynjolfsson, Li & Raymond, 2025 ↗Delegation architecture
The same profession contains operations with very different risks, required learning and verification possibilities. Classifying an entire position as “automatable” erases that diversity. Define a task, identify who receives its effects, and choose a case to see how the control design changes.
Recommended layout
You can automate parts, but you should not erase assertion authorship or traceability.
The matrix guides; does not certify. Legal risk, security, privacy and rights require specific evaluations.
Guidance self-diagnosis
Assess observable practices, not intentions or isolated documents. Ask whether there is an up-to-date inventory, authority to disagree, independent verification, unaided practice, rehearsed continuity and participation. The result helps start a conversation and choose priorities; it is not a validated scale or a compliance audit.
Indicative maturity
Controls are already appearing, although they still depend on local initiatives or specific people.
Define what AI can propose, who decides and what decisions are not delegated.
Periodically evaluate essential unassisted tasks and schedule deliberate practice.
Editorial heuristics. The six dimensions synthesize principles of human factors, organizational learning, and governance frameworks. They do not allow comparing companies or inferring clinical, legal or financial risk.
Operating system
The cycle references the Govern, Map, Measure, and Manage functions of the NIST AI RMF, and explicitly adds human capacity maintenance. Each layer must leave operational evidence: responsible parties, inventories, tests, thresholds, incidents and exercises that can be reviewed when the model or process changes.
Govern
Connect each use to a legitimate purpose, an acceptable level of risk, and a person with real authority.
A person can say who is responsible for each assisted decision.
Literacy and supervision are not synonymous with a generic course. Article 4 of the European AI Regulation calls for measures adapted to knowledge, experience, training and context; article 14 requires effective human oversight for high-risk systems. This case study does not determine legal classification or prove compliance.
Consolidated text ↗Development and organizational justice
Exposure depends on the tasks and also on access, training, language, disability, job security, experience and the real possibility of participating in the redesign. Measuring only the average can obscure who loses practice, endures more review, is excluded from learning tasks, or lacks a safe alternative.
ILO · global index, 2025
1 in 4Exposure means possibility of task transformation; it is not equivalent to job replacement.
of global employment
Literacy should be tailored to tasks, risks, and experience level, with protected time for practice and feedback on real cases. A generic course does not demonstrate competence.
Those doing and receiving the work detect tacit knowledge, exceptions, review burdens, and invisible costs that rarely appear in a demonstration.
Accessibility, language, digital competence, age, caring responsibility and employment status may require different supports to achieve an equivalent opportunity.
Productivity must also translate into learning, quality of work, reasonable time and opportunities for progression, not just more expected volume.
Team simulation
A short practice allows you to discover dependencies that do not appear in a policy: concentrated knowledge, missing permissions, symbolic verification or an alternative that no one has tried. Choose an incident, bring together the functions involved and respond without looking for blame; the goal is to redesign the system.
Simulation · 35 minutes
The volume of work continues to come in and several tasks no longer have visible alternative procedures.
Identify blocked processes, affected people and decisions that cannot wait.
Activate the approved alternative and reduce the range before improvising with another tool.
Assign priorities, authority and return criteria to normal service.
Record what knowledge was missing and schedule a new continuity test.
First cycle
Inventory real uses, select three priority tasks and appoint those responsible.
Compare assisted and unassisted results, incorporate controls and listen to the teams.
Approve criteria, train by function, rehearse incidents and review indicators.
Evidence and frameworks
Empirical results, classical frameworks, exposure indices, and normative references are separated because they answer different questions. An experiment can estimate performance on a task; a framework organizes controls; A standard establishes obligations. No single study demonstrates a general cognitive loss caused by AI or proves that an organization is well governed.
Organizational evidence still has limits. Many studies cover tasks, tools, and short periods. That is why this case study proposes to measure locally, publish limits and review decisions when the context changes.