The Aprentiz Proving Ground

Before you automate,
prove the work.

Two connected environments in development: one to establish how a professional workflow should operate, and another to test how its implementation will deliver that contract.

The purposeTurn a promising AI experiment into a defined workflow, measurable acceptance criteria and a reviewable implementation decision.

01 / Workflow Proving Ground

Establish what useful,
accountable work looks like.

Practitioners bring the context, difficult exceptions and judgement. Aprentiz is designed to make the process, evidence and decision boundaries explicit and testable.

01Define the real job

Name the question, business scope, useful outcome and responsible professional. Make exceptions and limits explicit.

02Admit the evidence

Identify the source versions and organisational material permitted for the work. Record what is missing or excluded.

03Test the workflow

Exercise the analysis, evidence review, human decisions and failure paths. Compare the proposed process with the current way of working.

04Agree the contract

Retain the stages, roles, evidence requirements, acceptance cases and unresolved questions that implementation must preserve.

THE INTENDED OUTPUT

A tested workflow contract.

Defined inputs. Named roles. Evidence standards. Failure states. Acceptance scenarios. A clear account of what remains unresolved.

02 / Implementation Proving Ground

Carry that contract
into the system.

The technical implementation inherits the workflow’s requirements. The design keeps the original need connected to each integration, control, test and later change.

01Translate into requirements

Map each workflow need to a system behaviour, data boundary, responsible owner and test.

02Build a bounded candidate

Define integrations, model roles, controls and recovery. Make the smallest change needed to test the chosen question.

03Challenge the whole journey

Test evidence quality, identity, isolation, missing dependencies, hostile inputs, interruption and recovery.

04Make the release decision

Bring the results, limitations and operational requirements to the authorised people. A successful demonstration is one piece of that decision.

THE INTENDED OUTPUT

An evidenced implementation decision.

Traceable requirements, integration maps, control tests, qualification results and the limitations an authorised release decision must consider.

The AI research underneath

Choose the model
through the work.

Model evaluation supports both Proving Grounds. The research question is which configuration can perform the defined work usefully, at an acceptable complete cost, within the required controls.

01 / QUALITY

Does the evidence hold?

Test source support, wrong or missing citations, contradictions, appropriate refusal and loss of conditions during context compression.

02 / OPERATION

Will it fit the workload?

Measure latency, active requests, context limits, memory, recovery and the full cost of a useful completed workflow.

03 / DECISION

Is the change worth keeping?

Freeze the baseline and evaluation criteria. Retain negative results. Improve, narrow, keep the baseline or stop when the evidence calls for it.

Research thresholds are agreed before testing. A model score, agreement between models or a small error-free test set does not establish professional correctness or readiness for release.