SETTING GLOBAL STANDARDS FOR TRUSTED AI CREDENTIALSAI Competence Framework v1.29 · current release
You are reading the current version of the framework, v1.29.Permanent address for this version

Produce a set of test cases covering expected, edge and failure conditions for a defined task.

Type Skill · introduced in version 0.1

Performance indicators

Normative. These state what would be observed in a person who holds the statement.

Includes cases expected to fail, not only cases expected to pass
Covers realistic inputs rather than convenient ones
States the expected outcome for each case before running it
Evidence examples

Non-normative. Illustrative of evidence an awarding body might accept; not a required form.

A worked artefact produced in the course of normal duties, with the reasoning recorded at the time
Attestation by a competent supervisor against the indicators above, not against a general impression
Relationships

Assumes

Statements a candidate is taken to hold already. Never at a higher level than this one.

This statement assumes no other statement. It is a starting point within its domain.

Assumed by

Derived inverse. Statements that take this one as given.

D3.L3.05

Evaluate a change to an instruction against a case set rather than by impression.

D4.L2.06

Test an automated workflow before it runs on real work.

D6.L2.04

Check whether a change to a prompt, model or configuration has altered previously acceptable output.

D6.L3.02

Design an evaluation set representative of production conditions and resistant to overfitting.

O3.L2.02

Test a stated control and record whether it operates as described.

X-SWE.L2.02

Test AI-generated code, including cases it was not written for.

Related

Cross-domain relationships, stated in both directions and typed in the content model.

D5.L2.05

Identify the evidence a decision to build would require.

Referenced by

Entries on the register of conformance claims whose coverage map cites this statement.

IBAIC Inc.

Accredited awarding body — full credential portfolio assessed against the AI Competence Framework · Tier III

Provenance
IntroducedVersion 0.1
Last modifiedNot modified since introduction
Statusactive · stable
Version displayedv1.29
Permanent URLaicertificationstandards.org/framework/statements/D6.L2.02
Cite this statement

AI Certification Standards (2026) AICF D6.L2.02, D6 Evaluation and assurance, version 1.29. Available at aicertificationstandards.org/framework/v1.29/statements/D6.L2.02 (accessed date).

Propose an amendment

Statements change through the published process, not by correspondence

An amendment to this statement, its indicators or its relationships is proposed through the contribution process. Every submission is answered on the record, and the reasoning for acceptance or rejection is published in the release record for the version that follows.

Propose an amendment to D6.L2.02

Writing rules and the controlled verb list that govern how this statement is worded are published in the methodology.

This page displays version 1.29 · last reviewed 30.08.2026