* Add 4 more AL/BC testing patterns from Luc van Vugt's fluxxus.nl blog Fourth batch from CURABIS ApS, mined from an external BC/NAV testing expert's blog archive (fluxxus.nl). Confirm+StrSubstNo interaction with ConfirmHandler, Table Relation Test's OnAfterRemoveTableRelation exclusion hook (verified against BCApps source, codeunit 134926), committing shared lazy-Initialize fixture data, and Assert.IsFalse vs asserterror for boolean checks. * Address Jesper Schulz-Wedde's review on PR #159 - transactionmodel-attribute-governs-test-transactions.md: the "Commit causes an error" behavior is specific to an explicitly declared AutoRollback attribute. A test method with no TransactionModel attribute at all is a distinct, valid shape — BCApps' own codeunit 134915 "ERM Online Mapping Setup" commits inside a lazy Initialize() with no attribute declared, cleaning up via a manual asserterror at the end. Evidence for commit-shared-test-fixture- inside-lazy-initialize.md (this PR), which is correct as submitted. - confirm-needs-strsubstno-before-confirmhandler-sees-substituted-text.md: reframe as a known, unconfirmed-fix platform defect (microsoft/ALAppExtensions#23935) rather than designed behavior; add the Message/MessageHandler asymmetry as supporting evidence. - table-relation-test-exclude-known-invalid-relations-via-event.md: note the test-app-only consumer dependency; correct "walks every TableRelation field property in the app" to the actual tenant-wide Table Relations Metadata scope across installed apps. - Wire confirm-needs-strsubstno, commit-shared-test-fixture-inside- lazy-initialize, and table-relation-test-exclude-known-invalid- relations-via-event into al-testing-review.md's candidate-selection cues. Co-Authored-By: Claude Sonnet 5 <noreply@anthropic.com> * Address second round of Jesper Schulz-Wedde's review on PR #159 - commit-shared-test-fixture-inside-lazy-initialize.md: fundamentally rewritten. AutoCommit is the documented default TransactionModel, not AutoRollback. Explains the real mechanism (Commit() protects a fixture from the test method's own later deliberate rollback, per Codeunit.Run/ TransactionModel-property semantics) and the TestIsolation dependency (Disabled/Codeunit survive across methods, Function does not). Fixtures rewritten to demonstrate the actual failure/success shape. - transactionmodel-attribute-governs-test-transactions.md: now states the AutoCommit default explicitly and agrees with the article above, closing the contradiction Jesper flagged between the two testing articles. - Deleted confirm-needs-strsubstno-before-confirmhandler-sees-substituted-text (.md/.good.al/.bad.al): the underlying platform bug (microsoft/ ALAppExtensions#23935) was closed as completed in Feb 2024; cannot be reproduced or bc-version-pinned on any currently supported version. - table-relation-test-exclude-known-invalid-relations-via-event.md: added the [Scope('OnPrem')] boundary verified against BCApps' Table Relation Test codeunit. - use-assert-isfalse-not-asserterror-for-boolean-checks.md: added a Scope section resolving the overlap with asserterror-needs-expectederror-and-code. - al-testing-review.md: fixed the shared-fixture cue to catch the actual anti-pattern instead of the compliant shape, added the missing cue for use-assert-isfalse-not-asserterror-for-boolean-checks, wired precedence between it and the generic asserterror rule, and removed the cue for the deleted article. - Added in-file Source provenance (specific fluxxus.nl post per article, with what was independently verified vs. taken from the post) to the three surviving externally-inspired articles, per Jesper's request that provenance live in the knowledge file itself, not only the PR description. Co-Authored-By: Claude Sonnet 5 <noreply@anthropic.com> * Fix remaining correctness issues from Jesper's 2026-09-15 re-review - commit-shared-test-fixture-inside-lazy-initialize: three sub-issues. Recommended TestIsolation = Codeunit instead of listing Disabled as an equal option - Disabled never rolls back at all ("tests are not isolated from each other" per the property's own docs), so a fixture this pattern commits under Disabled is permanent database contamination unless something else tears it down; Disabled is now only mentioned alongside that explicit teardown requirement. Added precedence in al-testing-review.md so the deliberate end-of-test asserterror Error(...) rollback sentinel isn't also flagged by the generic asserterror-needs-expectederror-and-code rule. Rewrote both fixtures to actually demonstrate the pattern: persisted fixture data (an Item record) instead of an empty comment, a second [Test] method that depends on the fixture surviving into it, and an explicit Subtype = TestRunner / TestIsolation = Codeunit runner codeunit. - table-relation-test-exclude-known-invalid-relations-via-event: the length/type rule was stated as one global requirement. Verified ValidateFieldRelation in codeunit 134926 directly (BCApps reference clone) and split it into the two branches the source actually has: a field with any unconditional relation needs exact length and exact resolved type; a field whose relations are all conditional only fails on being shorter (longer is fine) than the largest related field, and when the required type is specifically Code, a Text source passes too - a tolerance that does not apply on the unconditional side and does not extend to a required Text. Rebased onto upstream/main (one conflict in transactionmodel-attribute-governs-test-transactions.md - upstream had already linked its sample references via the READ convention, ours added a Source section; merged both). Also converted the 3 remaining plain-backtick sample references in this PR to the READ-convention markdown-link form, same fix as #156/#157/#158. * Fix four merge-critical issues from Jesper's 2026-09-22 review - al-testing-review.md: the generic ExpectedError cue's asserterror Assert.IsTrue/IsFalse exclusion was unconditional, but the specialized rule it deferred to only claims the pure-inversion shape. A test expecting the guarded Boolean-returning call itself to raise fell through both routes. Narrowed the exclusion to the same inversion-only condition the specialized cue already uses. - asserterror-needs-expectederror-and-code.md: the rollback-sentinel exception (a trailing asserterror Error(...) used purely to force a fixture rollback, not to verify a specific failure) previously lived only in skill routing prose. Encoded it directly in the article's Anti Pattern section so every consumer of the knowledge base sees it, not just this one skill. - commit-shared-test-fixture-inside-lazy-initialize.good.al/.bad.al: replaced hand-rolled Item.Init()/Insert(true) with LibraryInventory.CreateItem, so the canonical fixture doesn't itself trigger use-library-codeunits-for-test-fixtures. - table-relation-test-exclude-known-invalid-relations-via-event.good.al/ .bad.al: declared minimal "Sample Setup"/"Sample Header" tables inline instead of referencing undefined symbols, matching this repo's own convention that every fixture is self-contained. * Give the table-relation-test fixtures a real relation to exclude and one to protect The "Category Code" field had no TableRelation at all, so the good subscriber's RemoveTableRelation call targeted metadata that never existed - a no-op. Added a real TableRelation to "Sample Setup" on that field (the one known exception to exclude) and a second, ordinary self-referencing relation ("Parent No." -> "Sample Header"."No.") with no exception. The good fixture now removes only the first; the bad fixture's table-wide removal (field/related table/field all 0) now demonstrably also strips the second, showing the actual anti-pattern instead of removing nothing meaningful. --------- Co-authored-by: Claude Sonnet 5 <noreply@anthropic.com> Co-authored-by: Jesper Schulz-Wedde <jesper.schulzwedde@microsoft.com> Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>
20 KiB
| kind | id | version | title | description | inputs | outputs | bc-version | technologies | countries | application-area | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| action-skill | al-testing-review | 1 | AL testing review | Performs an AL testing review against guidance from BCQuality. |
|
|
|
|
|
|
AL testing review
Reviews AL source changes against the testing knowledge domain in BCQuality and emits a findings report. This is a leaf action skill: it invokes no sub-skills. It is one of the skills composed by al-code-review.
An orchestrator invokes this skill with a pr-diff, file-path, or folder-path. Testing findings are narrow by design — they apply when the review scope contains test codeunits, test runners, test methods, handlers, assertions, or fixture construction. The skill returns not-applicable when none of those apply.
Source
Use READ's Bounded retrieval for review skills workflow with -Domain testing. Consume every catalog page across enabled layers before applying this leaf's Relevance and Worklist; preserve each exact catalog path and open complete bodies only for exact paths selected by the Worklist. If the helper or prepared index is unavailable or invalid, use READ's explicit path-discovery and bounded native-read fallback.
Relevance
Apply the frontmatter matching rules defined in READ (Frontmatter matching semantics) against the task context:
bc-version— the target BC version from the PR branch'sapp.jsonor the orchestrator-supplied version. If unavailable, the dimension isunknown.technologies—[al].countries— the countries declared in the consuming app'sapp.json. Default to the orchestrator's configured context; if absent,unknown.application-area— the union of application areas declared by the changed objects. Pass the actual set; do not substitute[all]. If the area cannot be determined from the changes, the dimension isunknown.
Discard files that are not applicable. Retain conditionally applicable files (any dimension unknown) only when the orchestrator's configuration permits them; findings derived from those files MUST have confidence no higher than medium, AND the finding's message MUST name the dimension or dimensions that were unknown.
Worklist
Narrow the relevant files to the subset that applies to the changes under review. For each relevant file, compute overlap against:
- The changed AL object names and types — especially codeunits with
Subtype = Test, test runner codeunits withTestIsolation, test libraries, and codeunits that define UI handlers. - The changed methods and attributes, weighted toward
[Test],[TransactionModel(...)],[TestPermissions(...)],[HandlerFunctions(...)], handler attributes,asserterror,ExpectedError,ExpectedErrorCode, fixture initialization, and test-library calls. - Tokens extracted from the diff that relate to testing (
Subtype = Test,Subtype = TestRunner,TestIsolation,TestPermissions,Restrictive,NonRestrictive,Disabled,Permissions Mock,Library - Lower Permissions,TransactionModel,AutoRollback,AutoCommit,Commit,asserterror,ExpectedError,ExpectedErrorCode,HandlerFunctions,ConfirmHandler,MessageHandler,StrMenuHandler,ModalPageHandler,SendNotificationHandler,RecallNotificationHandler,Enqueue,Dequeue,AssertEmpty,Initialize,IsInitialized,OnTestInitialize,LibrarySetupStorage,Library Assert,LibraryVariableStorage,LibrarySales,LibraryPurchase,LibraryERM,LibraryInventory,LibraryRandom,Library - Utility,LibraryUtility,GenerateGUID,GenerateRandomCode,TestPage,.Visible(,.Enabled(,.Editable(,OpenNew,OpenView,OpenEdit,Init,Insert).
A file enters the candidate worklist when its keywords intersect the extracted tokens or its topic (derived from the index entry's path, title, and description) matches a changed object type. Read an article's full file — its ## Best Practice / ## Anti Pattern bodies — only after it makes the worklist; candidate selection uses the index alone. When the diff contains no testing-related changes by any of the above signals, return outcome: "not-applicable" without evaluating files.
The following targeted checks cover every current testing article. Treat each as a candidate-selection cue: when the signal appears in changed code, add the named article to the worklist and evaluate it in Action.
- A method in a
Subtype = Testcodeunit adds or changes[TransactionModel(...)], exercises code that callsCommitunderAutoRollback, defaults broadly toAutoCommit, or choosesNonefor a writing test —transactionmodel-attribute-governs-test-transactions. - A new or changed
[Test]procedure is added, whether or not it already carries[FEATURE]/[SCENARIO]/[GIVEN]/[WHEN]/[THEN]tags —test-feature-scenario-tags. A procedure with no tags at all, or a generic name likeTest1, is the anti-pattern signal; presence of the tags is the compliant shape, not the thing to search for. - A test codeunit calls
TestPagemethods (OpenNew,OpenView,OpenEdit) alongside[Test]procedures in the same codeunit that call business-logic procedures directly with noTestPageinvolved —ui-test-codeunit-naming. The anti-pattern signal is both kinds of test mixed into one codeunit (or, on a project using the_UTconvention, a UI-layer codeunit missing the suffix); a codeunit containing onlyTestPage-driven tests is not itself a violation. - A
[GIVEN]-tagged setup precedes a posting call or report execution and does not visibly set up posting-group/VAT setup records, an explicit date, or (for a report test) both an included and an excluded record —given-blocks-must-cover-full-precondition-chain. - A test procedure contains more than one
[WHEN]block, or more than one distinct action not labelled[GIVEN], without the procedure name declaring a flow/defect-then-fix shape —test-one-when-per-test. - A
BCPT*scenario codeunit is added and the PerformanceTest app's only other scenario codeunits are copies of Microsoft's shipped BCPT samples (BCPT Create Customer,BCPT Create Item Journal,BCPT Post GL Entries, etc.) with no scenario exercising the extension's own codeunits, FlowFields, or pages —bcpt-scenarios-must-be-app-specific. - An
AutoCommittest runs under aSubtype = TestRunnercodeunit that omitsTestIsolationor sets it toDisabled, leaving committed data between tests —testisolation-belongs-on-the-test-runner. Require runner/repository context; a standalone test file cannot prove which runner executes it. - A permission-sensitive test uses
TestPermissions = Disabled, claims to test a restricted user without"Permissions Mock"/"Library - Lower Permissions", or declares[TestPermissions(...)]without applying that context —permission-tests-must-lower-the-execution-context. - Test fixture code manually calls
Init/Insert, invents keys or prerequisite records, or bypasses availableLibrarySales,LibraryPurchase,LibraryERM,LibraryInventory,LibraryRandom, or equivalent library codeunits —use-library-codeunits-for-test-fixtures. - A test codeunit's
Initializeprocedure exits onIsInitializedbefore per-test reset such asLibraryVariableStorage.Clear,LibrarySetupStorage.Restore, orLibraryTestInitialize.OnTestInitialize, or a[Test]method in a codeunit using that pattern does not callInitialize()first —reset-per-test-state-before-the-isinitialized-guard. asserterroris added or changed without a followingAssert.ExpectedError,Assert.ExpectedErrorCode, or a purpose-built assertion such asExpectedTestFieldError—asserterror-needs-expectederror-and-code. Excludeasserterror Assert.IsTrue(...)/asserterror Assert.IsFalse(...)only when it is used solely to invert the guarded call's Boolean result (the same conditionuse-assert-isfalse-not-asserterror-for-boolean-checkscues on below, which wins for that shape) — not when the test expects the guarded Boolean-returning call itself to raise an error, which this rule still owns even though it happens to wrap anAssert.IsTrue/IsFalsecall. Also exclude a trailingasserterror Error(...)used purely as an end-of-test rollback sentinel after a lazyInitialize()fixture already committed — that shape belongs tocommit-shared-test-fixture-inside-lazy-initialize, which wins for it; the sentinel's own error text is not meant to be asserted against.asserterrorwrapsAssert.IsTrue(BooleanExpression, ...)(or theIsFalsemirror) solely to invert the boolean result of the guarded call, rather than to assert that call itself raises an error —use-assert-isfalse-not-asserterror-for-boolean-checks.- A shared/lazy
Initialize()-style fixture helper creates fixture data without a followingCommit(), in a test method whose body later forces its own rollback (for exampleasserterror Error(...)used for end-of-test cleanup) —commit-shared-test-fixture-inside-lazy-initialize. The presence ofCommit()after the fixture is the compliant shape, not the signal to look for; the missing-Commit()shape combined with a later deliberate rollback is the anti-pattern. Require runner/repository context for theTestIsolationvalue: a standalone test file cannot prove which runner executes it, and underFunction-level isolation this whole pattern is moot regardless ofCommit()— do not raise the finding when the executing runner'sTestIsolationis known to beFunction. - Changed code subscribes to
OnAfterRemoveTableRelation, callsRemoveTableRelation, or referencesCodeunit "Table Relation Test"/134926 —table-relation-test-exclude-known-invalid-relations-via-event. - Test fixture code assigns a hardcoded literal to a primary-key field or a field the test relies on as a unique lookup identifier, hand-builds a "unique" value for such a field (string concatenation, a counter,
Format(CurrentDateTime)), or truncatesLibraryUtility.GenerateGUID()'s result withCopyStrfor such a field shorter than 10 characters —use-generateguid-for-unique-test-fixture-values. CallingGenerateGUID()untruncated into a full-length field,GenerateRandomCodeWithLengthfor a shorter field needing real verified uniqueness, orGenerateRandomCode20specifically for aCode[20]field, is the compliant shape, not the signal to flag.GenerateRandomCode20is not a substitute forGenerateRandomCodeWithLengthon a shorter field — it truncatesGenerateGUID()'s sequential value down to the field's length by keeping the leftmost characters, which change the slowest, so retries against a short field can churn through the same truncated prefix far longer thanGenerateRandomCodeWithLength's equivalent. A hardcoded or deterministic value in an ordinary descriptive field is not this anti-pattern — that field carries no uniqueness constraint. Do not claimGenerateRandomCode(withoutWithLength/20) orGenerateRandomXMLTextverify uniqueness against the real table, or thatGenerateRandomCodeis collision-free even within one test run for a short field — none of that is true. - A test asserts against a
TestPagefield's.Visible()or.Enabled()—use-testpage-visible-enabled-to-verify-field-ui-state. When the assertion is against.Editable(), or the page is opened withOpenEdit()specifically to check editability —use-testpage-editable-to-verify-field-editability. - A test path raises UI and
[HandlerFunctions(...)]does not match the invoked handlers, or the test has no meaningful evidence of the UI result (for example, it treats a Boolean set before the action as proof of success) —ui-handlers-in-tests. A capture/reset/assert-after-RunModalpattern is valid. Enqueue/dequeue andAssertEmptyare required only when order, count, text, replies, or a scripted sequence is part of the contract. Only nonoptional handlers have to execute: a listed handler declared[SendNotificationHandler(true)]or[RecallNotificationHandler(true)]is optional by design, so do not treat it as unmatched when the run never raises the notification. - A test's
[GIVEN]/setup looks up a hardcoded code/number/name assumed to already exist instead of creating it, leaves a mandatory field on a created record empty, uses a value that doesn't satisfy the scenario's own explicit length/format requirement (for example a truncation test whose value never exceeds the field), or a scenario-defining value (amount, quantity, percentage, date, threshold, rounding precision) is generated/randomized instead of an explicit chosen value —test-data-must-be-random-and-complete. Generating incidental fixture values (identifiers, names, descriptions) via the standard library codeunits is the compliant shape, not the signal to flag, and neither is a short-but-valid value in an otherwise-unremarkable field.
Once the candidate worklist is known, resolve layer-precedence conflicts per READ. Drop lower-precedence files whose normative guidance (## Best Practice or ## Anti Pattern) directly contradicts a higher-precedence candidate, and record each dropped file in suppressed with reason: "layer-precedence". Files that would have been candidates but are hidden because their layer is disabled in consumer configuration are recorded with reason: "configuration". Files that never became candidates are NOT recorded in suppressed.
When the post-conflict worklist is empty because no applicable testing knowledge exists, or because configuration suppressed every candidate, emit outcome: "no-knowledge". When the worklist is empty because no applicable testing knowledge matched the changes, emit outcome: "completed" with an empty findings array.
Action
For each worklist entry, evaluate the diff against the file's ## Best Practice and ## Anti Pattern sections. Emit findings as follows:
- When the diff contains a clear match for an Anti Pattern, emit a finding with severity
majororblocker, a message summarizing the anti-pattern,locationpointing to the offending line or range, and areferencesentry pointing to the knowledge file. Useblockeronly when the test can pass while verifying the wrong behavior or can leave committed data that contaminates later tests; otherwise the ceiling ismajor. - When the diff contains code that contradicts a Best Practice without being a full anti-pattern, emit
minorwith the same reference shape. - Applicability alone is not a finding. Emit
infoonly for a concrete, non-actionable observation the article explicitly defines; otherwise emit nothing when no violation is present.
For ui-handlers-in-tests, use major when missing or incorrectly listed handlers make the test fail at runtime. Use minor when the test executes but lacks a meaningful semantic postcondition, including a pre-set Boolean used as proof. Do not escalate solely because a handler does not use queue storage or asserts inside the handler.
Set confidence to:
highwhen the detection is based on an unambiguous pattern match (attribute, handler declaration, assertion sequence, or fixture call).mediumwhen detection relies on heuristics or when any frontmatter dimension wasunknown.lowwhen the finding is an advisory derived only from applicability.
After evaluating each worklist entry, also consider whether the diff exhibits a testing defect the agent recognises from its general AL knowledge that no knowledge file in the worklist covers. Such candidates are agent findings within this skill's domain — emit them with references: [], an id slug prefixed with agent:, confidence capped at medium, severity capped at minor (agent findings are advisory and non-gating), and a message that is self-contained (describing both the issue and a concrete recommendation, since there is no knowledge-file footer for the consumer to fall back on). Hold every candidate to the precision bar in skills/do.md (Agent findings): emit only a concrete, material testing defect a knowledgeable BC reviewer would agree is wrong — steelman it first and drop anything stylistic, speculative, dependent on code outside the diff, or merely a valid alternative; when in doubt, omit. The scope is strictly AL testing; defects outside this domain belong to other leaves and MUST NOT be emitted here. Before emitting, check the worklist for a knowledge file that matches the candidate — if one exists, upgrade the candidate to a knowledge-backed finding instead. See skills/do.md for the full contract.
For every emitted finding, decide whether the fix is mechanical. A fix is mechanical when it is small, local, and unambiguous from the diff context (for example: add the matching ExpectedError assertion after asserterror; add or remove a handler name in HandlerFunctions, except that a listed optional notification handler must never be proposed for removal; add LibraryVariableStorage.Clear or AssertEmpty when queue/LVS intentionally verifies interaction order, count, text, replies, or a scripted sequence; or replace hand-rolled fixture creation with an evident library call). For mechanical findings, emit findings[].suggested-code with the literal replacement for the source lines indicated by location. The payload must be a verbatim replacement — no diff markers, no fences, no commentary — that the consumer can render as a one-click suggestion. When a .good.al companion exists and the diff context matches the .bad.al shape, adapt the .good.al replacement into suggested-code.
Omit suggested-code only when the appropriate fix depends on context the skill cannot determine, when multiple defensible replacements exist, or when the fix spans non-contiguous code. If a finding is mechanical-looking but you omit suggested-code, set findings[].suggested-code-omission-reason to a short explanation. See skills/do.md for the full contract.
Outcome selection:
completed— the skill evaluated every worklist item.no-knowledge— no applicable testing knowledge survived filtering.not-applicable— the diff touches no test codeunit, runner, method, handler, assertion, or fixture surface.partial— a budget was hit before the worklist was exhausted.failed— an unrecoverable error occurred.
Output
Output conforms to the DO output contract. Every finding this skill emits MUST set findings[].domain to "Testing". A populated example:
{
"skill": { "id": "al-testing-review", "version": 1 },
"outcome": "completed",
"summary": {
"counts": { "blocker": 0, "major": 1, "minor": 0, "info": 0 },
"coverage": { "worklist-size": 1, "items-evaluated": 1 }
},
"findings": [
{
"id": "microsoft/knowledge/testing/asserterror-needs-expectederror-and-code.md",
"severity": "major",
"message": "The negative test uses asserterror without checking the resulting message or error code, so any unrelated setup or permission error can make the test pass.",
"location": {
"file": "test/SalesPostingTests.Codeunit.al",
"line": 42
},
"references": [
{ "path": "microsoft/knowledge/testing/asserterror-needs-expectederror-and-code.md" }
],
"confidence": "high",
"domain": "Testing",
"suggested-code": "asserterror PostInvalidOrder();\nAssert.ExpectedError(ExpectedPostingErr);"
}
],
"suppressed": []
}
The empty-corpus case produces:
{
"skill": { "id": "al-testing-review", "version": 1 },
"outcome": "no-knowledge",
"summary": {
"counts": { "blocker": 0, "major": 0, "minor": 0, "info": 0 },
"coverage": { "worklist-size": 0, "items-evaluated": 0 }
},
"findings": [],
"suppressed": []
}