platform-apex-test-run: Salesforce Test Execution & Coverage Analysis
Use this skill when the user needs Apex test execution and failure analysis: running tests, checking coverage, interpreting failures, improving coverage, and managing a disciplined test-fix loop for Salesforce code.
When This Skill Owns the Task
Use platform-apex-test-run when the work involves:
sf apex run test workflows
- Apex unit-test failures
- code coverage analysis
- identifying uncovered lines and missing test scenarios
- structured test-fix loops for Apex code
Delegate elsewhere when the user is:
- writing or refactoring production Apex →
platform-apex-generate skill
- testing Agentforce agents →
agentforce-test skill
- testing LWC with Jest → experience-lwc-generate
Required Context to Gather First
Ask for or infer:
- target org alias
- desired test scope: single class, specific methods, suite, or local tests
- coverage threshold expectation
- whether the user wants diagnosis only or a test-fix loop
- whether related test data factories already exist
Recommended Workflow
1. Discover test scope
Identify:
- existing test classes
- target production classes
- test data factories / setup helpers
2. Run the smallest useful test set first
Start narrow when debugging a failure; widen only after the fix is stable.
3. Analyze results
Focus on:
- failing methods
- exception types and stack traces
- uncovered lines / weak coverage areas
- whether failures indicate bad test data, brittle assertions, or broken production logic
4. Run a disciplined fix loop
When the issue is code or test quality:
- delegate code fixes to
platform-apex-generate skill when needed
- add or improve tests
- rerun focused tests before broader regression
5. Improve coverage intentionally
Cover:
- positive path
- negative / exception path
- bulk path (251+ records where appropriate)
- callout or async path when relevant
High-Signal Rules
| Rule |
Rationale |
Default to SeeAllData=false |
Ensures test isolation; prevents reliance on org-specific data |
| Every test must assert meaningful outcomes |
Tests with no assertions prove nothing and give false confidence |
| Test bulk behavior with 251+ records |
Triggers process in batches of 200; 251 records crosses the boundary |
Use factories / @TestSetup when they improve clarity |
Consistent data creation in one place; rolled back between test methods |
Pair Test.startTest() with Test.stopTest() for async |
Ensures async operations (queueable, future) complete before assertions |
| Do not hide flaky org dependencies inside tests |
Prevents intermittent failures tied to org state |
Gotchas
| Issue |
Resolution |
| Test passes locally but fails in CI org |
Check for SeeAllData=true or undeclared dependencies on org-specific records |
| Coverage drops unexpectedly after refactor |
Run focused class-level tests first, then widen to RunLocalTests to confirm |
| "Uncommitted work pending" error in callout test |
DML and HTTP callouts cannot be mixed in the same test context without Test.startTest() wrapping |
| Mock not taking effect in test |
Ensure Test.setMock() is called before the code that makes the callout |
@TestSetup data missing in test method |
@TestSetup data is committed per test method — re-query it; do not store in static variables |
Output Format
When finishing, report in this order:
- What tests were run
- Pass/fail summary
- Coverage result
- Root-cause findings
- Fix or next-run recommendation
Suggested shape:
Test run: <scope>
Org: <alias>
Result: <passed / partial / failed>
Coverage: <percent / key classes>
Issues: <highest-signal failures>
Next step: <fix class, add test, rerun scope, or widen regression>
Cross-Skill Integration
| Need |
Delegate to |
Reason |
| Fix production code or author test classes |
platform-apex-generate skill |
Code generation and repair |
| Create bulk / edge-case test data |
platform-data-manage |
Realistic test datasets |
| Deploy updated tests to org |
platform-metadata-deploy |
Deployment workflows |
| Inspect detailed runtime logs |
platform-apex-logs-debug |
Deeper failure analysis |
Reference File Index
| File |
When to read |
references/cli-commands.md |
All sf apex run test command flags, output formats, async execution, and coverage commands |
references/test-patterns.md |
Test class templates — basic, bulk (251+), mock callout, and data factory patterns |
references/testing-best-practices.md |
Core testing principles — AAA pattern, naming conventions, bulk, negative, and mock strategies |
references/test-fix-loop.md |
Agentic test-fix loop implementation and failure analysis decision tree |
references/mocking-patterns.md |
HttpCalloutMock, DML mocking, StubProvider, and selector mocking patterns |
references/performance-optimization.md |
Techniques to reduce test execution time — DML mocking, SOQL mocking, loop optimizations |
assets/basic-test.cls |
Template: standard test class with @TestSetup, positive / negative / bulk / edge-case methods |
assets/bulk-test.cls |
Template: bulk test with 251+ records that crosses the 200-record trigger batch boundary |
assets/mock-callout-test.cls |
Template: HTTP callout mock using HttpCalloutMock |
assets/test-data-factory.cls |
Template: reusable TestDataFactory with create and insert helpers |
assets/dml-mock.cls |
Template: IDML interface + DMLMock implementation for database-free unit tests |
assets/stub-provider-example.cls |
Template: StubProvider-based dependency injection stub |
scripts/parse-test-results.py |
Post-tool hook — parses sf apex run test JSON output and formats failures for the auto-fix loop |
Score Guide
| Score |
Meaning |
| 108+ |
strong production-grade test confidence |
| 96–107 |
good test suite with minor gaps |
| 84–95 |
acceptable but strengthen coverage / assertions |
| < 84 |
below standard; revise before relying on it |
1---2name: platform-apex-test-run3description: Apex test execution, coverage analysis, and test-fix loops with 120-point scoring. Use when the user needs to run Apex tests, check code coverage, fix failing tests, or work with *Test.cls / *_Test.cls files. TRIGGER when: user runs Apex tests, checks code coverage, fixes failing tests, or touches *Test.cls / *_Test.cls files. DO NOT TRIGGER when: writing Apex production code (use platform-apex-generate), Agentforce agent testing (use agentforce-test), or Jest/LWC tests (use experience-lwc-generate).4---56# platform-apex-test-run: Salesforce Test Execution & Coverage Analysis78Use this skill when the user needs **Apex test execution and failure analysis**: running tests, checking coverage, interpreting failures, improving coverage, and managing a disciplined test-fix loop for Salesforce code.910## When This Skill Owns the Task1112Use `platform-apex-test-run` when the work involves:13- `sf apex run test` workflows14- Apex unit-test failures15- code coverage analysis16- identifying uncovered lines and missing test scenarios17- structured test-fix loops for Apex code1819Delegate elsewhere when the user is:20- writing or refactoring production Apex → `platform-apex-generate` skill21- testing Agentforce agents → `agentforce-test` skill22- testing LWC with Jest → [experience-lwc-generate](../experience-lwc-generate/SKILL.md)2324---2526## Required Context to Gather First2728Ask for or infer:29- target org alias30- desired test scope: single class, specific methods, suite, or local tests31- coverage threshold expectation32- whether the user wants diagnosis only or a test-fix loop33- whether related test data factories already exist3435---3637## Recommended Workflow3839### 1. Discover test scope40Identify:41- existing test classes42- target production classes43- test data factories / setup helpers4445### 2. Run the smallest useful test set first46Start narrow when debugging a failure; widen only after the fix is stable.4748### 3. Analyze results49Focus on:50- failing methods51- exception types and stack traces52- uncovered lines / weak coverage areas53- whether failures indicate bad test data, brittle assertions, or broken production logic5455### 4. Run a disciplined fix loop56When the issue is code or test quality:57- delegate code fixes to `platform-apex-generate` skill when needed58- add or improve tests59- rerun focused tests before broader regression6061### 5. Improve coverage intentionally62Cover:63- positive path64- negative / exception path65- bulk path (251+ records where appropriate)66- callout or async path when relevant6768---6970## High-Signal Rules7172| Rule | Rationale |73|------|-----------|74| Default to `SeeAllData=false` | Ensures test isolation; prevents reliance on org-specific data |75| Every test must assert meaningful outcomes | Tests with no assertions prove nothing and give false confidence |76| Test bulk behavior with 251+ records | Triggers process in batches of 200; 251 records crosses the boundary |77| Use factories / `@TestSetup` when they improve clarity | Consistent data creation in one place; rolled back between test methods |78| Pair `Test.startTest()` with `Test.stopTest()` for async | Ensures async operations (queueable, future) complete before assertions |79| Do not hide flaky org dependencies inside tests | Prevents intermittent failures tied to org state |8081---8283## Gotchas8485| Issue | Resolution |86|-------|------------|87| Test passes locally but fails in CI org | Check for `SeeAllData=true` or undeclared dependencies on org-specific records |88| Coverage drops unexpectedly after refactor | Run focused class-level tests first, then widen to `RunLocalTests` to confirm |89| "Uncommitted work pending" error in callout test | DML and HTTP callouts cannot be mixed in the same test context without `Test.startTest()` wrapping |90| Mock not taking effect in test | Ensure `Test.setMock()` is called before the code that makes the callout |91| `@TestSetup` data missing in test method | `@TestSetup` data is committed per test method — re-query it; do not store in static variables |9293---9495## Output Format9697When finishing, report in this order:981. **What tests were run**992. **Pass/fail summary**1003. **Coverage result**1014. **Root-cause findings**1025. **Fix or next-run recommendation**103104Suggested shape:105106```text107Test run: <scope>108Org: <alias>109Result: <passed / partial / failed>110Coverage: <percent / key classes>111Issues: <highest-signal failures>112Next step: <fix class, add test, rerun scope, or widen regression>113```114115---116117## Cross-Skill Integration118119| Need | Delegate to | Reason |120|------|-------------|--------|121| Fix production code or author test classes | `platform-apex-generate` skill | Code generation and repair |122| Create bulk / edge-case test data | [platform-data-manage](../platform-data-manage/SKILL.md) | Realistic test datasets |123| Deploy updated tests to org | [platform-metadata-deploy](../platform-metadata-deploy/SKILL.md) | Deployment workflows |124| Inspect detailed runtime logs | [platform-apex-logs-debug](../platform-apex-logs-debug/SKILL.md) | Deeper failure analysis |125126---127128## Reference File Index129130| File | When to read |131|------|-------------|132| `references/cli-commands.md` | All `sf apex run test` command flags, output formats, async execution, and coverage commands |133| `references/test-patterns.md` | Test class templates — basic, bulk (251+), mock callout, and data factory patterns |134| `references/testing-best-practices.md` | Core testing principles — AAA pattern, naming conventions, bulk, negative, and mock strategies |135| `references/test-fix-loop.md` | Agentic test-fix loop implementation and failure analysis decision tree |136| `references/mocking-patterns.md` | HttpCalloutMock, DML mocking, StubProvider, and selector mocking patterns |137| `references/performance-optimization.md` | Techniques to reduce test execution time — DML mocking, SOQL mocking, loop optimizations |138| `assets/basic-test.cls` | Template: standard test class with `@TestSetup`, positive / negative / bulk / edge-case methods |139| `assets/bulk-test.cls` | Template: bulk test with 251+ records that crosses the 200-record trigger batch boundary |140| `assets/mock-callout-test.cls` | Template: HTTP callout mock using `HttpCalloutMock` |141| `assets/test-data-factory.cls` | Template: reusable `TestDataFactory` with create and insert helpers |142| `assets/dml-mock.cls` | Template: `IDML` interface + `DMLMock` implementation for database-free unit tests |143| `assets/stub-provider-example.cls` | Template: `StubProvider`-based dependency injection stub |144| `scripts/parse-test-results.py` | Post-tool hook — parses `sf apex run test` JSON output and formats failures for the auto-fix loop |145146---147148## Score Guide149150| Score | Meaning |151|---|---|152| 108+ | strong production-grade test confidence |153| 96–107 | good test suite with minor gaps |154| 84–95 | acceptable but strengthen coverage / assertions |155| < 84 | below standard; revise before relying on it |