Install any skill in seconds. Free to start, no credit card required.
Get Started Free →SAP Analytics Cloud (SAC) automated testing skill for designing capability-gated browser discovery and deterministic Playwright test suites for SAC stories, dashboards, reports, planning workflows, comments, permissions, visual regression, and reusable QA automation. This skill should be used when building SAC end-to-end tests, onboarding SAC dashboards into Playwright, creating dashboard profiles or scenario YAML, using Microsoft Edge/CDP, Chrome DevTools MCP, Vercel Labs agent-browser, or manu
.claude/skills/secondsky-sap-sac-test-automation/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | 71% | 0% |
| case-02 | ✗→✓ | ▲ Improved | 531% | 0% |
| case-03 | ✗→✓ | ▲ Improved | 56% | 0% |
| case-11 | ✗→✓ | ▲ Improved | 94% | 0% |
| case-09 | ✗→✓ | ▲ Improved | 120% | 0% |
Design reusable SAC test automation as a capability-gated system: select the safest available discovery backend, require human review for profiles and baselines, then use reviewed Playwright code for deterministic execution, CI gating, reporting, and evidence.
Apply the core rule: discovery proposes, humans approve, Playwright executes, CI enforces.
@playwright/test suites, use this SAC skill as the test architecture guide and follow the local project's Playwright conventions.references/chrome-devtools-mcp.md for SAC-safe defaults and Edge boundaries.When the user is starting SAC automation or has not supplied a reviewed dashboard profile, route them through /sac-test-onboard or follow the same intake sequence manually. Use a two-stage intake:
Default to draft-only artifacts until the user explicitly confirms file creation and target directory. If writing is confirmed, use profiles/<profile-id>/intake.md, profiles/<profile-id>/dashboard.yaml, and profiles/<profile-id>/scenarios/read-only-smoke.yaml based on the bundled templates.
After a profile or scenario draft exists, route safety review to sac-test-profile-reviewer when available. Keep the reviewer focused on intake/profile/scenario safety, not broad Playwright suite implementation.
Use this skill to plan, implement, or review SAC automation involving:
Do not use this skill as the main guide for generic web applications. Use general Playwright guidance for non-SAC sites.
Apply this contract to reporting-only SAC stories and dashboards. It complements the profile-driven smoke scenario and does not authorize planning or model changes.
Browser access is an execution prerequisite, not evidence that the story exists. If browser initialization or discovery fails, including runtime errors such as agent is not defined:
sap-browser-automation and the approved Microsoft Edge/CDP recovery guidance in references/edge-cdp-enterprise.md.Treat AI/browser-agent output as a draft, not as the source of truth. Require human review for profile creation, selector approval, expected business values, visual/data baseline changes, permission matrices, and any destructive/writeback scenario.
Prefer profile-driven automation:
Load these references only as needed:
references/architecture.md: hybrid architecture, feasibility boundaries, reliable SAC test categories, and reusable project shape.references/tool-availability-and-deployment.md: backend decision matrix, Windows/restricted-environment checks, Firecrawl public-research policy, and no-tool fallbacks.references/chrome-devtools-mcp.md: Chrome DevTools MCP modes, SAC-safe configuration, tool categories, CLI usage, Windows/restricted deployment, and enterprise safety boundaries.references/edge-cdp-enterprise.md: SAC test-automation add-on for the shared sap-browser-automation authentication, profile-copy, Edge/CDP, and recovery layer.references/dashboard-profiles-and-scenarios.md: dashboard profile contract, scenario contract, adapter responsibilities, and onboarding flow.references/agent-browser-discovery.md: optional agent-browser read-only discovery workflow, command patterns, output artifacts, and human review checklist.references/playwright-execution.md: Playwright test runner guidance, auth, readiness, CI stages, and test category policy.references/governance-and-sac-testability.md: SAC testability contract, auth/SSO, planning/comment safety, baseline approval, and role governance.references/failure-triage-and-artifacts.md: required evidence, failure packet shape, root-cause categories, and performance/readiness metrics.templates/intake.md: guided intake packet for policy, tooling, roles, risk, baselines, and approvals.templates/dashboard-profile.yaml: starter dashboard profile with SAC metadata, readiness, components, roles, baselines, and risk policy.templates/scenario-read-only-smoke.yaml: selector-free starter smoke scenario using profile component IDs.When implementing against a live project, also inspect the project's existing Playwright config, package manager, CI, profile schema, and artifact conventions before adding new structure.
webSocketDebuggerUrl, or bypass Edge RemoteDebuggingAllowed policy. Attaching to or copying a daily user profile requires explicit approval and the shared skill's isolated-profile procedure./json/version or /json/list returning 404 as proof that Edge CDP is unusable; read DevToolsActivePort and use the direct browser WebSocket only when the harness supports it.Derived from incorporated SAC automated-suite planning content recorded in docs/project/sac-test-automation-source-review-2026-06-17.md, plus extracted profile/scenario templates bundled with this skill. The planning sources cite SAP Help, Vercel Labs agent-browser, and Playwright documentation. Edge/CDP and Chrome DevTools MCP guidance also considers the ChromeDevTools/chrome-devtools-mcp README, CLI docs, tool reference, troubleshooting guide, package metadata, bundled skills, issue #1235 and PR #1229, Microsoft Edge DevTools Protocol documentation, Microsoft Edge DevTools MCP guidance, Edge RemoteDebuggingAllowed policy, and Firecrawl public documentation for MCP/search/scrape safety. This skill is docs-audited only; live SAC tenant execution, Chrome DevTools MCP runtime behavior, SSO behavior, CI behavior, planning writeback, and visual baseline stability remain tenant-specific and must be validated before making runtime claims.
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→pass | 22,242 | 18,696 | -16% | 1 | 1 | 0% | 3,720 | 6,361 | +71% | 0 | 0 | — |
case-02 | fail→pass | 36,033 | 16,630 | -54% | 1 | 1 | 0% | 921 | 5,812 | +531% | 0 | 0 | — |
case-03 | fail→pass | 27,807 | 25,968 | -7% | 1 | 1 | 0% | 4,627 | 7,233 | +56% | 0 | 0 | — |
case-04 | pass→pass | 19,987 | 14,265 | -29% | 1 | 1 | 0% | 3,816 | 5,738 | +50% | 0 | 0 | — |
case-05 | fail→fail | 9,731 | 7,619 | -22% | 1 | 1 | 0% | 1,801 | 4,406 | +145% | 0 | 0 | — |
case-11 | fail→pass | 16,092 | 10,823 | -33% | 1 | 1 | 0% | 2,428 | 4,705 | +94% | 0 | 0 | — |
case-06 | fail→fail | 13,827 | 13,359 | -3% | 1 | 1 | 0% | 2,286 | 5,084 | +122% | 0 | 0 | — |
case-07 | pass→pass | 15,801 | 9,956 | -37% | 1 | 1 | 0% | 2,370 | 4,702 | +98% | 0 | 0 | — |
case-08 | pass→pass | 14,981 | 14,520 | -3% | 1 | 1 | 0% | 2,351 | 5,589 | +138% | 0 | 0 | — |
case-09 | fail→pass | 16,006 | 15,231 | -5% | 1 | 1 | 0% | 2,536 | 5,578 | +120% | 0 | 0 | — |
case-10 | fail→pass | 18,396 | 11,342 | -38% | 1 | 1 | 0% | 2,639 | 4,929 | +87% | 0 | 0 | — |
case-12 | pass→pass | 15,150 | 12,522 | -17% | 1 | 1 | 0% | 2,306 | 5,076 | +120% | 0 | 0 | — |
case-13 | fail→pass | 14,721 | 10,874 | -26% | 1 | 1 | 0% | 2,117 | 4,650 | +120% | 0 | 0 | — |
case-14 | fail→pass | 14,888 | 14,485 | -3% | 1 | 1 | 0% | 2,338 | 5,550 | +137% | 0 | 0 | — |
case-15 | pass→pass | 12,288 | 8,979 | -27% | 1 | 1 | 0% | 1,934 | 4,424 | +129% | 0 | 0 | — |
case-16 | pass→pass | 16,260 | 7,618 | -53% | 1 | 1 | 0% | 2,615 | 4,237 | +62% | 0 | 0 | — |
case-17 | pass→pass | 11,739 | 8,469 | -28% | 1 | 1 | 0% | 1,728 | 4,322 | +150% | 0 | 0 | — |
case-18 | fail→pass | 17,838 | 10,151 | -43% | 1 | 1 | 0% | 2,779 | 4,574 | +65% | 0 | 0 | — |
case-19 | pass→pass | 16,194 | 11,978 | -26% | 1 | 1 | 0% | 2,736 | 5,083 | +86% | 0 | 0 | — |
case-20 | pass→pass | 12,501 | 8,057 | -36% | 1 | 1 | 0% | 1,908 | 4,354 | +128% | 0 | 0 | — |
case-21 | fail→pass | 15,377 | 10,754 | -30% | 1 | 1 | 0% | 2,283 | 4,801 | +110% | 0 | 0 | — |
case-22 | fail→pass | 7,935 | 3,295 | -58% | 1 | 1 | 0% | 1,213 | 3,566 | +194% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 22 cases were attempted, and 21 counted toward the lift figure. The other 1 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +50 percentage points is the difference between those two pass rates over the 21 comparable cases.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
Other measured skills in the registry, with their headline benchmark lift.