Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Use when implementing a GitHub issue, bug fix, or feature in flutter_rust_bridge end-to-end: develop the change, add regression coverage, prepare and open a PR, monitor CI until green, request and resolve Gemini review, and keep following up on a 5 minute cadence until the PR is ready.
.claude/skills/fzyzcjy-frb-issue-to-green-pr/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-04 | ✗→✓ | ▲ Improved | 149% | 0% |
| case-05 | ✗→✓ | ▲ Improved | 84% | 0% |
| case-11 | ✗→✓ | ▲ Improved | 50% | 0% |
| case-16 | ✗→✓ | ▲ Improved | -17% | 0% |
| case-17 | ✗→✓ | ▲ Improved | 27% | 0% |
Use this skill as the top-level workflow when the user asks you to implement an issue, fix a bug, add a feature, or otherwise "handle it end to end" in flutter_rust_bridge, especially when they ask for an automatic PR or CI monitoring.
This skill orchestrates the narrower FRB skills. Do not duplicate their detailed instructions in memory; read them when their phase begins.
Always read these first:
frb-dev-env and any applicable user-level environment rules.frb-develop-feature for the reproduce -> fix -> final regression workflow.Read these when entering the matching phase:
frb-code-generation before running generation or changing Rust APIs, codegen, or example APIs.frb-test before running local tests.frb-lint before lint or format checks.frb-prepare-pr before pushing or opening the PR.frb-pr-review before treating a non-trivial PR as ready.frb-fix-ci before diagnosing any CI failure.tom-ci skill for CI waiting, status inspection, and GitHub Actions logs.frb-manual-test before writing a manual regression test report under tools/manual_tests/.frb-ci-filter before creating an intentional red CI reproduction PR or using filtered CI.frb-debugging when generated code is surprising or codegen behavior is unclear.tom-ci skill between CI processing actions until the PR is ready.git status --short and do not disturb unrelated user or multi-agent changes.frb-develop-feature.Reproduce ISSUE_SUMMARY with intentional red CI.Add manual reproduction for ISSUE_SUMMARY.frb-ci-filter and make the reproduction PR an intentional red CI PR: unchanged fix code, minimal reproducer or workflow adjustment, mandatory focused ci_filter run, and a failure whose error matches the user's report.frb-manual-test and make the independent reproduction PR add or update tools/manual_tests/NAME.md with a normal manual test procedure and mechanical execution steps an agent or human can run.frb-develop-feature.frb-develop-feature Final Placement Gate: final regression coverage belongs in frb_example/pure_dart with generated pure_dart_pde coverage, not only in frb_example/dart_minimal.frb-prepare-pr.frb_example/dart_minimal. If it does, stop PR preparation and migrate it to frb_example/pure_dart first.Close #1234, unless the active PR workflow explicitly requires an empty body.frb-pr-review for the full PR review gate, including correctness review and test-weakening review.frb-fix-ci and the user's tom-ci skill, diagnose the latest relevant failure, fix it, commit, push, and continue monitoring.tom-ci skill to wait on the PR URL after each push, rerun, or handled CI event.tom-ci skill.| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-01 | fail→fail | 13,525 | 17,287 | +28% | 1 | 1 | 0% | 1,227 | 2,030 | +65% | 0 | 0 | — |
case-02 | fail→fail | 31,505 | 14,528 | -54% | 1 | 1 | 0% | 4,218 | 2,054 | -51% | 0 | 0 | — |
case-03 | fail→fail | 20,857 | 13,446 | -36% | 1 | 1 | 0% | 254 | 2,127 | +737% | 0 | 0 | — |
case-04 | fail→pass | 27,105 | 24,495 | -10% | 1 | 1 | 0% | 910 | 2,265 | +149% | 0 | 0 | — |
case-05 | fail→pass | 12,822 | 8,717 | -32% | 1 | 1 | 0% | 1,092 | 2,014 | +84% | 0 | 0 | — |
case-06 | pass→pass | 16,573 | 13,362 | -19% | 1 | 1 | 0% | 1,582 | 2,829 | +79% | 0 | 0 | — |
case-07 | fail→fail | 13,919 | 12,767 | -8% | 1 | 1 | 0% | 1,234 | 2,798 | +127% | 0 | 0 | — |
case-08 | pass→pass | 16,028 | 10,241 | -36% | 1 | 1 | 0% | 1,683 | 3,253 | +93% | 0 | 0 | — |
case-09 | pass→pass | 14,546 | 6,882 | -53% | 1 | 1 | 0% | 1,493 | 2,391 | +60% | 0 | 0 | — |
case-10 | pass→pass | 14,482 | 9,207 | -36% | 1 | 1 | 0% | 1,474 | 2,221 | +51% | 0 | 0 | — |
case-11 | fail→pass | 17,680 | 12,005 | -32% | 1 | 1 | 0% | 1,861 | 2,787 | +50% | 0 | 0 | — |
case-12 | pass→pass | 18,685 | 10,845 | -42% | 1 | 1 | 0% | 1,949 | 2,399 | +23% | 0 | 0 | — |
case-13 | pass→pass | 12,846 | 5,811 | -55% | 1 | 1 | 0% | 1,215 | 2,340 | +93% | 0 | 0 | — |
case-14 | pass→pass | 12,358 | 2,638 | -79% | 1 | 1 | 0% | 1,873 | 1,977 | +6% | 0 | 0 | — |
case-15 | fail→fail | 13,714 | 2,795 | -80% | 1 | 1 | 0% | 1,202 | 2,054 | +71% | 0 | 0 | — |
case-16 | fail→pass | 15,390 | 7,769 | -50% | 1 | 1 | 0% | 2,360 | 1,949 | -17% | 0 | 0 | — |
case-17 | fail→pass | 15,706 | 19,855 | +26% | 1 | 1 | 0% | 2,077 | 2,641 | +27% | 0 | 0 | — |
case-18 | fail→pass | 41,798 | 4,704 | -89% | 1 | 1 | 0% | 1,709 | 2,324 | +36% | 0 | 0 | — |
case-19 | fail→fail | 14,294 | 13,607 | -5% | 1 | 1 | 0% | 1,437 | 2,859 | +99% | 0 | 0 | — |
case-20 | fail→fail | 20,440 | 3,729 | -82% | 1 | 1 | 0% | 2,209 | 2,028 | -8% | 0 | 0 | — |
case-21 | pass→pass | 18,531 | 32,687 | +76% | 1 | 1 | 0% | 3,150 | 6,035 | +92% | 0 | 0 | — |
case-22 | pass→pass | 19,309 | 15,191 | -21% | 1 | 1 | 0% | 2,573 | 3,326 | +29% | 0 | 0 | — |
case-23 | pass→pass | 16,637 | 15,320 | -8% | 1 | 1 | 0% | 1,989 | 3,462 | +74% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 23 cases were attempted, and 19 counted toward the lift figure. The other 4 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +26 percentage points is the difference between those two pass rates over the 19 comparable cases. 1 case got worse with the skill loaded, and it is included in that figure.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
| Model | Method | Date | Lift |
|---|---|---|---|
| gemini-3.6-flash | verified | 9/1/2026 | +32% |
| gemini-3.6-flash | verified | 8/11/2026 | +30% |
| gemini-3.6-flash | verified | 8/9/2026 | +59% |
Other measured skills in the registry, with their headline benchmark lift.