Files
dotnet__skills/plugins
Amaury Levé 8fb17964bc Improve test smell skill quality and eval power (#1056)
* Improve test smell skill quality

Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>

* Use conventional empty class bodies

Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>
Copilot-Session: 718f824b-6b86-4d7f-9428-2d7a8908e95b

* Improve test smell calibration

Align workspace discovery and false-positive decisions with the losing eval transcripts, correct contradictory fixtures, and make graders outcome-focused.

Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>

Copilot-Session: 718f824b-6b86-4d7f-9428-2d7a8908e95b

* Make eval regexes multiline-safe

Allow outcome evidence to match across line breaks in generated review output.

Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>

Copilot-Session: 718f824b-6b86-4d7f-9428-2d7a8908e95b

* Make notification fixtures observable

Record notification identifiers so post-wait assertions can fail, while preserving fixed sleeps as the intentional smell under evaluation.

Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>

Copilot-Session: 718f824b-6b86-4d7f-9428-2d7a8908e95b

* Strengthen test smell stop conditions

Require workspace discovery, preserve formal skip and file classifications, prevent clean-suite false positives, and reduce lexical grader coupling.

Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>

Copilot-Session: 718f824b-6b86-4d7f-9428-2d7a8908e95b

* Use conventional exception class body

Keep the fixture compatible with compilers that do not accept semicolon-only class declarations.

Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>

Copilot-Session: 718f824b-6b86-4d7f-9428-2d7a8908e95b

* Remove brittle eval gates

Rely on outcome rubrics instead of narrow lexical matches and keep the Sensitive Equality fixture culture-stable.

Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>

Copilot-Session: 718f824b-6b86-4d7f-9428-2d7a8908e95b

* Add JUnit eval exit check

Fail fast on empty or failed trial output while dropping a redundant severity-word matcher.

Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>

Copilot-Session: 718f824b-6b86-4d7f-9428-2d7a8908e95b

* Make remaining eval regexes multiline-safe

Allow concise verdict and async-fix patterns to match wrapped model output across line breaks.

Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>

Copilot-Session: 718f824b-6b86-4d7f-9428-2d7a8908e95b

* Preserve non-catalog validity findings

Keep formal smell classification while separately reporting proven test-validity defects that do not belong to the taxonomy.

Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>

Copilot-Session: 718f824b-6b86-4d7f-9428-2d7a8908e95b

---------

Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>
Copilot-Session: 718f824b-6b86-4d7f-9428-2d7a8908e95b
2026-08-25 13:57:41 +00:00
..
2026-08-03 10:19:29 +00:00
2026-08-03 10:19:29 +00:00
2026-08-24 09:18:10 +00:00
2026-08-03 10:19:29 +00:00
2026-08-17 09:16:22 +00:00
2026-08-03 10:19:29 +00:00
2026-08-03 10:19:29 +00:00
2026-08-17 09:16:22 +00:00
2026-08-24 09:18:10 +00:00
2026-08-03 10:19:29 +00:00