Skip to content

Include low-severity findings in PR Review benchmarks - #885

Open
Wenjie Fan (gggdttt) wants to merge 1 commit into
mainfrom
gggdttt-bcquality-plugin-comparison
Open

Wenjie Fan (gggdttt) wants to merge 1 commit into
mainfrom
gggdttt-bcquality-plugin-comparison

Conversation

@gggdttt

Copy link
Copy Markdown
Collaborator

Summary

Default the BC PR Review benchmark to Low instead of Medium, so its review scope includes low-severity gold findings.

Pass the configured floor to both MINIMUM_SEVERITY and AGENT_MINIMUM_SEVERITY. Previously, the adapter set only the latter, leaving the pinned engine's prompt and knowledge-backed findings at its default Medium floor. Both gates now honor the benchmark configuration or explicit local override, regardless of inherited severity environment variables.

Update the CLI help and code-review documentation, and cover the default, explicit overrides, ambient environment values, and Low output mapping with regression tests. Bump BC-Bench from 0.13.0 to 0.14.0 because the default review behavior changes.

Scope

No changes to the dataset, scoring, historical results, or engine/BCQuality pins. Production BC-ALAgents configuration is handled separately. This PR does not claim a measured score improvement.

Apply the configured severity to both engine gates and default to Low so low-severity gold findings remain eligible. Bump the benchmark to 0.14.0 for the changed review scope.

Co-authored-by: Copilot App <223556219+Copilot@users.noreply.github.com>
Sun Haoran (haoranpb) added a commit that referenced this pull request Sep 15, 2026
Force GitHub Actions to index an experiment ref with the exact PR #885 tree.

Co-authored-by: Copilot <223556219+Copilot@users.noreply.github.com>
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant