open-swe/.github/workflows
Adam Moussa 0d21fb8c47
fix: align reviewer eval with published findings (#1713)
* fix: make reviewer eval reflect published findings

Serialize and deduplicate finding persistence, align review calibration around the final six-finding publication, and make judge matching order-independent and auditable.

* fix: honor reviewer eval limits

Forward configured caps into publication snapshots and keep recall-at-cap bounded for diagnostic all-findings runs.

(cherry picked from commit 71e3b8183882bcc42e318f3f220c291617ebcb67)
Co-authored-by: Johannes du Plessis <johannes@langchain.dev>
2026-07-16 17:36:18 -04:00
..
ci.yml feat(open-swe): re-add Fable 5 behind an admin toggle (Bedrock) (#172) 2026-07-10 14:05:20 -04:00
dependency-review.yml chore(deps): bump actions/dependency-review-action from 4 to 5 (#67) 2026-06-30 23:04:43 +00:00
labeler.yml ci: align workflows with Sea Haven CI/CD handbook (#29) 2026-06-27 21:47:25 -04:00
pr_lint.yml chore: update pr title lint to include dashboard and triage scopes (#170) 2026-07-10 13:05:51 -04:00
promote-to-main.yml ci: re-home prod promotion into a gated promote-to-main workflow (PR1: managed-LGC migration) (#63) 2026-06-29 19:49:21 -04:00
reviewer-eval.yml fix: align reviewer eval with published findings (#1713) 2026-07-16 17:36:18 -04:00
upstream-ledger-sync.yml ci: use GitHub App token to open upstream-ledger-sync PR (#165) 2026-07-10 12:25:59 -04:00