 Garry TanandOpenAI Codex
|
4a3c6a8a3c
|
v1.87.0.0 feat: add verified CSO audits and replayable repair bundles (#2852)
* feat(cso): add verified audits and replayable repair bundles
* fix(cso): harden qualification and setup boundaries
* fix(cso): assemble security canaries at runtime
* fix(cso): bound release proof and maintenance work
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* fix(cso): require complete evaluation reports
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* fix(cso): replay expired snapshots from supplied source
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* test(cso): synchronize DNS cancellation assertion
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* chore(ship): exempt repository owner from liveness proof
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* test(cso): make recheck retention overlap deterministic
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* chore: bump version and changelog (v1.85.0.0)
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* fix(cso): pass native release gates
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* chore: move release to v1.86.0.0
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* fix(cso): resolve rechecks by finding
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* chore: move release to v1.87.0.0
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* fix(cso): pass macOS and Windows release gates
Normalize BSD wc output, compare Windows paths by filesystem identity, preserve portable snapshot race coverage, and narrow POSIX-only Windows fixtures.
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* fix(cso): harden native verification gates
* fix(cso): refine Windows native diagnostics
* test(cso): isolate Windows Git startup failure
* test(cso): stabilize Windows native diagnostics
* fix(cso): support hardened Git on Windows
* fix(cso): close final verification gaps
* test(cso): bound cold Docker fixture setup
* fix(cso): restore cross-platform free-suite gates
---------
Co-authored-by: OpenAI Codex <noreply@openai.com>
|
2026-09-14 15:14:58 -07:00 |
|
 Garry TanandOpenAI Codex
|
71f6048e8a
|
v1.84.1.0 fix: default Codex and Claude to frontier models (#2835)
* fix: default cross-model workflows to frontier models
* chore: bump version and changelog (v1.82.1.0)
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* fix: repair frontier eval budgets and workflow instructions
Preserve frontier models and quality thresholds while fixing truncated judge output, ordered section expansion, consent checks, QA scoring, and ship audit gates. Add regression coverage and refresh generated docs.
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* fix: resolve workflow gaps exposed by frontier evals
Clarify plan-review ordering and fallback modes, preserve deploy readiness gates, honor configured merge methods, correct benchmark and canary contracts, and restore vendored installs on setup failure. Cover recovery with real-shell regressions.
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* fix: use agent capture budgets for deploy evals
Multi-turn deploy and benchmark sessions were incorrectly limited to the single-call judge timeout. Use the existing capture tier and leave outer-test cleanup headroom, with a free policy regression test. Keep all behavioral assertions and frontier models unchanged.
Co-Authored-By: OpenAI Codex <noreply@openai.com>
* fix: clarify retro workflow and evaluate compare instructions
Include compare mode in the frontier judge excerpt, define metric sources and snapshot ordering, and preserve the existing prompt-size budget.
Co-authored-by: OpenAI Codex <noreply@openai.com>
* fix: make documentation release review and publication consistent
Review before commit, clarify changelog safeguards and unavailable reviewer modes, and preserve raw PR bodies across separate shell calls. Keep title sync in one shell and add regression coverage.
Co-authored-by: OpenAI Codex <noreply@openai.com>
---------
Co-authored-by: OpenAI Codex <noreply@openai.com>
|
2026-09-09 08:57:21 -07:00 |
|