feat(xreview-scout-codex): run on gpt-6.1-sol at high reasoning effort - #469
Conversation
The scout pinned gpt-5.6-sol. gpt-6.1-sol is now the Codex CLI's default model, released 2026-09-29. The scout also set no reasoning effort, so it ran at the harness default. It now sets `executor.reasoning_effort: high`. The deployed server (omnigent 0.16.0.dev0) reads `executor.reasoning_effort`. The bundle check in CI pins omnigent 0.9.0, which parses the bundle but drops the field, so CI does not assert it. Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
PR SummaryMedium Risk Overview CI is aligned with the server parser: Reviewed by Cursor Bugbot for commit fd4cefb. Bugbot is set up for automated code reviews on this repo. Configure here. |
There was a problem hiding this comment.
Moves the codex scout from gpt-5.6-sol to gpt-6.1-sol and sets executor.reasoning_effort: high. Nothing blocks: codex's one finding, that gpt-6.1-sol is not a supported model id, was dropped because OpenAI released that model on 2026-09-29 and Codex CLI 0.159.1 made it the default; codex raised nothing else, and the only open risk is the effort field, which nothing checks before deploy.
Non-blocking
1 finding on the changed lines, as inline comments.
1 nit, not posted on the code
agents/xreview-scout-codex/config.yaml:31— The new reason ("the Codex CLI's default model") doesn't fit this bundle well: it runs thecodexharness, and line 42 says the image carries no Codex CLI. The removed reason, matching the model ai-review's codex pass resolves to, applied directly to this bundle. If that is still the goal, say so here.
seidroid review · decision approve · session 9ebbe45d0f8b408199ec3249d21e48ff · turn resp_claude_b6fb989970bc0366ac5a17513858c035 · item b9fa60044aed5e6d93d9e998958c793f
Findings: 0 blocking | 1 non-blocking | 1 posted inline
…effort The bundle check pinned omnigent 0.9.0, which parses a bundle but drops executor.reasoning_effort. It now pins 0.16.0, the release the deployed server tracks, and fails if xreview-scout-codex does not read effort high. Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
There was a problem hiding this comment.
Cursor Bugbot has reviewed your changes using default effort and found 1 potential issue.
❌ Bugbot Autofix is OFF. To automatically fix reported issues with cloud agents, enable autofix in the Cursor dashboard.
Reviewed by Cursor Bugbot for commit 9267267. Configure here.
|
@seidroid review |
There was a problem hiding this comment.
Since my last review, the bundle check now runs on omnigent 0.16.0 and asserts reasoning_effort == 'high' for the codex scout. Running the check on 0.16.0 passes as committed and fails once the field is removed, which resolves my open thread, so nothing blocks; codex's only finding, the model-id claim, is dropped again because OpenAI released gpt-6.1-sol on 2026-09-29 and Codex CLI 0.159.1 made it the default.
1 nit, not posted on the code
agents/xreview-scout-codex/config.yaml:31— The new reason ("the Codex CLI's default model") doesn't fit this bundle well: it runs thecodexharness, and line 42 says the image carries no Codex CLI. The removed reason, matching the model ai-review's codex pass resolves to, applied directly to this bundle. If that is still the goal, say so here.
seidroid review · decision approve · session 9ebbe45d0f8b408199ec3249d21e48ff · turn resp_claude_a6883a6d3220e6a05fd6858a54c83843 · item 537bd76159bc5556bc8c12ef7fd35519
Findings: 0 blocking | 0 non-blocking | 0 posted inline
The comment named the Codex CLI's default as the reason, but this bundle runs the codex harness without the Codex CLI. Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
|
@seidroid review |
There was a problem hiding this comment.
The only change since my last review rewords the model comment in the codex scout config. It no longer leans on the Codex CLI default (my earlier nit) and now gives a reason that fits this bundle; the workflow assertion I checked last time is unchanged, so nothing blocks. Codex's only finding, that gpt-6.1-sol is not a valid model id, is dropped for the third time: OpenAI released gpt-6.1-sol on 2026-09-29 and it is a supported Codex model.
seidroid review · decision approve · session 9ebbe45d0f8b408199ec3249d21e48ff · turn resp_claude_0060ab2fc60494bb7fd51db2b6b08d09 · item a84cab1600bd5ce396b689e66b92d787
Findings: 0 blocking | 0 non-blocking | 0 posted inline
The PR gate's EXPECT map now requires effort high for xreview-scout-codex, but the publish gate's map did not, so a later drop of reasoning_effort could pass the in-image parse that production uses. Both gates now assert it. Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
|
@seidroid review |
There was a problem hiding this comment.
New since my last review: ecr-server.yml's deploy gate now also asserts reasoning_effort == 'high' for the codex scout, matching the PR gate. Nothing blocks; of codex's two findings, the model-id claim is dropped again (OpenAI released gpt-6.1-sol on 2026-09-29 and it is a supported Codex model), and its low-severity point on ecr-server.yml:128 holds and is kept below as a nit.
1 nit, not posted on the code
.github/workflows/ecr-server.yml:128— (Raised by codex.) The comment at line 100 says None means a bundle sets no effort, butwant_effort is not Noneskips the check for those bundles, so an effort added to any of them would still pass. Either assert None for them too, or reword the comment to say None means not checked.
seidroid review · decision approve · session 9ebbe45d0f8b408199ec3249d21e48ff · turn resp_claude_6c1d379ffa8ea3852bd048b782813e97 · item ac7352086a92598484bee4ab0c137f5d
Findings: 0 blocking | 0 non-blocking | 0 posted inline
Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
|
@seidroid review |
There was a problem hiding this comment.
The only change since my last review rewords the ecr-server.yml comment to say None skips the effort check, which settles the nit I carried from codex; nothing else moved, and nothing blocks. Codex's one finding this round, that gpt-6.1-sol is not a supported model id, is dropped again because OpenAI released gpt-6.1-sol on 2026-09-29 and it is a supported Codex model.
seidroid review · decision approve · session 9ebbe45d0f8b408199ec3249d21e48ff · turn resp_claude_78d76ba04dcf554df7e7c79b34be8ed0 · item 6057bf25d1d55835a773a6984cf52f17
Findings: 0 blocking | 0 non-blocking | 0 posted inline

Summary
The codex scout moves from
gpt-5.6-soltogpt-6.1-soland setsexecutor.reasoning_effort: high. Before this, it set no effort and ran at the harness default.Why this model
gpt-6.1-sol(released 2026-09-29) is the Codex CLI's default model as of v0.159.1, and OpenAI describes it as "near-Astra performance for complex work at a lower cost". The comment that tied the pin toopenai/codex-action's old default is replaced.Verification
omnigent==0.9.0:okfor all 4 bundles.0.16.0.dev0, which hasExecutorSpec.reasoning_effort. CI's pinned0.9.0parses the bundle but drops the field, so CI does not assert it.gpt-6.1-sol. The first review after deploy settles both: its log must showscout reported, not a scout failure note.Rollout
Ships on the next bundle deploy. Independent of the driver PRs.
🤖 Generated with Claude Code