ouroboros/devtools/benchmarks/programbench/settings_base.json
Ouroboros af3523466d v7next F3.1-D ABI-10: remove the reviewer comma-list migration read (owner 5.4=A)
The structured OUROBOROS_REVIEWER_SLOTS is the ONE reviewer configuration
surface. The legacy migration-read block (reviewer_slot_config.py:
_shared_session_route_spec / _legacy_rows / _legacy_config — comma-lists,
phase-5 per-row/advisory route envs, global-effort copying) is DELETED; with
no structured value the loader serves the SHIPPED DEFAULT PANEL: api_chat
triad/scope rows over get_review_models()/get_scope_review_models() (the
derived env plane — identical models to the old legacy defaults on every
config class, incl. single-direct-provider adaptation and a bench launcher's
env override), default advisory row, source="default", historical positional
slot ids (receipts keep lining up) and the unchanged legacy fingerprint
identity for the unconfigured panel (skill-review replay authority survives
the upgrade).

Settings vocabulary: OUROBOROS_REVIEW_MODELS / OUROBOROS_SCOPE_REVIEW_MODELS /
OUROBOROS_SCOPE_REVIEW_MODEL plus the phase-5 route envs
(OUROBOROS_REVIEW_ROUTES / OUROBOROS_SCOPE_REVIEW_ROUTES /
OUROBOROS_ADVISORY_REVIEW_ROUTE) leave SETTINGS_DEFAULTS and join
RETIRED_SETTING_KEYS (ghost purge on load; an install that configured
reviewers only through them gets the default panel — the RC auditor names
this migration, per plan). The derived runtime projection
(project_reviewer_slots_into_env) STAYS, floored from
OPENROUTER_REVIEW_DEFAULTS; get_review_models/get_scope_review_models keep
serving the API-pinned surfaces from that env plane. server_runtime's
provider-defaults migration now normalizes a ghost comma value it is FED but
never INTRODUCES a retired key (the direct-provider review adaptation lives
on the read side); the singular→plural promotion in model_slots is gone
(purged before it could run). gateway/settings candidate probes read the
structured candidate else the live derived config.

Bench templates migrated to structured slots with the SAME models
(continual_learning, gaia, swe_bench_pro base/example/probe/profile; comma
keys dropped everywhere incl. osworld/cybergym/programbench which already
carried structured values); run_tb metadata defaults now come from
OPENROUTER_REVIEW_DEFAULTS; manifests record OUROBOROS_REVIEWER_SLOTS.

tests/test_comma_list_sweep.py is the phase CI gate (ABI-10 hook): no
migration-read branches, retired vocabulary pinned, bench templates clean,
prose rewritten, derived projection alive. ARCHITECTURE deltas same commit.

(cherry picked from commit e7578133758f01eac584e8e6522b3f03c4d09818)
2026-08-31 16:54:17 +00:00

29 lines
2.2 KiB
JSON

{
"OPENROUTER_API_KEY": "",
"OPENAI_API_KEY": "",
"ANTHROPIC_API_KEY": "",
"OUROBOROS_MODEL": "openai/gpt-5.5",
"OUROBOROS_SUBAGENTS": "{\"enabled\":true,\"items\":[{\"subagent_id\":\"benchmark-model\",\"recommended_use\":\"Use for substantial implementation, difficult debugging, research synthesis, and end-to-end ownership. Prefer when overall quality and continuity matter more than speed.\",\"route\":{\"kind\":\"api_model\",\"target_id\":\"openai/gpt-5.5\"}}]}",
"OUROBOROS_MODEL_LIGHT": "openai/gpt-5.5",
"OUROBOROS_MODEL_VISION": "openai/gpt-5.5",
"OUROBOROS_MODEL_CONSCIOUSNESS": "openai/gpt-5.5",
"OUROBOROS_MODEL_FALLBACKS": "openai/gpt-5.5",
"OUROBOROS_MODEL_DEEP_SELF_REVIEW": "openai/gpt-5.5",
"USE_LOCAL_MAIN": false,
"USE_LOCAL_LIGHT": false,
"USE_LOCAL_FALLBACK": false,
"USE_LOCAL_CONSCIOUSNESS": false,
"OUROBOROS_WEBSEARCH_MODEL": "openai/gpt-5.5",
"OUROBOROS_REVIEWER_SLOTS": "{\"triad\":[{\"slot_id\":\"benchmark-triad-1\",\"route\":{\"kind\":\"api_chat\",\"target_id\":\"openai/gpt-5.5\"},\"effort\":\"medium\"},{\"slot_id\":\"benchmark-triad-2\",\"route\":{\"kind\":\"api_chat\",\"target_id\":\"openai/gpt-5.5\"},\"effort\":\"medium\"},{\"slot_id\":\"benchmark-triad-3\",\"route\":{\"kind\":\"api_chat\",\"target_id\":\"openai/gpt-5.5\"},\"effort\":\"medium\"}],\"scope\":[{\"slot_id\":\"benchmark-scope-1\",\"route\":{\"kind\":\"api_chat\",\"target_id\":\"openai/gpt-5.5\"},\"effort\":\"high\"},{\"slot_id\":\"benchmark-scope-2\",\"route\":{\"kind\":\"api_chat\",\"target_id\":\"openai/gpt-5.5\"},\"effort\":\"high\"},{\"slot_id\":\"benchmark-scope-3\",\"route\":{\"kind\":\"api_chat\",\"target_id\":\"openai/gpt-5.5\"},\"effort\":\"high\"}],\"advisory\":{\"enabled\":false,\"route\":{\"kind\":\"api_chat\",\"target_id\":\"\"},\"effort\":\"low\"}}",
"OUROBOROS_TASK_REVIEW_MODE": "required",
"OUROBOROS_REVIEW_ENFORCEMENT": "blocking",
"OUROBOROS_REVIEW_MAX_CYCLES": "unlimited",
"OUROBOROS_EFFORT_TASK": "high",
"OUROBOROS_EFFORT_REVIEW": "medium",
"OUROBOROS_EFFORT_SCOPE_REVIEW": "high",
"OUROBOROS_POST_TASK_EVOLUTION": "false",
"OUROBOROS_RUNTIME_MODE": "pro",
"OUROBOROS_SAFETY_MODE": "light",
"OUROBOROS_MAX_WORKERS": 4,
"OUROBOROS_MAX_ROUNDS": 1000
}