Execution attempt · A30 · rerun
claude-fable-5-1-claude-high: six-game smoke then full 460-game high-effort baseline using corrected denial audit; launch strictly after preceding arm verifies.
Pinned source and configuration
https://github.com/stw2/zendo-lab @ 105a353c9aa58ea4dae8f368cbca97e9072acd7d
Reference checked 2026-10-08 13:38 UTC. Later commits, branches or plan changes do not retarget this attempt.
- experiments/E04-claude-baselines/DESIGN.md
- experiments/E04-claude-baselines/scripts/claude_backend.py
- experiments/E04-claude-baselines/scripts/01_all.sh
- experiments/E04-claude-baselines/scripts/01_run.sh
- experiments/E04-claude-baselines/scripts/01_play.py
- experiments/E04-claude-baselines/scripts/04_denials.py
- Command
- bash experiments/E04-claude-baselines/scripts/01_run.sh claude-fable-5-1-claude-high
- Working directory
- .
- Configuration paths
- experiments/E04-claude-baselines/DESIGN.md
- Parameters
- {"reasoning_effort": "high", "thinking": "adaptive", "max_tokens": 128000, "sampling": "provider defaults", "manifest": "dev", "games": 460, "batch": 12, "player": "model", "backend": "function", "seed": "manifest episode seeds; model unseeded", "model": "claude-fable-5-1", "sdk_identity": "You are a Claude agent, built on Anthropic's Claude Agent SDK.", "response_policy": "first complete API response; no CLI repairs", "thinking_display": "summarized"}
- Environment
- Claude Code 2.1.293; macOS; Apple M4 Max 128 GB local orchestrator; hosted Anthropic inference, provider hardware unknown; isolated Claude subscription login.
- Output directory
- Not recorded
Inputs
Materials the registrant named when registering this attempt, by content identity; obtainability is derived from their location reports. Nothing is fetched or verified.
- ZendoBench 1.0.0 dev manifest · evaluation data · M1 zendo-bench-1.0.0-dev-manifest · dataset · Download
Delivered events
Registered
#1claude-fable-5-1-claude-high: six-game smoke then full 460-game high-effort baseline using corrected denial audit; launch strictly after preceding arm verifies.
Preflight failed before launch
#2A30 was never launched. Superseded before inference to pin provider-error retry classification and the dedicated OAuth config-lock permission; no model calls or measurements.
Error: superseded_source_before_launch
Report an event
The registrant’s capture tool normally delivers start and outcome events. Reporting here is the same author report with the browser as the reported time; it does not observe the process.
This attempt has a delivered outcome. A new execution is a new attempt.