Skip to content

Rubric v3 and the repeated Pro run: decision rule not met - #29

Merged
TheCryptoDonkey merged 6 commits into
mainfrom
docs/rubric-v3-and-pro-run
Sep 23, 2026
Merged

TheCryptoDonkey merged 6 commits into
mainfrom
docs/rubric-v3-and-pro-run

Conversation

@TheCryptoDonkey

Copy link
Copy Markdown
Member
  • Rubric version 3 (docs/experiments/rubric-v3-20260923/): required points limited to what each prompt asks.
  • Repeated Pro run protocol, driver and summariser, with two stopped attempts recorded; the summariser was committed before the first result.
  • Results (docs/experiments/repeated-pro-20260923/RESULTS.md): 72 cells on DeepSeek V4 Pro. Context used more executor input than plain (geometric mean ratio 1.17, 90% interval 0.79 to 1.36) and Graphify (1.31, 0.94 to 1.64), and accepted 3 of 8 tasks against 5 for each. The decision rule is not met; no saving is claimed. Input follows turns: Context returns fewer tool bytes but takes more turns.
  • run.mjs records the rotated arm position with --order-index.

🤖 Generated with Claude Code

https://claude.ai/code/session_01KzXbuBSq88i3Re82qt6c4L

@TheCryptoDonkey
TheCryptoDonkey merged commit 1520303 into main Sep 23, 2026
1 check passed
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant