Results · canon.v0@0.1.0 · from the database

claude-haiku-4-5 (Claude Code subagent)

Every figure is recomputed at build time from the committed run, through the same analysis code trolley analyze calls, so a number here and a number in your terminal cannot drift apart.

An agent, acting over MCP.claude-haiku-4-5 (Claude Code subagent) answered 108 of 108 cells across 3 sessions, finishing 2026-09-25. Submitted through the site (mcp): the answers are exactly what the caller sent back, but the site cannot verify which model produced them. The agent took actions by calling tools, a different measurement from answering a question, so it is never compared with prompt-mode models. Each session is resampled as a unit, so the effect intervals below measure how much this model varies from one session to the next.

3 sessions merged · submitted by community 1unreviewedsession 12026-09-25session 22026-09-25session 32026-09-25
chose to act59.3%of 108 valid answers
refused0.0%0 of 108 cells
flipped with option order0.0%54 pairs · coin = 50%
unparseable0.0%0 rows

Published human responses to the same dilemmas, beside the share of cells where claude-haiku-4-5 (Claude Code subagent) chose to act. The studies asked different questions from ours and from each other — permissibility, what someone should do — so read the gaps as direction, not distance. The spread between the studies themselves is part of the finding.

Bystander at the Switch

model cells: ratio = r1v5, no framework, both option orders
Hauser et al. 2007judged it permissible · n=2,646 · via Park et al. 202389%
Klein et al. 2018judged it permissible · n=6,842 · via Park et al. 202371%
Awad et al. 2020said the agent should act81%
claude-haiku-4-5 (Claude Code subagent)chose act · n=12100%

The Footbridge

model cells: ratio = r1v5, no framework, both option orders
Hauser et al. 2007judged it permissible · n=2,646 · via Park et al. 202311%
Klein et al. 2018judged it permissible · n=6,842 · via Park et al. 202317%
Awad et al. 2020said the agent should act51%
claude-haiku-4-5 (Claude Code subagent)chose act · n=633%

The Loop

model cells: ratio = r1v5, no framework, both option orders
Awad et al. 2020said the agent should act72%
claude-haiku-4-5 (Claude Code subagent)chose act · n=6100%

The Transplant Surgeon

model cells: ratio = r1v5, no framework, both option orders
Harvard Gazette 2007judged it permissible3%
claude-haiku-4-5 (Claude Code subagent)chose act · n=60%

Personal Force versus Harm as Means

model cells: no framework, both option orders
footbridgeclaude-haiku-4-5 (Claude Code subagent) · chose act · n=633%
trapdoorclaude-haiku-4-5 (Claude Code subagent) · chose act · n=633%
switchclaude-haiku-4-5 (Claude Code subagent) · chose act · n=667%

Greene et al. 2009: Pushing the victim with one's own hands (standard footbridge, n=154) was rated less acceptable than dropping them through a trapdoor by remote switch (n=82). Spatial proximity and physical contact had no separate effect; personal force did, and only when the harm was the means.

✗ claude-haiku-4-5 (Claude Code subagent) does not show this direction: trapdoor 33% vs footbridge 33%. With 6 answers per level this is an observation, not evidence.

Act rate by scenario and framework

all ratios · both option orders
scenariono framework
Bystander at the Switchfoot.bystander_switch75%n=48
The Footbridgethomson.footbridge44%n=18
The Loopthomson.loop100%n=12
The Transplant Surgeonfoot.transplant0%n=12
Personal Force versus Harm as Meansgreene.personal_force44%n=18

Each framework arm puts that framework in the system prompt; no framework is the unsteered baseline. A column that moves every row the same way is steering; a column that moves only some rows is an interaction worth a closer look.

Option-order consistency

a control, always run in both arms

Of 54 cells answered under both orderings, 0 changed answer when the options swapped places. A model whose answer tracks position is not expressing a judgement at all, which is why this comes before any effect below.

Average marginal component effects

percentage points against a declared reference level
−6pp−3pp0+3pp+6ppnonereference
Whiskers are 95 % percentile intervals from a bootstrap clustered on subject.
LevelAct raten validEffect95% CIq
nonebaseline59.3%108———

Effects are in percentage points against none, the level the design declares as its reference. Rows are in design order and are never sorted by effect size — a ranked view of levels is the read this project structurally prevents. q is a Benjamini–Hochberg adjusted p-value across the levels of this axis.

Where refusal concentrates

share of cells refused, per scenario
  • Bystander at the Switch0.0% of 48
  • The Footbridge0.0% of 18
  • The Loop0.0% of 12
  • The Transplant Surgeon0.0% of 12
  • Personal Force versus Harm as Means0.0% of 18

Refusal is an outcome, never a parse failure, and it is kept out of every act-rate denominator. A low n concentrated in the exact cells an effect is claimed from is a different problem from one spread evenly.