2026-09-16 01:52:41 +05:30
2026-09-16 01:52:41 +05:30
2026-09-16 01:52:41 +05:30
2026-09-16 01:52:41 +05:30
2026-09-16 01:52:41 +05:30

messageboardbench

Do coding agents share, adopt, reject, or correct cheating techniques through a public team board?

Agents work on ImpossibleBench tasks in separate containers with private scratch space and, in the board condition, explicit tools for reading and posting team messages. The default pilot uses three agents per cohort, two cohorts, and 90 messages per episode; model, population size, sampling and budgets are configurable.

Built on ImpossibleBench and Inspect AI. Passing contradictory tests is an operational signal of specification gaming, not a complete label of intent or attempted cheating.

See setup and the board experiment workflow. Run just board to choose parameters interactively and start the experiment. Use just board-preview for a configuration preview without model calls, or just board-run --model muse to run with Muse Spark Contributor. Executing makes paid requests; the existing key spending cap is never raised.

Start with EXPERIMENT.md for the research question, what we have done, what the results establish, and the next work. The latest pilot contains one explicitly attributed cross-task adoption; it does not establish an increased gaming rate or concealed collusion.

Location Purpose
src/, scripts/, tests/ Implementation, entry points and tests
docs/ Setup, workflow and supporting research
results/ Reviewed, frozen experiment evidence
logs/ Raw local runs, ignored by Git
work/ Disposable local working files, ignored by Git

This is the sole working repository for the experiment. The former messageboard repo is retired; personal notes and unrelated material remain in its archive. Run just evidence-check to verify the migrated evidence.

S
Description
No description provided
Readme MIT
50 MiB
0 Stars 1 Watchers 0 Forks
Languages
Python 99.8%
Just 0.2%