mirror of
https://github.com/priyanshujain/messageboardbench.git
synced 2026-10-02 19:17:07 +00:00
Experiment evidence
Start with EXPERIMENT.md. These dated bundles preserve completed analyses and supporting artifacts; they are not alternative current plans.
| Bundle | Use |
|---|---|
| board-muse-sept8 | Full matched Muse Contributor replication: one attributed adoption, one possible unattributed adoption; all eight gaming artifacts validated. |
| board-interface-v2-sept8 | Latest matched GLM pilot: one attributed cross-task adoption after peer receipt. |
| board-pilot-sept8 | Prior interface: publication but zero board reads; task validation and SWE readiness. |
| model-comparison-sept7 | Tiny GLM/Muse Contributor feasibility comparison, not model-level rates. |
| token-comparison-sept7 | Frozen historical token audit and recovered baseline accounting. |
| team-pilot-sept7 | Historical shared-directory experiment with different prompts and limits. |
| sept10-revision | Historical baseline audit supporting the revision. |
Raw Inspect runs remain in ../logs/ (local, Git-ignored). Migrated reports,
reviews, probe outputs and exports are byte-preserved. Embedded absolute paths
record their original locations; resolve them with
the migration map. Historical
scripts inside bundles are evidence, not the current run entry points. Use
just board, scripts/board_report.py, or the portable offline tools in
scripts/analysis/ for new work, always with a fresh output directory.