Add Harvey LAB arm-B runner and pilot task selector

↗ view on GitHub · Eli Ziff · 2026-07-28 · 59769d6c

lab-beaver-arm.ts drives the real /chat transport (anonymous local mode,
same recipe as eval-beaver-arm.ts) on a LAB task and writes LAB-layout
results so harvey-labs evaluation.run_eval judges both arms identically.
lab-select-tasks.ts picks a seeded, practice-area-stratified pilot set
restricted to deliverables the beaver arm can export (docx/md/txt).

Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_016yXxuitxX5h4wm1v88hnga
Repository eliziff/Beaver
Author Eli Ziff <eliasziff@gmail.com>
Authored
Parents 3009c9bd
Stats 2 files changed , +404
Part of Evaluation harness: Beaver-CAN and LegalBench-RAG adapters

Capture this commit into my fork

Download a Markdown prompt that tells Claude how to port this exact commit into your working tree. Run it via claude -p < capture-commit-59769d6c.md from inside the repo you want the change in.

⬇ Download capture-commit-59769d6c.md