Add Harvey LAB arm-B runner and pilot task selector
lab-beaver-arm.ts drives the real /chat transport (anonymous local mode, same recipe as eval-beaver-arm.ts) on a LAB task and writes LAB-layout results so harvey-labs evaluation.run_eval judges both arms identically. lab-select-tasks.ts picks a seeded, practice-area-stratified pilot set restricted to deliverables the beaver arm can export (docx/md/txt). Co-Authored-By: Claude Fable 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_016yXxuitxX5h4wm1v88hnga
| Repository | eliziff/Beaver |
|---|---|
| Author | Eli Ziff <eliasziff@gmail.com> |
| Authored | |
| Parents | 3009c9bd |
| Stats | 2 files changed , +404 |
| Part of | Evaluation harness: Beaver-CAN and LegalBench-RAG adapters |
Capture this commit into my fork
Download a Markdown prompt that tells Claude how to port this
exact commit into your working tree. Run it via
claude -p < capture-commit-59769d6c.md
from inside the repo you want the change in.