40 tasks. August 4, 2026 closed at 14.0x weighted leverage across 786.5 human-equivalent hours in 3,362 minutes of wall-clock time. Supervisory leverage came in at 386.8x.
That is 19.7 weeks of human-equivalent throughput in 56.0 hours. The ceiling was 62.8x; the floor was 1.7x. 27 of the 40 entries came from a single project.
Task Log
| # | Task | Human Est. | Claude | Sup. | Factor |
|---|---|---|---|---|---|
| 1 | Documentation | 68.0h | 65m | 2m | 62.8x |
| 2 | Audit and review | 28.0h | 30m | 3m | 56.0x |
| 3 | Documentation | 32.0h | 40m | 5m | 48.0x |
| 4 | Infrastructure | 28.0h | 38m | 2m | 44.2x |
| 5 | Create /fable-planner skill - Fable-gated program planner producing four docs and phased Opus/Sonnet execution plans | 2.5h | 4m | 5m | 37.5x |
| 6 | Infrastructure | 32.0h | 63m | 3m | 30.5x |
| 7 | Deployment | 26.0h | 68m | 3m | 22.9x |
| 8 | Fable planning: model-scorecard program (research + four docs x3 repos + phased plan + state + Fable review gate) | 26.0h | 70m | 2m | 22.3x |
| 9 | Documentation | 34.0h | 96m | 6m | 21.2x |
| 10 | Documentation | 14.0h | 42m | 3m | 20.0x |
| 11 | Testing | 16.0h | 50m | 5m | 19.2x |
| 12 | Model-Scorecard program: P0 baseline + P1.T1 benchmarks package + P4 phase-scoped routing live on the daemon + 5 production defects found/fixed with regression tests + 2 real benchmark runs | 130.0h | 445m | 12m | 17.5x |
| 13 | Testing | 16.0h | 57m | 2m | 16.8x |
| 14 | 13-site UI/UX audit remediation: canon-driven facts fleet-wide, dropdown a11y rebuild, responsive tables, CTA/posture truthing, [unreleased product] shell migration, GED+Certs compact pilots, Playwright+axe… | 100.0h | 362m | 3m | 16.6x |
| 15 | iOS plan approval pass: decisions D1-D5 recorded, WP rewrites, STATE+ledger scaffolding, ASC runbook, labs-deferral card | 2.5h | 10m | 2m | 15.0x |
| 16 | iOS productization steer: D9 per-product builds + icon family, D10 price parity, WP-2.0/4.0 added, ASC runbook rewrite | 2.0h | 8m | 2m | 15.0x |
| 17 | Coding | 7.0h | 28m | 2m | 15.0x |
| 18 | Electron: canon-driven LaunchPromoBanner + 7 stale-content fixes across 2 surfaces, 33 tests (A-ELEC Opus agent) | 8.0h | 33m | 2m | 14.5x |
| 19 | Phase-2 content repair: R4 complete, dedupe 14665->5265, deictic stems -26%/explanations -41%, math markup pushed, 2 transform bugs + 1 platform bug caught pre-push, [cost]spend (R-EXEC Opus agent) | 26.0h | 115m | 2m | 13.6x |
| 20 | avian-content-checks: JudgeClient/BudgetMeter thread safety + gate_report schema-pin test (A-CHECKS Opus agent) | 5.0h | 24m | 2m | 12.5x |
| 21 | run-smoke self-heal + smoke-user deletion forensics (audit_logs empty — actor unidentifiable) + 3 follow-up cards (A-SMOKE Opus agent) | 6.5h | 32m | 2m | 12.2x |
| 22 | [engine subsystem]: benchmark spend->ledger + findings current-version default + hub key-scrub done right (kept benchmark keys, filed doc-contradiction card) (A-[engine subsystem] Opus agent) | 9.0h | 45m | 2m | 12.0x |
| 23 | iOS plan: WP-1.4 src/ reorganization added (sequencing, registry/tooling path scope, exit gates) | 1.0h | 5m | 1m | 12.0x |
| 24 | Testing | 1.5h | 8m | 2m | 11.2x |
| 25 | iOS parity WP-0.2: catalog stale-while-revalidate (disk cache, explicit load state, skeleton/error+Retry), host out of Swift into xcconfig, Associated Domains repointed; 20 tests | 12.0h | 64m | 4m | 11.2x |
| 26 | [ip-cluster] first-3 cards: [engine subsystem] judge-role swap + regression guard; ANN debt cleared in 2 test files; [engine subsystem] CLAUDE.md benchmark-keys exception; board-state corrected (56/81 already… | 5.0h | 27m | 1m | 11.1x |
| 27 | iOS parity WP-4.0: environment-scoped App Store Server Notifications V2 (per-env SignedDataVerifier + data-layer guard), 66 tests incl. real x5c chain | 10.0h | 58m | 4m | 10.3x |
| 28 | Testing | 18.0h | 105m | 6m | 10.3x |
| 29 | Phase-2 drain: dedupe 14665->13, deictic stems -58%, explanations -89%, ISACA matcher root-cause corrected + husks repaired, closing audit locked, 3 findings filed (R-FINISH Opus agent) | 32.0h | 190m | 2m | 10.1x |
| 30 | Design and frontend | 12.0h | 85m | 3m | 8.5x |
| 31 | iOS parity WP-0.3: anonymous public domain catalog in avian-api (optional-user dep, public_only forced for anonymous) + iOS OIDC retirement note; live-probed electron/[unreleased product] prod sign-in breakage | 4.0h | 33m | 4m | 7.3x |
| 32 | Audit and review | 5.0h | 43m | 2m | 7.0x |
| 33 | Debugging | 5.0h | 46m | 2m | 6.5x |
| 34 | Debugging | 14.0h | 142m | 4m | 5.9x |
| 35 | Deployment | 10.0h | 105m | 2m | 5.7x |
| 36 | Testing | 5.0h | 59m | 3m | 5.1x |
| 37 | Engine: unmapped question-type WARNING + evidence-based code-scenarios->mcq mapping, fleet re-audit clean (A-ENG Opus agent) | 3.5h | 43m | 2m | 4.9x |
| 38 | Debugging | 14.0h | 240m | 2m | 3.5x |
| 39 | iOS parity WP-0.6: prod-API posture + launch-time unreachable-gateway banner (3s Task-cancellation deadline, dedicated session, offline/unreachable mutual exclusion); root-caused an XCUIAccessibilityAudit… | 10.0h | 175m | 4m | 3.4x |
| 40 | Product-listing regroup across 14 sites: Available Now / Free Forever / Coming Soon (launch-date order) on shared products page + gateway router, August→September launch slip via canon, full-family redeploy +… | 6.0h | 209m | 2m | 1.7x |
Aggregate Statistics
| Metric | Value |
|---|---|
| Total tasks | 40 |
| Total human-equivalent hours | 786.5 |
| Total Claude minutes | 3,362 |
| Total supervisory minutes | 122 |
| Total tokens | 28,090,247 |
| Weighted average leverage factor | 14.0x |
| Weighted average supervisory leverage factor | 386.8x |
| Human-equivalent weeks | 19.7 |
Analysis
The highest factor of the day came in at 62.8x and the lowest at 1.7x, a spread of 36.4 times between the two. That is a wide range for a single day, and it usually means the day mixed mechanical work with work that needed real judgement.
The largest single entry accounted for 68.0 of the 786.5 human-equivalent hours, or 9 percent of the day. No single task dominated the total, so the weighted average is representative.
Supervisory time was 122 minutes against 3,362 minutes of execution, a ratio of about 1 to 28. Supervisory leverage of 386.8x is the figure I find most honest, because it measures the hours I actually spent rather than the hours a machine spent on my behalf.
Every figure here is recorded at the time the work is done rather than reconstructed afterwards. The human estimate is my own judgement and carries the uncertainty that implies; the minutes and tokens are measured. The full dataset, including this day, is available for download.