29 tasks. July 11, 2026 closed at 21.0x weighted leverage across 762.0 human-equivalent hours in 2,172 minutes of wall-clock time. Supervisory leverage came in at 205.9x.
That is 19.1 weeks of human-equivalent throughput in 36.2 hours. The ceiling was 60.0x; the floor was 5.5x. 24 of the 29 entries came from a single project.
Task Log
| # | Task | Human Est. | Claude | Sup. | Factor |
|---|---|---|---|---|---|
| 1 | Deployment | 90.0h | 90m | 8m | 60.0x |
| 2 | Design and frontend | 140.0h | 140m | 8m | 60.0x |
| 3 | TTU Biology + Anatomy & Physiology full-year coverage analysis (2 courses, 4 syllabi PDFs) vs live AVIAN domains, mirroring CHEM1B playbook | 32.0h | 33m | 8m | 58.2x |
| 4 | TTU Biology + Anatomy and Physiology full-year coverage analysis vs live AVIAN domains | 32.0h | 33m | 8m | 58.2x |
| 5 | Temporal migration design doc for [engine subsystem]/[engine subsystem] job substrate (3-way NATS/Temporal/Procrastinate research pivoted mid-task into a full Temporal migration plan per Cooper's decision:… | 20.0h | 30m | 12m | 40.0x |
| 6 | Temporal migration design doc ([engine subsystem]/[engine subsystem] job substrate) | 20.0h | 30m | 12m | 40.0x |
| 7 | Deployment | 44.0h | 100m | 4m | 26.4x |
| 8 | Judge-in-the-loop regeneration gate for [engine subsystem]-service action.regen_items (checks-lib mcq.judge.option_truth per-option judge + independent confirm adjudication + node_ambiguity/generation_quality… | 14.0h | 40m | 8m | 21.0x |
| 9 | Judge-in-the-loop regeneration gate ([engine subsystem] regen_items: routed judge + v4-pro confirm in retry loop, unrepairable classification) | 14.0h | 40m | 8m | 21.0x |
| 10 | Temporal migration Phase 1 for AVIAN [content generation] job layer: deploy integrated Temporal server into the shared-Postgres supporting-services stack (server+UI, namespace/720h retention/ports 7233+8080,… | 26.0h | 82m | 6m | 19.0x |
| 11 | Documentation | 34.0h | 110m | 18m | 18.5x |
| 12 | Infrastructure | 80.0h | 270m | 15m | 17.8x |
| 13 | Deployment | 48.0h | 175m | 3m | 16.5x |
| 14 | Bio Fundamentals + Anatomy&Physiology [engine subsystem]: synthesize+validate+ship v5 (blood unit / plant systems / succession / biomolecules) | 24.0h | 90m | 10m | 16.0x |
| 15 | Fix systemic [engine subsystem]-service [engine subsystem]/resume [content generation] hang: scope resume validation to PENDING nodes (Defect B) + hard per-node timeout bound on Pass-3 adversarial evaluation… | 14.0h | 55m | 8m | 15.3x |
| 16 | Fix [engine subsystem]-service [engine subsystem]/resume [content generation] hang (scoped validation + hard per-node timeout) | 14.0h | 55m | 8m | 15.3x |
| 17 | [engine subsystem] coverage baseline + latency integration test tier + 11 pre-existing test fixes | 14.0h | 60m | 8m | 14.0x |
| 18 | Except-stem uniqueness generation guard (mcq.gen.except_uniqueness) in [engine subsystem]-service: new [content generation]/questions/except_uniqueness.py reusing the mcq.judge.option_truth per-option judge… | 6.0h | 27m | 8m | 13.3x |
| 19 | Except-stem uniqueness generation guard (mcq.gen.except_uniqueness) in [engine subsystem]-service | 6.0h | 27m | 8m | 13.3x |
| 20 | SAA-C03 node-edit pilot: 125-item diagnosis via 5 sub-agents + dispatch-bug discovery + judge-gated regen (61% conversion) + findings doc | 28.0h | 160m | 10m | 10.5x |
| 21 | Close CHEM1B hardlink-corruption hazard in [engine subsystem]-service: hard refusal + isolated-build-then-finalize default across 11 [content generation]/pipeline call sites, 29 new tests, CLAUDE.md doc;… | 14.0h | 100m | 4m | 8.4x |
| 22 | Close CHEM1B hardlink-corruption hazard: hard refusal + isolated-build default ([engine subsystem]-service) | 14.0h | 100m | 4m | 8.4x |
| 23 | Testing | 6.0h | 50m | 3m | 7.2x |
| 24 | Debugging | 6.0h | 50m | 3m | 7.2x |
| 25 | Build isolated Temporal SPIKE proving worker-heartbeat self-heal (kill-worker auto-resume) for AVIAN [content generation] layer | 6.0h | 55m | 8m | 6.5x |
| 26 | Root-cause + fix live regression: build_real_action_llm hardcoded pool_size=1 (discarding config.llm_pool_size) broke QuestionValidator.map_structured for [engine subsystem]-service action.regen_items… | 3.0h | 30m | 5m | 6.0x |
| 27 | Root-cause + fix build_real_action_llm pool_size regression (hardcoded pool_size=1 predating regen kinds; regression tests on real construction seam) | 3.0h | 30m | 5m | 6.0x |
| 28 | Root-cause + fix live no-op regression: regen_items exhausted-retry logic silently discarded judge/confirm-confirmed defect evidence whenever a later attempt failed for an unrelated gate reason; reproduced… | 5.0h | 55m | 6m | 5.5x |
| 29 | Root-cause + fix regen evidence-loss no-op (mixed-retry case; real-package repro; seam tests through serialize chain) | 5.0h | 55m | 6m | 5.5x |
Aggregate Statistics
| Metric | Value |
|---|---|
| Total tasks | 29 |
| Total human-equivalent hours | 762.0 |
| Total Claude minutes | 2,172 |
| Total supervisory minutes | 222 |
| Total tokens | 21,015,000 |
| Weighted average leverage factor | 21.0x |
| Weighted average supervisory leverage factor | 205.9x |
| Human-equivalent weeks | 19.1 |
Analysis
The highest factor of the day came in at 60.0x and the lowest at 5.5x, a spread of 11.0 times between the two. That is a wide range for a single day, and it usually means the day mixed mechanical work with work that needed real judgement.
The largest single entry accounted for 90.0 of the 762.0 human-equivalent hours, or 12 percent of the day. No single task dominated the total, so the weighted average is representative.
Supervisory time was 222 minutes against 2,172 minutes of execution, a ratio of about 1 to 10. Supervisory leverage of 205.9x is the figure I find most honest, because it measures the hours I actually spent rather than the hours a machine spent on my behalf.
Every figure here is recorded at the time the work is done rather than reconstructed afterwards. The human estimate is my own judgement and carries the uncertainty that implies; the minutes and tokens are measured. The full dataset, including this day, is available for download.