23 tasks. April 13, 2026 closed at 32.7x weighted leverage across 858.0 human-equivalent hours in 1,576 minutes of wall-clock time. Supervisory leverage came in at 287.6x.
That is 21.4 weeks of human-equivalent throughput in 26.3 hours. The ceiling was 200.0x; the floor was 6.7x. 15 of the 23 entries came from a single project.
Task Log
| # | Task | Human Est. | Claude | Sup. | Factor |
|---|---|---|---|---|---|
| 1 | Build 3-tier course catalog system for accelastudy.ai: data build script, courses.jinja catalog, category-page.jinja provider pages, course-page.jinja detail pages, subscription detection JS. 892 total pages. | 40.0h | 12m | 5m | 200.0x |
| 2 | Full tools fleet consistency audit (auth, API, DB, theme, frontend) across 15 repos + canonical standards doc in tools/CLAUDE.md | 40.0h | 12m | 5m | 200.0x |
| 3 | Testing | 160.0h | 55m | 3m | 174.5x |
| 4 | Build @avian/app-shell package + migrate all 15 tool frontends (auth, theme, login, FOUC, fonts, favicons) — 244 test files, 0 failures | 120.0h | 45m | 8m | 160.0x |
| 5 | Coding | 40.0h | 30m | 5m | 80.0x |
| 6 | Analyze service gaps across 2048 labs, build service config registry for 322 slugs, AdaptiveServiceDashboard (all cloud providers), TerminalView (git/docker/k8s/cloud-shell), CodeEditorView… | 80.0h | 70m | 15m | 68.6x |
| 7 | Build and run Playwright E2E tests for all 2048 console-sim labs: config, data-testid attributes, test manifest generator, guided smoke spec, watch mode spec, run full suite 2048/2048 pass | 24.0h | 25m | 3m | 57.6x |
| 8 | Fix [engine subsystem] broken prompts, garbage RAG chunks, and busted UI — rewrote MCQ pipeline to use question bank loader, fix correctness detection, store factual chunks, add event IDs, fix calibration… | 16.0h | 18m | 5m | 53.3x |
| 9 | Refactor console-sim to single generic executor: deploy all 2048 labs, create generic executor factory with UI automation, delete 290 hand-written executor files, simplify registry | 40.0h | 45m | 8m | 53.3x |
| 10 | Fix 6 CSS defects in avian-app-web + avian-activities-react (font sizes, flashcard occlusion, timed recall instability, SVG overflow). Add 26 data-testids across 7 activity components. Update 9 page object… | 24.0h | 30m | 3m | 48.0x |
| 11 | Audit and review | 40.0h | 55m | 5m | 43.6x |
| 12 | Debugging | 16.0h | 25m | 3m | 38.4x |
| 13 | AVIAN [engine subsystem] night session: ran 45-day simulation (ongoing), fixed datetime bugs, MCQ dual-variant selectors, verifier method/options bugs, tour dismissal, activity filtering. Added Valkey prompt… | 60.0h | 240m | 20m | 15.0x |
| 14 | Testing | 40.0h | 180m | 15m | 13.3x |
| 15 | Testing | 8.0h | 45m | 5m | 10.7x |
| 16 | Debugging | 6.0h | 35m | 5m | 10.3x |
| 17 | Design and frontend | 3.0h | 18m | 8m | 10.0x |
| 18 | Autonomous [engine subsystem] simulation iteration — fixed 15+ issues across study pipeline, RAG ingestion, exam trigger, exam execution, and worker completion until Alex Chen ran full study→exam journey… | 80.0h | 480m | 30m | 10.0x |
| 19 | Testing | 3.0h | 20m | 5m | 9.0x |
| 20 | Testing | 4.0h | 28m | 5m | 8.6x |
| 21 | Coding | 6.0h | 45m | 8m | 8.0x |
| 22 | Deployment | 6.0h | 45m | 5m | 8.0x |
| 23 | Coding | 2.0h | 18m | 5m | 6.7x |
Aggregate Statistics
| Metric | Value |
|---|---|
| Total tasks | 23 |
| Total human-equivalent hours | 858.0 |
| Total Claude minutes | 1,576 |
| Total supervisory minutes | 179 |
| Total tokens | 6,464,000 |
| Weighted average leverage factor | 32.7x |
| Weighted average supervisory leverage factor | 287.6x |
| Human-equivalent weeks | 21.4 |
Analysis
The highest factor of the day came in at 200.0x and the lowest at 6.7x, a spread of 30.0 times between the two. That is a wide range for a single day, and it usually means the day mixed mechanical work with work that needed real judgement.
The largest single entry accounted for 40.0 of the 858.0 human-equivalent hours, or 5 percent of the day. No single task dominated the total, so the weighted average is representative.
Supervisory time was 179 minutes against 1,576 minutes of execution, a ratio of about 1 to 9. Supervisory leverage of 287.6x is the figure I find most honest, because it measures the hours I actually spent rather than the hours a machine spent on my behalf.
Every figure here is recorded at the time the work is done rather than reconstructed afterwards. The human estimate is my own judgement and carries the uncertainty that implies; the minutes and tokens are measured. The full dataset, including this day, is available for download.