Skip to main content
AI29 April 2026

Coding and design and frontend

20 tasks. April 29, 2026 closed at 7.6x weighted leverage across 119.2 human-equivalent hours in 939 minutes of wall-clock time. Supervisory leverage came in at 115.4x.

That is 3.0 weeks of human-equivalent throughput in 15.7 hours. The ceiling was 45.0x; the floor was 3.7x. 13 of the 20 entries came from a single project.

Task Log

#TaskHuman Est.ClaudeSup.Factor
1auth-service: Apple Sign In server-to-server notification endpoint (verify signed JWT + handle email-disabled/enabled, consent-revoked, account-delete with session revocation and GDPR deletion path) + 14…6.0h8m3m45.0x
2auth-service: wire Apple + Google social login through Dockerfile + buildspec.yml + populate 7 SSM params (incl. multiline .p8) + verify (391 tests green)3.0h6m1m30.0x
3Add flashcard [content generation] stage to avian-engine (generator, writer, loop, REST endpoint, regression tests, standalone runner)6.0h22m8m16.4x
4avian-admin: hard-delete students, login-method column, Reports tab + login-methods report (auth-service + admin-service + frontend)4.0h15m3m16.0x
5Multi-piece overhaul: scenario 500 fix + cross-device user-state store + activity preferences (model+UI+filter) + Settings tabbed redesign + FAQ/guide content + flashcard [engine subsystem] scaffolding18.0h92m4m11.7x
6auth-service: Apple developer-domain-association well-known endpoint + placeholder file + test0.5h3m1m10.0x
7Debugging3.0h18m4m10.0x
8Diagnose + fix broken avian-app-web prod CI/MCQ (root cause: empty @avian/activity-ui publishes), add Convoy cross-course UI (card + InfoButton + modal), 2 new FAQ entries, autopilot guide section, dedicated…9.0h60m6m9.0x
9students plan/access bubbles + edit modal + bulk entitlements endpoint; engine postgres durability fix (asyncpg+ssl+ssm-loaded creds)5.0h35m4m8.6x
10Wire failed flashcard goals into [engine subsystem] pipeline — FlashcardTribunalRepair class, prompts/schemas, write_escalated_flashcards writer, --[engine subsystem] flag in run_flashcards.py,…3.0h22m5m8.2x
11cloud-fidelity Phase 1.a: AWS Cloudscape primitives — install + Button/Alert/StatusIndicator/KeyValuePairs wrappers + runtime dispatch shim; 290 AWS labs zero score regressions vs Phase 0 baseline7.0h55m1m7.6x
12engine: spot->on-demand cutover (clone instance, TG swap, drain, terminate spot) + ConfirmModal replacing window.confirm3.0h25m2m7.2x
13Fix audit_lab_instructions.py glob bug (missed 86 type-B labs); update canonical.json + 5 doc files for 2048→2134 / 935→1021 strict-pass; reconcile lab-manifest.ts with disk (added 7 missing entries:…3.0h25m4m7.2x
14Triage + fix all 7 watch-sweep failures (3 dashboard runtime crashes; 3 initialResources schema bugs; 1 score=0 placeholder). Bring 2,134-lab corpus to strict-pass green6.0h50m3m7.2x
15Lab post-crash recovery: cleanup + commit 1,508 in-flight changes (testId sweep + multi-checkpoint executor + sidebar nav + 90 type-B labs); fix all 9 remaining cloud-cert audit issues across AWS/GCP and add…4.0h35m5m6.9x
16cloud-fidelity Phase 3 (GCP Material 3) + Phase 4 lite (favicon/title/LICENSING) + Modal for Azure & GCP via custom-div + vendor tokens (skip vendor Dialog entirely); 304 GCP + 427 Azure + 290 AWS = 1021…14.0h145m1m5.8x
17Documentation0.8h8m3m5.6x
18cloud-fidelity Phase 1.b: AWS Cloudscape Modal + LabRunner z-index 50→9000 (above Cloudscape 5000) + force-unmount on visible=false; root-caused via Cloudscape display:none vs lab autoplay waitForGone…6.0h65m1m5.5x
19cloud-fidelity Phase 0: primitive shim + cloud detection + testid baseline + bulk import refactor (267 view files); full 2,134-lab Watch sweep regression-free; fixed pre-existing bedrock test mismatch;…10.0h120m2m5.0x
20cloud-fidelity Phase 2: Azure Fluent UI v9 — install + Button/Alert/StatusIndicator/KeyValuePairs wrappers + FluentProvider mount + runtime dispatch; Modal deferred (Fluent Dialog blocked autoplay DOM…8.0h130m1m3.7x

Aggregate Statistics

MetricValue
Total tasks20
Total human-equivalent hours119.2
Total Claude minutes939
Total supervisory minutes62
Total tokens2,687,000
Weighted average leverage factor7.6x
Weighted average supervisory leverage factor115.4x
Human-equivalent weeks3.0

Analysis

The highest factor of the day came in at 45.0x and the lowest at 3.7x, a spread of 12.2 times between the two. That is a wide range for a single day, and it usually means the day mixed mechanical work with work that needed real judgement.

The largest single entry accounted for 6.0 of the 119.2 human-equivalent hours, or 5 percent of the day. No single task dominated the total, so the weighted average is representative.

Supervisory time was 62 minutes against 939 minutes of execution, a ratio of about 1 to 15. Supervisory leverage of 115.4x is the figure I find most honest, because it measures the hours I actually spent rather than the hours a machine spent on my behalf.

Every figure here is recorded at the time the work is done rather than reconstructed afterwards. The human estimate is my own judgement and carries the uncertainty that implies; the minutes and tokens are measured. The full dataset, including this day, is available for download.