Skip to main content
AI27 February 2026

Mostly coding

19 tasks. February 27, 2026 closed at 49.7x weighted leverage across 246.0 human-equivalent hours in 297 minutes of wall-clock time. Supervisory leverage came in at 191.7x.

That is 6.2 weeks of human-equivalent throughput in 5.0 hours. The ceiling was 192.0x; the floor was 16.0x. 17 of the 19 entries came from a single project.

Task Log

#TaskHuman Est.ClaudeSup.Factor
1Hand-craft 78 [ip] figure SVGs + 11 Figures PDFs80.0h25m8m192.0x
2Infrastructure design document for 8x8 export (16 sections: VPC/ECS/S3/IAM/EventBridge/CloudWatch/CICD/SNS/Terraform/costs/security)16.0h8m5m120.0x
3Cloud knowledge benchmark: repo scaffolding + harness + 106 AWS questions + 3 model benchmark runs40.0h30m8m80.0x
4Coding8.0h8m4m60.0x
5Coding8.0h12m4m40.0x
6Coding8.0h12m4m40.0x
7Stage-isolated pipeline with per-stage model override8.0h12m4m40.0x
8Coding8.0h15m4m32.0x
9Coding8.0h15m4m32.0x
10Coding6.0h12m3m30.0x
11Coding6.0h12m3m30.0x
12Coding6.0h12m3m30.0x
13Coding6.0h12m3m30.0x
14Coding6.0h12m3m30.0x
15Standalone [quality gate] tool (run_validation.py) with batch mode and shared model instances6.0h15m3m24.0x
16Implement Pass 3 adversarial challenge generator (§440) with borderline-only filtering and hybrid Mercury+Sonnet config8.0h25m4m19.2x
17Draft 3800-word cloud knowledge benchmark article + AI detection scan + Work With Me CTA redesign + staging/production deploy8.0h25m4m19.2x
18Legacy answer support: schema + scorers + 6 question files + reporter/formatter + tests + 4 benchmark reruns + commit/push6.0h20m3m18.0x
19Coding4.0h15m3m16.0x

Aggregate Statistics

MetricValue
Total tasks19
Total human-equivalent hours246.0
Total Claude minutes297
Total supervisory minutes77
Total tokens1,985,000
Weighted average leverage factor49.7x
Weighted average supervisory leverage factor191.7x
Human-equivalent weeks6.2

Analysis

The highest factor of the day came in at 192.0x and the lowest at 16.0x, a spread of 12.0 times between the two. That is a wide range for a single day, and it usually means the day mixed mechanical work with work that needed real judgement.

The largest single entry accounted for 80.0 of the 246.0 human-equivalent hours, or 33 percent of the day. No single task dominated the total, so the weighted average is representative.

Supervisory time was 77 minutes against 297 minutes of execution, a ratio of about 1 to 4. Supervisory leverage of 191.7x is the figure I find most honest, because it measures the hours I actually spent rather than the hours a machine spent on my behalf.

Every figure here is recorded at the time the work is done rather than reconstructed afterwards. The human estimate is my own judgement and carries the uncertainty that implies; the minutes and tokens are measured. The full dataset, including this day, is available for download.