METR finds frontier AI software-task horizons doubling about every seven months
Using a 170-task suite calibrated to expert human completion times, METR introduced the 50% task-completion time horizon and estimated that the frontier for model agents on software and research tasks doubled approximately every seven months from 2019 to early 2025.
Claim: ConfirmedEvidence: Very StrongReview: Stable