aimpg scoreboard

Do token-savers, prompts, models and agents really save money on real code? Each result reruns a developer's own past commits both ways, and their own tests judge them. Made with aimpg verify; 0 records so far.

Reproduced

Counted only when someone else reran a public record and got the same answer.

Nothing here yet: no claim has enough repos and people behind it.

Self-reported (not audited)

Nothing here yet: no claim has enough repos and people behind it.

Collecting data

—

A number appears once at least 5 repos and 3 people back it, with no one supplying more than 50% of the runs. Each account counts for at most 3 repos per claim. Results use the median repo, never an average.

Disputed

Public records (rerunnable)

TierClaimAnswerByWeek
No public records yet.