Skip to content
Preview. Launch500 opens to the public soon; payments here are in test mode and no money moves.

Run by 99 Developers Ltd. Listings and Crash Tests pay for it. Money never buys a score or a place on the list. What money cannot buy

Categories

0 of 0 tested · updated 20 September 2026

The best ai coding agents and automated pull request tools

Agents that take a written task or an issue, change code across a repository, run tests, and open a pull request for human review.

What has to be true to appear here

  • Operates across a repository, not a single file or an editor autocomplete
  • Runs the project's own tests before proposing a change
  • Opens a reviewable diff rather than writing directly to a default branch
  • Documents what repository access it requires

Inclusion is not for sale, and neither is order. Nine of the products we cover have no affiliate programme at all.

What every product here was put through

  1. 1.Assign 40 real issues from three open-source repositories of different sizes and measure the share of pull requests merged without human edits
  2. 2.Measure how often the agent claims tests pass when they do not
  3. 3.Review 20 diffs for scope creep, unrelated files touched, formatting churn, dependency additions
  4. 4.Test behaviour on an under-specified issue: does it ask, or does it guess and commit?

Ranked by Crash Test score

The shortlist

Tested products come first, ordered by score. Untested products appear below them regardless of their community rating, because we have not looked at them. Nothing on this page can read a vendor’s subscription tier.
ProductCrash TestVerified reviewsFrom (5 seats)Best for

Feature matrix

What each one actually does

Confirmed in testing. Where the honest answer is a sentence rather than a tick, we write the sentence. A safety-relevant feature that is off by default may not be recorded as a plain yes.
Capability
Runs your test suite
Pull request only, never direct push
Asks when under-specified
Self-hosted runner
Scope guard on diffs
Languages supported

Real questions

What buyers actually ask

Sourced from launch threads, review text and search data, not invented to fill a schema block. Each answer is two to four sentences and says a number where we have one.

Why this page has no winner badge

“Best overall” is a question about your situation, not about the software. The table gives you a tested score, a real price at your team size, and a one-line statement of who each product is wrong for. The last of those is usually the one that decides it.

Found something out of date?

Pricing moves constantly in this category and we miss things. Send a correction with a dated link and it goes in the public log, with your handle on it if you want it there.

Corrections log and policy