Tool families

AI coding tools: evaluate an assistant in your repository

Generation speed does not show whether code fits a project. A useful trial takes one bounded task in a real repository, with tests, review and a way to undo changes.

Development workstation with a code screen
AI-generated illustration.

When to explore this family

Explain an area of code before proposing an edit.

Fix a reproducible defect and verify the affected behaviour.

Build a limited prototype, then inspect its dependencies.

How to compare tools

  1. Compare the proposal with existing conventions and dependencies.

  2. Run meaningful tests and review the diff line by line.

  3. Check permissions, repository access and data handling.

A repeatable trial

Starting situation. In a small test repository, fix a known defect without changing neighbouring behaviour. Prepare a test that fails before the fix.

  1. Baseline

    Give two assistants the same ticket, files and failing test. Note the commands each proposes to run.

  2. Diff review

    Inspect changes line by line: scope, new dependencies, secret access, error handling and readability.

  3. Evidence

    Run the regression test and project checks. Measure human review time and how easily the change can be rolled back.

Evidence for a decision

Output
A bounded, readable change validated in the target repository.
Checks
Reproduction, diff, existing tests and adjacent behaviour.
Evidence
Branch, instructions, final diff, commands and test results.
Failure signal
A fix without reproduction or with out-of-scope changes.
Compare actual cost and keep a decision record →
47 active services in the directory

Services to examine

These active services are starting points from the directory. Editorial highlighting and score order this selection; they do not replace a trial on your own task.

Code & buildaccount

Amazon Q Developer

coding agent

Amazon Web Services · US

Official site

Related guides

Questions before choosing

Is compiling code enough?

No. It must solve the requested case, preserve nearby behaviour and pass relevant tests.

What else should be measured besides generation speed?

Count rework, useful tests, out-of-scope changes and review time.

What access should an agent receive?

Limit access to the repository and actions needed for the test; keep secrets and sensitive operations out of reach.

Explore other tool families