Skip to content
RESULT-FIRST DECISION GUIDEDECIDE FROM THE DELIVERABLE

Which kind of AI should you use when you want work done?

Choose chat, a one-off general agent, recurring automation, or a coding agent from the deliverable first. Then check the official entry point, permissions, cost, and stop conditions. The product name can come later.

For: Everyday workers, independent creators, and non-specialist builders who have tried AI but are easily misdirected by trending names, tutorials, and overlapping features

ANSWER THIS FIRST

Do you need an answer, a one-off finished artifact, a recurring workflow, or maintainable software?

No installation required yetPoint to one of the four deliverables below. If none fits yet, narrow the task instead of searching for a more powerful tool.
Possible routes
4
First decision
Deliverable
Product sources
10
Independent comparison
Not yet

01 / SEE THE DECISION FIRST

You do not need the “best tool.” You need the right kind of delivery

Chatting, operating a website, running on a schedule, and maintaining code carry different responsibility boundaries. Define the result, then check whether you can supply context, supervise actions, accept permissions, and maintain the outcome. More features can otherwise mean more costly trial and error.

Smallest useful decision: Stop at chat when content is enough. Move to an agent only when AI must act. Automate only after the same task repeats reliably. Use a coding agent only when the deliverable itself is software.

02 / MATCH YOUR SITUATION

Consider task, readiness, and cost together

Each route names when to use it and when to stop or switch. Products appear only as examples checked against first-party sources; they do not replace your own minimal test.

01

Need an answer: start with chat

The deliverable is an explanation, comparison, outline, draft, or edit suggestion. AI returns content; you judge, copy, and act on it without granting access to external systems.

Task and user readiness
Use it when the task can end in the chat and you can judge whether the response is correct. It is the lowest-cost starting point for a first AI task or an unclear request.
Permissions, cost, and stop/switch condition
You still open files, fill websites, and send content yourself. Once the goal becomes “change and deliver it for me,” stop piling on prompts and switch to a bounded one-off task.

Product examples to verify further

FIRST-PARTY SOURCE CHECKED · NOT INDEPENDENTLY TESTED

Names and links come from first-party product pages. Recheck capability, price, permissions, and fit before proceeding.

02

Need it once: use a general agent

The deliverable is a specific file, one web task, a research result, or a reviewable set of changes. Let an agent perform bounded steps while you inspect the outcome.

Task and user readiness
Use it when the task has a clear finish line, can be retried after failure, and you can supervise a small non-sensitive trial. Coding skill is not required.
Permissions, cost, and stop/switch condition
Website, file, and account access is broader than chat, and each run may consume usage. Stop at sign-in, sending, payment, deletion, or unexplained extra steps. Consider automation only after the same task repeats reliably.

Product examples to verify further

FIRST-PARTY SOURCE CHECKED · NOT INDEPENDENTLY TESTED

Names and links come from first-party product pages. Recheck capability, price, permissions, and fit before proceeding.

03

Need it again: build recurring automation

The deliverable is not only today's result but a workflow that can run again: defined inputs, checks, run history, failure alerts, and a pause control.

Task and user readiness
Use it after the same task has succeeded manually or under supervision several times, inputs and acceptance criteria are stable, and someone can inspect logs, handle failures, and update connections.
Permissions, cost, and stop/switch condition
It needs a persistent runtime, model or API budget, account authorization, and maintenance. Pause when sites, APIs, rules, or owners change. If every input and judgment differs, return to a one-off agent.

Product examples to verify further

FIRST-PARTY SOURCE CHECKED · NOT INDEPENDENTLY TESTED

Names and links come from first-party product pages. Recheck capability, price, permissions, and fit before proceeding.

04

Need maintainable software: use a coding agent

The deliverable is a script, website, application, integration, or repository change. Completion includes execution, tests, diff review, and future maintenance—not code generation alone.

Task and user readiness
Use it when an existing tool cannot cover the need and you have a runnable project, explicit acceptance criteria, and someone who can review tests and maintain the result.
Permissions, cost, and stop/switch condition
Repository, terminal, network, and deployment access is broader, with subscription, compute, dependency, and maintenance costs. Stop if there is no test baseline or maintainer. Return to a general agent when only one result is needed.

Product examples to verify further

FIRST-PARTY SOURCE CHECKED · NOT INDEPENDENTLY TESTED

Names and links come from first-party product pages. Recheck capability, price, permissions, and fit before proceeding.

03 / AVOID THE FIRST MISTAKES

3-minute check before you start

No long questionnaire. If one item is unclear, narrow the task, permissions, or tool category first.

01

Name and official entry point

Confirm the standard product name, developer, and official domain. Follow official links to the installer, store, or repository.

Stop when: You only have a tutorial, chat-group link, shared installer, or a similar name that cannot be tied to the developer.
02

Final deliverable

State whether this task needs an answer, one finished artifact, a repeatable workflow, or maintainable software.

Stop when: You can only say “have AI do it” but cannot name what should exist when it is done.
03

Acceptance check

Define one observable success result and prepare a minimal sample for the first attempt.

Stop when: The only success evidence is the agent saying it is done.
04

Permissions and external actions

List the files or accounts it will read and any send, edit, delete, payment, or publish actions.

Stop when: The first test requests a primary account, whole-disk access, payment, or automatic sending.
05

Pause and takeover

Confirm where to pause, how to inspect the raw result, and who takes over or rolls back after failure.

Stop when: The run is invisible, cannot be interrupted, or failure only leads to broader authorization and another attempt.
Ready for a minimal test

The deliverable, entry point, permissions, acceptance check, and pause path are clear. Start with a non-sensitive sample.

Verify one gap first

The category broadly fits, but identity, cost, data path, or the result standard still needs verification.

Do not authorize the run yet

Stop when the name or entry point cannot be confirmed, the deliverable is unclear, or the first attempt requires irreversible access.

04 / SOURCES AND EVIDENCE BOUNDARY

Know what makers say—and what we have not proved

BitShovel has not independently compared the named products on the same everyday-user task and has no real-user completion, total-cost, or long-term maintenance data.

FIRSTHAND DEVELOPER PRACTICE

Simon Willison

When a coding agent enters an existing project, have it run the existing tests first.

A test baseline teaches the agent how the project defines success and gives the changed result the same yardstick. Unexecuted code cannot be accepted from a summary alone.

Open the original article

Next, run one low-risk sample

Do not install all four categories. Pick one route, confirm the official entry point, and get one observable result with non-sensitive data before continuing.

Check tool profiles
Which kind of AI should you use when you want work done? — BitShovel