The comparison pages on this site cover thirteen AI agents, coding tools and app builders, each against TODO for AI and each against the rest of the field. They are generated from a single data file, so every page uses the same facts, the same six dimensions, and the same ranking rule. This guide explains that rule and indexes the pages.

Disclosure: TODO for AI builds these pages about its own competitors. Every fact carries a verification date and the scoring weights are published below so the ranking can be checked rather than trusted.

The tools covered

ToolCategoryModelsFromCompare
Claude Codecoding agentAnthropic only$20/movs · alternatives
OpenAI Codexcoding agentOpenAI only$8/movs · alternatives
CursorIDEmulti-vendor, metered$20/movs · alternatives
OpenCodecoding agentany provider (BYOK)freevs · alternatives
ClineIDE extensionany provider (BYOK)freevs · alternatives
Aidercoding agentany provider (BYOK)freevs · alternatives
OpenHandscoding agentany providerfree (self-host)vs · alternatives
Claude Coworkgeneral agentAnthropic only$20/movs · alternatives
ChatGPTgeneral agentOpenAI onlyfreevs · alternatives
Lindygeneral agentmost major models$30/movs · alternatives
Manusgeneral agentvendor-selected$20/movs · alternatives
Lovableapp buildervendor-selected$25/movs · alternatives
Base44app buildervendor-selected, choice on paid plans$16/movs · alternatives

Prices are public list prices at the verification date shown on each page (15–18 September 2026). “BYOK” tools are free to run but bill tokens through your own provider account.

How a page is scored

Each tool receives an integer score from 0 to 5 on six dimensions. The same six apply to TODO for AI.

DimensionWhat it measures
Model freedomWhich vendors and models you can run
Long-running workPersistent machine, scheduling, parallel tasks
Beyond codeMarketing, ops, browser, email — not only repositories
OpennessSource available, self-host, extend
Team & permissionsRoles, pooled usage, tool approvals
Cost predictabilityFlat plan versus metered tokens

A direct comparison (/vs/<tool>) shows both tools’ scores side by side with a feature matrix and the pricing ladder. It does not rank; it presents the two profiles and states who should choose which.

How a role guide ranks

The six audience guides — founders, solo developers, agencies, marketing teams, non-technical teams, startups — are ranked, and the ranking is computed rather than written.

1 · Gate
hard capability check
  • edits code?
  • acts in business systems?
  • usable without a terminal?
2 · Weight
per role
  • each dimension gets w ∈ [0.25, 3]
  • unlisted dimensions weigh 1
3 · Score
weighted sum
  • tools sorted descending
Srole(tool)=∑d∈Dwrole,d⋅stool,d,s∈{0,…,5}S_{\text{role}}(\text{tool}) = \sum_{d \in D} w_{\text{role},d}\cdot s_{\text{tool},d}, \qquad s \in \{0,\dots,5\}

where DD is the six dimensions above and wrole,dw_{\text{role},d} defaults to 1 for any dimension the role does not list.

The gate runs first, for roles that declare one, because a weighted sum lets a strength compensate for a task the tool cannot perform at all. The marketing-teams guide requires both acts in business systems and usable without a terminal; a tool that fails either is listed as excluded rather than ranked low, however well it edits code. The founders and startups guides declare no gate and rank the full field.

Two example weight vectors from the published data:

Rolemodelsautonomyscopeopennessteamcost
Founders1.52.53112
Solo developers320.520.251.5

The same tool lands in different positions on different guides because the weights differ, not because the facts do. Each guide page states its weights in prose so the reader can disagree with them.

Reading the pages

  • Start from the tool you already use: the vs page answers “what changes if I switch”; the alternatives page answers “what else is there”.
  • Start from the job: the role guide answers “which tool first”, with the weighting made explicit.
  • For an editor evaluated on its own terms rather than against us, read the Cursor review.
  • For the benchmark number that appears on the coding-agent pages, read the Terminal-Bench 2.1 methodology — including why it is not directly comparable with the leaderboard values shown beside it.

One-page versions

Five of the comparisons condensed to a single table each, as public Google Docs: AI to-do list: the 5-item test, Claude Cowork alternatives for teams, back office: who finishes the job, agent in your logged-in browser, Claude Code for non-coding tasks. Index sheet.