Each reviewed tool carries four scores on a 0–100 scale. The first three are measured dimensions; the fourth is an editorial verdict that deliberately is not their average.
Speed
score_speed · 0–100Raw responsiveness in day-to-day use: completion and generation latency, indexing time on a real mid-size repository, cold-start time, and how the tool behaves under a slow connection. Measured hands-on during actual work sessions, not synthetic benchmarks.
Privacy
score_privacy · 0–100What leaves your machine and under which terms: training on your code, prompt retention windows, telemetry defaults and opt-outs, self-hosting options, and whether a zero-retention or enterprise mode exists. Claims are checked against the vendor’s own privacy policy and documentation, not marketing pages.
Developer Experience
score_dev_experience · 0–100How it feels to work with: setup friction, editor and CLI integration quality, configurability, documentation depth, failure modes, and how well it fits into an existing workflow rather than demanding a new one.
Overall
score_overall · 0–100The editor’s holistic verdict on a 0–100 scale. It weighs the three dimensions above together with pricing fairness and project health (maintenance activity, license, longevity risk). It is a considered editorial judgment, not an automated average — two tools with identical sub-scores can earn different overall scores when their pricing or trajectory differs.