Skip to main content
AIDiveForge AIDiveForge

How We Rate AI Tools

Two layers: verification for every listing, editorial scores only for tools we have actually used. This page is the methodology — so you can decide how much to trust what you read here.

How tools enter the directory

New tools arrive from scheduled discovery and from submissions (readers, vendors, and our own research). Every candidate starts unpublished. Nothing goes live on discovery alone.

How data is verified

Before a tool becomes a public listing, its homepage is fetched and checked for reachability. Structured fields are taken only from what that page states:

  • Company or vendor name
  • Pricing model and visible price points
  • Free trial length in days, if offered
  • Supported platforms (web, macOS, Windows, Linux, iOS, Android, API)
  • Whether a public API is advertised
  • Open-source status

If a field can’t be confirmed from the live page, it stays empty rather than guessed. When a tool claims to be open source, that claim is checked against the public repository. No public repo, no open-source flag — regardless of what the marketing site says.

What we deliberately don’t invent

A lot of AI directories publish performance numbers — accuracy scores, hallucination rates, tokens per second — that nobody sourced. We don’t.

  • Performance metrics stay only when a vendor or a named public benchmark published them. No published number, no number here.
  • Unverified claims are not repeated. “Industry-leading,” “most accurate,” and “10x faster” don’t appear unless there’s a citation.
  • Rankings are not for sale. Listings are not ordered by who pays us. Labeled Sponsored units are advertising; they do not move rank or score.

Freshness

Every listing records a last-updated timestamp. We re-fetch homepages on a schedule. When a vendor changes pricing, drops a platform, or goes offline, the listing reflects that on the next run.

Editorial rating: nine criteria, 1–10 each

On top of verification, tools we have personally used get an Editorial Rating — a 1–10 score on each of nine criteria, shown as a radar chart and a row of bars. The overall score is the simple average of the nine. We publish the individual numbers so you can see where a tool is strong, not just whether it earned four stars.

  1. Ease of Use — How quickly a brand-new user can become productive.
  2. Output Quality — Accuracy, polish, and usefulness of generated results, measured against best-in-class.
  3. Pricing Value — What you get per dollar versus comparable tools.
  4. Feature Depth — Breadth and sophistication of capabilities (deep specialization counts).
  5. Documentation — Clarity, completeness, and freshness of official docs.
  6. Support — Responsiveness and quality of help channels.
  7. Integration Ecosystem — Native connectors, public APIs, and third-party reach.
  8. Performance — Speed, reliability, and uptime in real-world use, weighted over advertised benchmarks.
  9. Update Cadence — Frequency and substance of product updates as a leading indicator of viability.

Listings without an editorial rating have been verified against the live site but have not been hand-reviewed. We don’t score tools we haven’t used, and we don’t take payment for higher scores. Sponsored placements are labeled and have zero influence on the rating.

Corrections

Outdated pricing, a broken link, a misclassified category, a tool marked open source that isn’t — email [email protected]. Vendors may request corrections to their own listings. We don’t charge for that and we don’t require anything in return.

Last reviewed

This methodology was last reviewed in September 2026.