Where our data comes from
Rankwell is built on public and openly licensed data rather than a private database of guesses. Here is each source, what we use it for, and what it cannot tell you.
- Our own crawlerSite audit
Fetches your pages directly, follows internal links and your sitemap, and obeys robots.txt. Identifies itself as RankwellBot.
Limit: Sees the HTML the server sends. Content added later by JavaScript is not yet rendered.
- Google AutocompleteKeyword ideas
The public suggestion list that appears as people type into Google, for any language and country.
Limit: Shows what is commonly typed, not how many times. Exact volumes come only from Google Ads or your own Search Console.
- Google Autocomplete comparisonsDomain overview
Who people compare a site with, from what they type after its name: “semrush vs ahrefs”, “linear app vs jira”.
Limit: Only covers brands searched often enough for Google to suggest them. Brands named after ordinary words can pick up unrelated comparisons.
- Domain overview
Turns a competitor’s name into its official website, from the open knowledge base behind Wikipedia.
Limit: Smaller companies are often missing; we then only accept a .com whose title contains the name.
- Domain overview
Every HTTPS certificate is logged publicly. The names on them reveal a site’s subdomains, such as its blog, shop or app.
Limit: Shows subdomains that ever had a certificate; we keep only those that still resolve.
- Keyword ideas
Monthly readers of the article closest to your topic, published openly by the Wikimedia Foundation. Wikipedia’s related-article search also suggests neighbouring topics.
Limit: Measures interest in a subject, not searches for your exact phrase.
- Keyword ideas
What each country is searching for today, with the news stories behind each trend, from Google’s public trending feed.
Limit: Shows sudden rises, not steady everyday searches, and only the top trends per country.
- Robots files, homepages and WikidataAI search readiness
Which AI crawlers robots.txt allows, whether the firewall lets them through, structured data, llms.txt, sitemap freshness and whether the brand is a known Wikidata entity.
Limit: Measures whether AI systems can read and understand a site, not whether a specific AI currently cites it.
- Domain overview
How people in a country browse (devices, systems, browsers, connection speed) and which AI crawlers are busiest across the web.
Limit: Country-wide figures, not about any one site. Published under CC BY-NC 4.0.
- Domain overview
A top-million sites ranking built by researchers from several traffic lists averaged over 30 days, designed to resist manipulation.
Limit: Covers only the top million sites and ranks whole domains, not pages.
- Domain overview
Registration dates and registrar, straight from the domain registries. The modern replacement for WHOIS.
Limit: Some country domains publish few or no dates.
- Domain overview
The earliest saved copy of the site in the Wayback Machine.
Limit: A site can exist for a while before the archive first visits it.
- Domain overview
Google runs Lighthouse on the homepage and scores its performance, SEO, accessibility and best practices, listing the SEO checks it failed.
Limit: A single test of the homepage on a simulated mid-range phone, so scores vary a few points between runs.
- Domain overview
How fast the site really was for Chrome users over the last 28 days, with weekly history. This is the data behind Google’s Core Web Vitals.
Limit: Only published for sites with enough Chrome visitors, and covers whole sites rather than single pages here.