Documentation
SEO growth guide
What each area actually means, what Crawlix does about it, and what only you can do.
Work in this order
Search visibility is built bottom-up. Working on a higher layer while a lower one is broken is decorating a house with no foundation.
- Technical foundation → crawlability → indexability
- Site architecture → on-page SEO → content quality
- Topical authority → internal linking → structured data
- Entity and brand signals → authority and links
- Then AEO and GEO, then continuous monitoring
Technical SEO
Status codes, redirect chains, HTTPS problems, response headers. This is the strongest part of Crawlix.
Crawlix automates: security headers, render-blocking resources, LCP preload, lazy loading. Crawlix recommends: status code and redirect problems. You do: server and hosting changes.
Be careful here
Redirects are never automated. A wrong redirect is one of the fastest ways to lose traffic, and unpicking a redirect chain afterwards is painful.
Crawlability
Whether search engines can reach your pages at all. Crawlix checks robots.txt conflicts, crawl traps, and orphan pages with no internal links pointing at them.
robots.txt changes always require your review. One wrong line can remove an entire site from search.
Indexability
Whether pages that can be crawled are actually eligible to appear. Crawlix checks noindex directives, canonical conflicts, and — when Search Console is connected — Google's own URL Inspection verdict for the page.
A page that cannot be indexed cannot rank, so every other fix on it is wasted. Fix indexation first, always.
Metadata
Titles, meta descriptions, Open Graph and Twitter cards. Titles influence ranking and click-through; meta descriptions influence click-through only.
Crawlix automates: titles, meta descriptions and Open Graph, subject to the risk check. You do: brand voice — read them before approving.
Schema and structured data
Machine-readable markup describing what a page is about. Crawlix detects and generates many types: FAQ, HowTo, Product, Image, Video, Audio, entity and organisation markup, and E-E-A-T signals.
Schema fixes are additive and on the safe auto-apply list. The exception is anything asserting facts — LocalBusiness address, phone, hours and geo always require your review, because a wrong opening time is a customer problem rather than an SEO one.
Be careful here
Never approve schema asserting something you have not verified. Structured data is a claim you are making to search engines in machine-readable form.
Images
Missing alt text, lazy loading, and preloading the largest image so it renders sooner.
All three are on the safe auto-apply list. Alt text is one of the highest-volume, lowest-risk wins available — if you automate one thing, automate this.
Internal linking
How pages connect. Crawlix finds orphan pages, measures click depth, and suggests specific links with anchor text.
Internal link fixes always require review. Apply them one at a time — a bad linking pattern applied in bulk is hard to unpick.
Aim for nothing important being more than three clicks from your homepage.
Content opportunities and topical authority
Crawlix identifies thin content, duplication, content gaps, decaying pages, and cannibalisation — two of your pages competing for the same query. It can generate content briefs.
Be careful here
Crawlix does not auto-publish body content, and through most CMS integrations it cannot publish it at all — only the GitHub pull-request route can write body content. Plan to write it yourself.
Topical authority comes from covering a subject properly rather than from any single page. Crawlix can show you the gaps; filling them is editorial work.
Core Web Vitals
Crawlix uses two different sources, and the difference matters:
| Source | What it is | Requires |
|---|---|---|
| Lab data | A simulated load via Google PageSpeed Insights | Nothing from you |
| Field data | Real measurements from your actual visitors | The Crawlix snippet installed, plus real traffic |
Field data is the one Google uses. If you have installed the snippet but see no field data, the usual reason is simply that not enough real visits have happened yet.
Google Search Console
Connecting Search Console gives Crawlix your queries, impressions and click-through data, plus the ability to ask Google directly whether a URL is indexed.
Worth knowing
The Search Console property must match your project domain exactly. www, the bare domain, and http versus https are all separate properties in Google's eyes. Data can take up to 48 hours to appear.
Google Analytics 4
Connecting GA4 gives Crawlix page traffic and key events, which is what lets a change be measured against something other than its own score.
Not currently supported
Crawlix will not add a GA4 tag to your site for you. It can detect that analytics is missing, but injecting a placeholder measurement ID would silently break your analytics rather than fix anything — so it reports the problem and leaves the real ID to you.
Keywords, rankings and competitors
Crawlix can track real SERP positions for keywords you choose, identify keyword opportunities, and compare you against competitors.
Worth knowing
All of this requires a rank-data provider configured on your plan. Without one, every figure is null and the interface says so rather than estimating. That is deliberate: an invented number is worse than a blank one.
When the problem is not on your site
For priority keywords, Crawlix reports which of five factors is holding you back — including when the answer is one it cannot fix, such as referring domains or brand search demand. It will not diagnose from a single signal, and where the constraint is off-site it says plainly that no amount of on-page work will move it.