WordPress SEO audit
Find the thin archive pages, attachment URLs, ?replytocom duplicates and plugin-injected schema conflicts that accumulate on a WordPress site — and fix them at the template.
WordPress generates a lot of URLs you never asked for: tag, category, author and date archives, an attachment page for every uploaded image, comment-reply links with ?replytocom, and feeds for everything. Add a theme and a dozen plugins that each inject their own markup and the result is a site with far more crawlable pages than real content. CrawlX crawls it all in the cloud, runs 150+ technical checks and groups the findings by the template or plugin that produces them.
WordPress issues CrawlX detects
The problems that show up again and again on WordPress sites — grouped by root cause and ranked by how much of the site they touch.
01Thin tag, category, author and date archives
Every tag with one post gets an archive page; so does every author and every month. CrawlX identifies thin and near-duplicate archive pages, checks whether they are indexable, and reports them as one issue per archive type rather than hundreds of URLs.
02Attachment pages indexed as content
Each media upload can have its own attachment page containing little more than the file. CrawlX detects attachment-style URLs that return 200 with thin content and reports whether they are canonicalised, redirected or left indexable.
03Plugin-injected schema that conflicts
SEO plugins, WooCommerce, review plugins and themes each add JSON-LD. Pages end up with two Organization blocks, competing WebPage types or duplicate Product schema. CrawlX validates every JSON-LD block and flags duplicates and conflicts on each page.
04?replytocom comment-reply duplicates
Threaded comments add a ?replytocom= link per comment, creating a duplicate of the post for each. CrawlX groups the parameter variants, checks canonical and noindex handling, and lists the posts that expose them.
05Paginated archives canonicalised to page one
Themes and plugin settings sometimes point /page/2/ at the first page's canonical, hiding the posts only linked from later pages. CrawlX validates canonicals on paginated series and flags non-self-referencing ones.
06Redirect chains from permalink changes and plugins
Changing the permalink structure, forcing HTTPS at the host and adding a redirect plugin all stack hops. CrawlX follows each chain, reports the hop count and final status, and lists the internal links still pointing at the first hop.
07Uploaded images without alt text
Alt text is set per media item and blank by default. CrawlX reports images with missing alt attributes across posts and templates so the media library can be fixed in bulk.
08Feed and API URLs linked from every page
The default head includes links to RSS feeds and the REST API discovery endpoint. CrawlX reports crawlable non-HTML URLs that consume crawl budget and any that return errors after a plugin or host change.
Crawl in the cloud, fix where the code lives
Theme and plugin code in a connected GitHub repo gets fixes as pull requests. Settings that live in wp-admin get a ready-to-apply change with the affected pages listed.
The WordPress workflow
- CrawlX crawls the live site from the cloud — no plugin to install and nothing to configure in wp-admin.
- JavaScript rendering (Professional and above) audits content produced by page builders and plugins that render on the client.
- Many WordPress themes and custom plugins are version-controlled in GitHub. Connect that repository and CrawlX drafts template-level fixes — canonical logic, schema output, archive noindex rules — as pull requests on a branch. It never pushes to your default branch.
- For settings that live in the WordPress admin or in a plugin's configuration, CrawlX gives you the exact change to make and the pages it affects instead of a pull request.
WordPress SEO questions
Do I need to install a WordPress plugin to use CrawlX?
No. CrawlX crawls your site from the cloud the way a search engine does. Nothing is added to WordPress and nothing runs on your server.
Can CrawlX fix WordPress SEO issues as pull requests?
If your theme or custom plugin is in a GitHub repository you connect, yes — CrawlX drafts the fix as a reviewable pull request on a branch and never pushes to your default branch. Changes that live in wp-admin or a plugin's settings are delivered as ready-to-apply instructions with the affected pages. AI drafting is bring-your-own-key.
Will CrawlX conflict with my SEO plugin?
No. CrawlX runs outside WordPress and audits the published HTML, including what your SEO plugin outputs. It will tell you when that output conflicts with what the theme or another plugin adds — duplicate canonicals, meta tags or schema — so you can choose which one wins.
How does CrawlX handle WooCommerce stores?
The same crawl covers product, category and tag pages, faceted filter URLs and Product structured data. Findings are grouped by template so a problem on one product template is reported once, with every affected product listed.
SEO audits for other stacks
Audit your WordPress site now
Your first cloud crawl runs free — 500 URLs, no credit card, nothing to install.