Fixed-price audit · ₹12,500 +91 89254 44011 WhatsApp 6 am – 11 pm
Free · the whole checklist

The 42-Point SEO Audit Checklist — Published in Full, Including the Six Nobody Else Runs

This is the checklist behind our ₹12,500 audit, published rather than gated. If you have a developer and a free afternoon, you can run most of it yourself — and you should, before you pay anybody.

In short

Forty-two checks in six sections: crawl and access, indexation and duplication, structure and linking, rendering and speed, structured data, and AI crawler access. The last six are the ones almost no Chennai audit includes. Running it yourself takes a focused afternoon; we charge ₹12,500 to do it in 10 working days.

Reviewed 20 September 2026 · revised twice a year; the AI-crawler section is reviewed quarterly as user agents and tokens change

Why publish the thing we sell

Because the checklist is not the product. Anyone can read a list of 42 checks; what takes ten working days is running them properly on a real site, evidencing each finding with a URL or a screenshot, and — the part that actually matters — ordering the fixes by effort against impact so a developer can start on Monday. The list is the easy half and gating it would be a marketing decision, not an honest one.

There is a second reason, and we will state it because this site publishes its reasoning: a page like this earns citations. Procedural content carries a 2.8× schema multiplier with HowTo markup and how-to queries average 5.1 AI citations against 3.1 for commercial pages. Publishing the checklist is both the right thing and the profitable thing, which is a comfortable position to be in and worth being transparent about.

The 42 checks

Six sections. Run them in order — there is no point optimising a page that is not being crawled.

Section#CheckWhy it matters
A. Crawl and access1robots.txt does not block anything that earns moneyA staging rule pushed live is the single most expensive one-line mistake in SEO
2XML sitemap lists only canonical, indexable, 200-status URLsA sitemap full of redirects and noindex pages teaches a crawler to distrust it
3No orphan pages — every important URL is linked from somewhereIf your own site does not link to it, nothing else will find it either
4Money pages sit within three clicks of the homepageCrawl depth is a proxy for how important your own architecture says a page is
5No 5xx errors in the last 30 days of server logsIntermittent server errors quietly reduce crawl rate for weeks afterwards
6Soft 404s identified — pages returning 200 with nothing on themCommon on empty category and filter pages; they dilute everything around them
7No redirect chains or loopsEvery hop loses time and some signal, and loops lose the page entirely
B. Indexation and duplication8Indexed page count roughly matches the number of pages you meant to publishA 40,000-URL index on a 600-page site means parameters are being crawled
9No noindex or nofollow on a page that is supposed to rankUsually a leftover from a rebuild; costs months before anyone notices
10Canonicals are self-referencing and consistentConflicting canonical and sitemap signals are resolved by the engine, not by you
11Parameter and faceted URLs are controlled deliberatelyThe commonest source of index bloat in ecommerce
12No two pages target the same queryCannibalisation splits signal. Two of the largest Chennai agencies do this to themselves
13Paginated sets handled with a stated approach“Load more” without URLs makes page two invisible
14Internal search result pages are not indexableThey generate infinite low-value URLs from nothing
C. Structure and linking15Exactly one H1 per pageThree of the twelve largest Chennai agency sites fail this on their own money pages
16Heading hierarchy is logical, not decorativeHeadings are the outline a retrieval system reads before the prose
17Internal links point to money pages with meaningful anchorsThe cheapest ranking lever most sites have never pulled
18Anchor text varies and describes the destination“Click here” tells an engine nothing about the page it points at
19Breadcrumbs present and marked upHelps both the user and the BreadcrumbList that appears in results
20URL structure is consistent — case, trailing slashes, separatorsInconsistency creates duplicates that nobody intended
21Mobile and desktop expose the same content and linksContent hidden on mobile is content the mobile-first index may not weigh
D. Rendering, speed, mobile22Core Web Vitals read from field data, not a lab scoreLab scores flatter. Field data is what users actually experienced
23Key content exists before JavaScript hydrationIf the answer only appears after hydration, assume it does not exist
24Images sized, compressed and in a modern formatUsually the largest single weight on an Indian SME site
25Layout shift sources identified — images, ads, late fontsCheap to fix once you know which element is moving
26Tap targets and viewport behave on a real phoneTest on a mid-range Android on mobile data, not on your desktop
27Server response time measured, not assumedOn WooCommerce sites hosting is usually the constraint, not the theme
28Third-party scripts audited for what they costChat widgets and analytics stacks routinely outweigh the page they sit on
E. Structured data and meta29Every schema type in use validates without errorsInvalid markup is worth nothing at all
30One Organization or LocalBusiness entity, referenced by @idA fresh copy on each page makes the entity harder to resolve, not easier
31FAQPage markup matches the questions visible on the pageMarking up questions a user cannot see is a guidelines problem
32Prices in Product, Service or Offer match the visible priceA wrong price in schema is worse than no schema — AI answers will quote it
33Titles are unique, descriptive and front-load the termStill the highest-leverage 60 characters on any page
34Meta descriptions written for the click, not for a keywordThey do not rank you; they decide whether the ranking is worth anything
35Canonical, Open Graph and hreflang agree with each otherHalf-done internationalisation is worse than none
F. AI access and citability36GPTBot allowed in robots.txtChecked by name — a blanket allow can still be overridden further down the file
37ClaudeBot allowedSame check, different token
38PerplexityBot allowedPerplexity cites densely and is often the first engine a new site wins
39Google-Extended allowedSeparate from Googlebot; blocking it removes you from some AI surfaces while leaving search intact
40CDN, WAF and bot rules tested from outside the networkA robots.txt that allows and an edge rule that blocks is the commonest false pass
41llms.txt published and pointing somewhere usefulAn hour of work, no guarantees, and no reason not to
42Answer-first structure and a 10-run baseline citation testYou cannot show a citation improvement without a starting rate
Checklist version 1.0, September 2026Checks 36 to 42 are the AI-access section. We have not found another Chennai audit that includes them.

How to prioritise what you find

A crawler will hand you a list of a thousand issues and no sense of proportion. The order that works is: anything blocking access first — robots.txt, noindex, crawler rules, server errors — because nothing else you do matters while a door is shut. Then anything splitting signal: cannibalised pairs, canonical conflicts, parameter bloat. Then anything missing: schema, internal links, answer-first structure. Then anything slow, which is usually the section everyone starts with and rarely the one that moves rankings.

Score each finding on effort and impact and do the low-effort, high-impact ones this week. In practice that is usually three or four fixes, and they are usually in the first two categories. The thousand-item list can then be triaged honestly: most of it is one rule applied a thousand times.

If you run this yourself and find something alarming

Send it to us and we will tell you whether it matters, free, with no obligation and no call required. We would rather answer a question from somebody who ran the checklist themselves than sell an audit to someone who did not need one. If it turns out you do need the full thing, it is ₹12,500, ten working days, and credited against a retainer started within 30 days.

Checklist questions

Can I really run this myself?

Most of it, in a focused afternoon, if you have Search Console access and a crawler. The parts that are hard without practice are prioritisation, log analysis and the AI-crawler tests from outside your own network. The list itself is not secret and never should have been — what takes ten working days is doing it properly on a real site.

What tools do I need?

Search Console, a crawler such as Screaming Frog, a structured data validator, and a way to see field Core Web Vitals rather than lab scores. For the AI-access checks you need to test robots.txt and your CDN rules from outside your network, because an edge rule can block what robots.txt allows.

Which checks matter most?

Anything that blocks access — robots.txt, stray noindex tags, crawler rules, server errors. Then anything splitting signal, like cannibalised page pairs and parameter bloat. Speed usually matters least of the four for ranking, though it matters plenty for users. Fix in that order and most sites see something move within six weeks.

Why are the AI-crawler checks separate?

Because GPTBot, ClaudeBot, PerplexityBot and Google-Extended each have their own robots.txt token and each has to be checked by name. We find at least one blocked on roughly half the sites we audit, including sites paying another agency ₹40,000 a month, and it is almost never a decision anybody made deliberately.

How is your paid audit different from this list?

Evidence and order. Every finding arrives with a URL or a screenshot, a severity, an effort estimate and a position in a fix queue, plus a video walkthrough and a written answer on whether the site is worth a retainer at all. Ten working days, ₹12,500, credited if you start a retainer within 30 days.

How often should a site be audited?

Annually for a stable site, and immediately after any migration, replatform or CMS upgrade. Migrations are where crawler access, canonicals and schema break silently — we have seen all four AI user agents blocked by a default rule the morning after a replatform and nobody notice for six weeks.

Can I give this checklist to my developer?

Please do. That is what it is for, and it is also what a good freelance brief looks like — a numbered list with a reason attached to each item. If your developer disagrees with an item, they are probably right about your specific setup and worth listening to.

Related pages

Send the URL. Get a tier, a price and a start date.

One working day, on WhatsApp, from a person who has already opened your site in Search Console-shaped eyes. No discovery call required before you get a number.