GUESSS India / www.guesssindia.in
Audit AI crawlability, GEO files, indexation, and site-level signals.
Audited website
https://www.guesssindia.in
Last analyzed 8/15/2026, 10:46:55 PM
Technical health score
78
out of 100
Improvements recommended
Across 9 scored checks
Audit result
6
Healthy technical signals
Audit result
2
Improvements recommended
Audit result
1
Visibility blockers found
Action plan
Start here for the changes most likely to improve discovery and eligibility.
No accessible sitemap.xml was found.
Recommended next step
Publish a valid XML sitemap containing live canonical URLs, include accurate lastmod dates, and remove redirected or broken pages.
2 format issues detected.
Recommended next step
Keep the file focused, add one H1 and a short blockquote summary, and use only absolute links to your strongest canonical pages.
0 of 5 AI agents have an explicit policy.
Recommended next step
Add explicit rules for retrieval and training crawlers. Allow the retrieval bots you want to surface your content and make a deliberate choice for training bots.
AI-readable files, crawler access, bot policy, and sitemap readiness.
2 format issues detected.
A concise Markdown guide that helps AI systems understand your organization and find its most important pages.
A concise Markdown guide that helps AI systems understand your organization and find its most important pages.
A clear, valid file can reduce ambiguity when AI assistants discover and interpret your content. It is an emerging GEO signal, not a guarantee of inclusion.
Keep the file focused, add one H1 and a short blockquote summary, and use only absolute links to your strongest canonical pages.
Technical values collected from the live website.
Bytes
Has H1
Issues
Link Count
Http Status
Has Blockquote
Absolute Urls Only
Optional extended file is not present.
An optional extended reference containing deeper documentation or supporting context for AI systems.
An optional extended reference containing deeper documentation or supporting context for AI systems.
This can help content-heavy sites provide richer context, but its absence does not prevent crawling or indexing.
Add it only when the short llms.txt cannot represent the site clearly. Keep the main file concise and link to the extended version.
Technical values collected from the live website.
Http Status
0 of 5 AI agents have an explicit policy.
Checks whether major AI crawlers have clear, intentional access rules instead of relying only on a generic wildcard rule.
Checks whether major AI crawlers have clear, intentional access rules instead of relying only on a generic wildcard rule.
Blocking retrieval crawlers can prevent your pages from appearing in AI answers. Unclear rules also make your AI visibility policy harder to control.
Add explicit rules for retrieval and training crawlers. Allow the retrieval bots you want to surface your content and make a deliberate choice for training bots.
Explicit means the crawler has its own rule. Inherited means it follows the general * rule.
OAI-SearchBot
retrieval · allowed
Inherited from *
PerplexityBot
retrieval · allowed
Inherited from *
GPTBot
training · allowed
Inherited from *
ClaudeBot
training · allowed
Inherited from *
Google-Extended
ai control · allowed
Inherited from *
Technical values collected from the live website.
Retrieval and training crawler policies are shown separately.
Separates crawlers used to retrieve content for answers from crawlers that may collect content for model improvement.
Separates crawlers used to retrieve content for answers from crawlers that may collect content for model improvement.
Retrieval access is directly connected to whether supported AI search products can discover current pages. Training access is a separate business choice.
Review each bot individually. Avoid blocking retrieval by accident, and document the training policy that matches your organization’s preference.
Explicit means the crawler has its own rule. Inherited means it follows the general * rule.
OAI-SearchBot
retrieval · allowed
Inherited from *
PerplexityBot
retrieval · allowed
Inherited from *
GPTBot
training · allowed
Inherited from *
ClaudeBot
training · allowed
Inherited from *
Google-Extended
ai control · allowed
Inherited from *
No accessible sitemap.xml was found.
Checks whether search and AI crawlers can find a valid list of canonical site URLs and understand when pages changed.
Checks whether search and AI crawlers can find a valid list of canonical site URLs and understand when pages changed.
Missing, invalid, or broken sitemap URLs can slow discovery and leave important pages outside search and AI retrieval indexes.
Publish a valid XML sitemap containing live canonical URLs, include accurate lastmod dates, and remove redirected or broken pages.
Technical values collected from the live website.
Url Count
Http Status
Content Type
Lastmod Count
Sampled Url Count
Not required for the detected sitemap size.
Checks for an index that organizes multiple sitemap files when the site is too large for one sitemap.
Checks for an index that organizes multiple sitemap files when the site is too large for one sitemap.
Large sites without a sitemap index can become harder to crawl completely. Smaller sites do not need one.
Use a sitemap index when you have multiple sitemaps or approach protocol limits; otherwise no action is required.
Technical values collected from the live website.
Required
Http Status
Sitemap Count
Index controls, canonicals, redirects, and hidden crawl restrictions.
1 valid user-agent group parsed.
Validates crawler rules and looks for directives that may accidentally hide CSS, JavaScript, or important site assets.
Validates crawler rules and looks for directives that may accidentally hide CSS, JavaScript, or important site assets.
Incorrect rules can block entire site areas or prevent crawlers from rendering pages accurately, reducing indexation and answer visibility.
No urgent change is needed. Continue monitoring this check after major site updates.
Technical values collected from the live website.
Group Count
No restrictive directives found across 1 sampled key pages.
Looks for page-level instructions such as noindex, nofollow, and noarchive across a sample of important URLs.
Looks for page-level instructions such as noindex, nofollow, and noarchive across a sample of important URLs.
A noindex directive removes a page from eligible search results and can also eliminate it as a source for AI retrieval.
No urgent change is needed. Continue monitoring this check after major site updates.
Technical values collected from the live website.
Sampled Pages
Sampled pages use consistent canonical hostnames.
Checks whether each sampled page points to the preferred version of itself and uses one consistent hostname.
Checks whether each sampled page points to the preferred version of itself and uses one consistent hostname.
Missing or conflicting canonicals split authority across duplicate URLs and can cause crawlers to index the wrong version.
No urgent change is needed. Continue monitoring this check after major site updates.
Technical values collected from the live website.
Sampled Pages
HTTP requests permanently redirect to HTTPS.
Verifies that insecure HTTP visits are permanently sent to the secure HTTPS version.
Verifies that insecure HTTP visits are permanently sent to the secure HTTPS version.
A missing or temporary redirect can create duplicate versions, dilute page signals, and reduce trust for users and crawlers.
No urgent change is needed. Continue monitoring this check after major site updates.
Technical values collected from the live website.
Canonical hostname is consistently guesssindia.in.
Checks that www and non-www addresses resolve to one preferred hostname instead of behaving like separate sites.
Checks that www and non-www addresses resolve to one preferred hostname instead of behaving like separate sites.
Two accessible hostnames can duplicate content and divide links, crawl activity, and ranking signals.
No urgent change is needed. Continue monitoring this check after major site updates.
Technical values collected from the live website.
1 sitemap directive found.
Checks whether robots.txt tells crawlers where the sitemap is located.
Checks whether robots.txt tells crawlers where the sitemap is located.
The directive provides another reliable discovery path. Its absence is usually a missed optimization rather than a blocking issue.
No urgent change is needed. Continue monitoring this check after major site updates.
Technical values collected from the live website.
Directives
Sitemap URL health uses a bounded sample to keep audits fast and safe. Each completed run is saved to this brand's audit history.