CrawlCheck

Directory · ML platforms and experiment tracking

Comet — what a crawler receives from comet.com

Website measurement, not a product review. This page records what machine clients received from comet.com. It does not evaluate Comet software, features, pricing or support.

75
Dgrade

Answer engines are being served something other than this site.

AI visibility 75/100, measured 2026-09-18 02:47 UTC from outside its network, as twelve named crawler identities plus an unnamed client and two mobile-browser controls. An engine can reach and read this site. What is left is giving it something specific to say.

Held at D A high-severity finding sets the letter.

Open the full report comet.com ↗

2

findings

llms.txt on its host

entity map on its host

1

scans on record since 2026-09-18

The three stages, in the order an engine hits a site

Each stage gates the next. An engine that cannot reach a page never reads it; one that cannot read it never quotes it. That is why reach is a ceiling and not just a weight.

88

Stage 1 · 40% of the score

Reach

Can a named answer-engine crawler get your pages at all?

Every named crawler was served the same page a browser gets.

57

Stage 2 · 30% of the score

Read

Once it has the bytes, can it find the words?

The content is buried in markup, or the files that guide a crawler are missing.

76

Stage 3 · 30% of the score

Quote

Is there a specific fact it can state and attribute?

An assistant would have to paraphrase your page instead of quoting a fact.

The stage to fix first is read, because the three run in order.

What to fix

2 findings on this site, 2 of them scored serious. The top one is below in full, free. Which the other 1 is, what we saw, and the bytes behind each one are what a licence buys — or we implement the whole list for you.

1 High 1 Medium

1robots.txt points at a sitemap that failsrobots.txt declares this sitemap and it resolves to the site's catch-all page (text/html; charset=utf-8, 6,694 bytes) - the same page a path that cannot exist rHighDeveloper

What we saw

robots.txt declares this sitemap and it resolves to the site's catch-all page (text/html; charset=utf-8, 6,694 bytes) - the same page a path that cannot exist receives. /sitemap_index.xml

Why it matters

robots.txt advertises a sitemap URL that does not resolve. Crawlers that trust the declaration and fail get nothing.

How to fix it

robots.txt names a sitemap that does not resolve. Fix the URL in robots.txt or publish the sitemap at the path it names.

Developer: a template, theme or application change — a developer touches it

Evidence bytes
<!doctype html><html lang="en"><head><meta charset="utf-8"><title>Comet | Supercharging Machine Learning</title><meta name="keywords" content="comet, machine learning"/><meta name="author" content="Comet"><meta name="viewport" content="width=device-width,initial-scale=1,shrink-to-fit=no,maximum-scale=1"><html itemscope itemtype="http://schema.org/Webpage"><meta itemprop="image" content="https://cdn.comet.com/img/facebook-1200x630.png"><meta name="twitter:card" content="summary"><meta name="twitter:site" content="@cometml"><meta name="twitter:title" content="Comet - Supercharging Machine Learni

DECLARED_SITEMAP_BROKEN · full transcript

2🔒 Medium finding, held backMedium

See the other 1

Every finding, what we saw, why it matters, how to fix it, and the evidence bytes — on this report and every re-scan. From $29 a month.

Unlock the findings

Or have them fixed

We implement the list on your site and re-measure, with a sealed before and after. One site, $749 once.

Get this fixed — $749

The chain a crawler follows

The chain breaks at llms.txt. Everything after that point is only reachable by a crawler guessing the conventional path.

robots.txtEvery crawler reads this first
sitemapNamed in robots.txt and resolves
llms.txtPoints onward, links stay on this host
entity graphParses as JSON and points home

A machine file that does not parse is worth less than one that is absent, because it looks present. Rows behind this.

Where the bytes go

Of the 295,239 decompressed bytes the homepage delivers, 4% is text a reader or a model can actually use. Most of what a crawler downloads here is not words.

  • Readable text 4% · 11,713 B
  • Markup & attributes 33.2% · 98,039 B
  • Structured data (JSON-LD) 0.3% · 957 B
  • Inline CSS 30.9% · 91,150 B
  • Inline JavaScript 30.4% · 89,726 B
  • Inline SVG 1% · 3,034 B
  • HTML comments 0.2% · 620 B

The individual blocks that weigh the most, so the fix is a search rather than a hunt:

  • Inline JavaScript · 80,780 B (27.4%) — unnamedopens (window.NREUM||(NREUM={})).init={privacy:{cookies_enabled:true},ajax:{
  • Inline CSS · 32,597 B (11%) — #global-styles-inline-cssopens :root{--wp--preset--aspect-ratio--square: 1;--wp--preset--aspect-ratio
  • Inline CSS · 9,624 B (3.3%) — #core-block-supports-inline-cssopens .wp-elements-1 a:where(:not(.wp-element-button)){color:var(--wp--prese
  • Inline CSS · 8,945 B (3%) — WordPress theme or block stylesopens .wp-block-image>a,.wp-block-image>figure>a{display:inline-block}.wp-bl
  • Inline CSS · 5,920 B (2%) — WordPress theme or block stylesopens .scroll-hero{--scroll-top-offset:0;--scroll-top-padding:0;position:rel

What changed

No change recorded between the readings on file yet.

1 scan on record, first 2026-09-18. Score movements are shown only between readings taken under the same score version; finding codes are comparable across all of them.

Is this your site?

This listing reads the public record. Claim it and Watch keeps that record on your terms: every list this page holds back, finding age, and twice-daily re-measurement. Free for 30 days, no card.

Claim the record for comet.com

Is this your company?

This listing came from a fetch, not a submission, so nobody was asked. If comet.com is yours you can prove it in a minute and for nothing — a DNS TXT record or one meta tag on the home page, both checkable by anyone.

Claiming marks you as the verified owner and lets you correct the category and the company name. It moves no score, no rank and no position.

Claim this listing — free or ask about a verified profile

Dispute or re-scan

If a reading here is wrong, it is our instrument that is wrong, and we want to know. Tell us what is wrong and the site is re-measured; the result is published as measured. Nothing about a listing, a payment or a request changes a number.

By email instead: hello@crawlcheck.io. Category “ML platforms and experiment tracking” is our label for what the product is primarily sold as; it is not scored. Back to the directory.