Directory · AI coding tools
Sourcegraph — what a crawler receives from sourcegraph.com
Website measurement, not a product review. This page records what machine clients received from sourcegraph.com. It does not evaluate Sourcegraph software, features, pricing or support.
Answer engines can reach and read this site. It needs more it can quote.
AI visibility 80/100, measured 2026-09-18 03:14 UTC from outside its network, as twelve named crawler identities plus an unnamed client and two mobile-browser controls. An engine can reach and read this site. What is left is giving it something specific to say.
Held at C A medium-severity finding sets the letter.
3
findings
✗
llms.txt on its host
✗
entity map on its host
1
scans on record since 2026-09-18
The three stages, in the order an engine hits a site
Each stage gates the next. An engine that cannot reach a page never reads it; one that cannot read it never quotes it. That is why reach is a ceiling and not just a weight.
Stage 1 · 40% of the score
Reach
Can a named answer-engine crawler get your pages at all?
Every named crawler was served the same page a browser gets.
Stage 2 · 30% of the score
Read
Once it has the bytes, can it find the words?
The content is buried in markup, or the files that guide a crawler are missing.
Stage 3 · 30% of the score
Quote
Is there a specific fact it can state and attribute?
An assistant would have to paraphrase your page instead of quoting a fact.
The stage to fix first is quote, because the three run in order.
What to fix
3 findings on this site, 2 of them scored serious. The top one is below in full, free. Which the other 2 are, what we saw, and the bytes behind each one are what a licence buys — or we implement the whole list for you.
2 Medium 1 Note
1Links on this site point at pages that do not answer1 of 18 sampled internal links are gone: /cdn-cgi/content (404, linked from /).MediumContent
What we saw
1 of 18 sampled internal links are gone: /cdn-cgi/content (404, linked from /).
Why it matters
A link on this site was followed and the target answered 4xx or 5xx. A crawler spends budget on every one of these and learns nothing; a reader following one gets an error page. The usual causes are a route renamed without updating the links to it, and a nav or footer that renders links to authenticated-only routes for everyone — those answer 404 when signed out, which is what a crawler always is. Only 404 and 410 count here. A link answering 401, 403 or 429 means something declined to answer this scanner — bot protection, a login wall, a rate limit — which is a fact about the request, not about the link, so those are counted and never reported as broken. Links are sampled rather than exhaustively crawled, so this is a floor and not a full link audit.
How to fix it
Fix or remove the internal links that return 404 so crawlers and visitors are not sent to dead pages.
Content: words, markup or images on a page — no deploy
BROKEN_INTERNAL_LINK · full transcript
See the other 2
Every finding, what we saw, why it matters, how to fix it, and the evidence bytes — on this report and every re-scan. From $29 a month.
Unlock the findingsOr have them fixed
We implement the list on your site and re-measure, with a sealed before and after. One site, $749 once.
Get this fixed — $749The main product
We fix the 3 findings on sourcegraph.com. $749, once.
2 of them are scored serious. One site. You keep the grade for free — this is the part where the list stops being a list.
- Every finding implemented on the site itself: machine files, schema, NAP, the lot
- An entity map authored for your business, not a template with the name swapped
- Measured again afterwards — the before and the after, sealed, so it is evidenced rather than claimed
- Thirty days of monitoring included
- Work starts within two business days
What it covers · one payment, nothing recurring
The chain a crawler follows
The chain breaks at llms.txt. Everything after that point is only reachable by a crawler guessing the conventional path.
A machine file that does not parse is worth less than one that is absent, because it looks present. Rows behind this.
Where the bytes go
Of the 220,397 decompressed bytes the homepage delivers, 3.4% is text a reader or a model can actually use. Most of what a crawler downloads here is not words.
- Readable text 3.4% · 7,477 B
- Markup & attributes 27.6% · 60,796 B
- Inline JavaScript 1.8% · 4,072 B
- Inline SVG 65% · 143,350 B
- HTML comments 2.1% · 4,702 B
The individual blocks that weigh the most, so the fix is a search rather than a hunt:
- Inline SVG · 78,026 B (35.4%) — unnamedopens
<g clip-path="url(#a)"><path stroke="#202020" d="M385 486h326m26-177.5 - Inline SVG · 8,100 B (3.7%) —
.perspective-gridopens<!--[--><line x1="360" y1="140" x2="360" y2="460" stroke="currentColor - Inline SVG · 3,665 B (1.7%) —
.pillars-svgopens<defs><radialGradient id="pillars-glow" cx="50%" cy="50%" r="50%"><sto - Inline SVG · 3,147 B (1.4%) — unnamedopens
<path d="M4 2H64.5C66.7091 2 68.5 3.79086 68.5 6V108C68.5 110.209 70.2 - Inline SVG · 2,631 B (1.2%) —
.insight-cardopens<rect x="0.5" y="0.5" width="239" height="169" rx="6" fill="#fff" stro
What changed
No change recorded between the readings on file yet.
1 scan on record, first 2026-09-18. Score movements are shown only between readings taken under the same score version; finding codes are comparable across all of them.
Is this your site?
This listing reads the public record. Claim it and Watch keeps that record on your terms: every list this page holds back, finding age, and twice-daily re-measurement. Free for 30 days, no card.
Claim the record for sourcegraph.comIs this your company?
This listing came from a fetch, not a submission, so nobody was asked. If sourcegraph.com is yours you can prove it in a minute and for nothing — a DNS TXT record or one meta tag on the home page, both checkable by anyone.
Claiming marks you as the verified owner and lets you correct the category and the company name. It moves no score, no rank and no position.
Claim this listing — free or ask about a verified profile
Dispute or re-scan
If a reading here is wrong, it is our instrument that is wrong, and we want to know. Tell us what is wrong and the site is re-measured; the result is published as measured. Nothing about a listing, a payment or a request changes a number.
By email instead: hello@crawlcheck.io. Category “AI coding tools” is our label for what the product is primarily sold as; it is not scored. Back to the directory.