Independent SiteAuditLint analysis
SiteAuditLint crawled 100 URLs on cnn.com. Most critical items trace to one source: all 14 CNN Underscored URLs returned HTTP 403, and the sitewide navigation links to them from 81 pages. Page weight is the other large pattern, with section pages between 3 and 5 MB of HTML.
| Metric | Result |
|---|---|
| Pages crawled | 100 |
| Crawl time | 47s |
| Health score | 82 / 100 (Good) |
| Issue types detected | 28 |
| Critical / warnings / opportunities / info | 95 / 156 / 512 / 7 |
| Audit date | October 4, 2026 |
What stood out
Overall assessment
cnn.com's core news pages loaded with 200 status codes, self-referencing canonicals and hreflang. The critical count of 95 comes almost entirely from one section: every /cnn-underscored/ URL returned HTTP 403 to the crawler, and the global navigation links to them, so 81 pages register as linking to a broken internal page.
A 403 limited to one section points to access rules on that section rather than missing pages. It is still worth confirming, because those URLs also appear in the XML sitemap.
Two sitewide patterns deserve attention after that. A help.cnn.com link in the shared footer returned 404 on 81 pages, and section pages ship between 3.1 and 5.2 MB of HTML, which is roughly ten times the 500 KB threshold SiteAuditLint uses.
Audit methodology
SiteAuditLint started from the website's homepage and crawled up to 100 URLs by following discoverable internal links and the URLs listed in the XML sitemap. Results represent the URLs reached during the crawl and do not constitute a complete audit of the entire website.
| Starting URL | https://cnn.com/ |
|---|---|
| Crawl limit | 100 URLs |
| URLs crawled | 100 (98 HTML) |
| Crawl date | October 4, 2026, 18:04 UTC |
| Crawl duration | 47s |
| Crawl method | Homepage start, internal link and sitemap discovery, robots.txt respected, raw HTML analysis |
| Tool | SiteAuditLint Desktop SEO Crawler |
| Scope | Publicly accessible URLs reached during the crawl |
Technical SEO scorecard
Scores are the category scores SiteAuditLint calculates as part of its Site Health Score.
| Area | Score | Key finding |
|---|---|---|
| Overall health | 82 | Good rating from SiteAuditLint |
| Technical | 84 | 14 Underscored URLs returned 403 |
| Content | 96 | 2 near-duplicate market pages, 1 thin page |
| Performance | 88 | 46 slow responses, 81 pages over 500 KB |
| AEO | 93 | robots.txt blocks all 9 AI crawlers checked |
Major findings
1. CNN Underscored URLs returning 403
HighFinding: All 14 URLs under /cnn-underscored/ returned HTTP 403. All 14 are in the XML sitemap, and the global navigation links to them from 81 pages.
Why it matters: A 403 tells clients they may not fetch the page. If search engine crawlers receive the same response, the section cannot be crawled despite being in the sitemap. See status codes.
| URL | Result |
|---|---|
https://www.cnn.com/cnn-underscored | HTTP 403 |
https://www.cnn.com/cnn-underscored/deals | HTTP 403 |
https://www.cnn.com/cnn-underscored/electronics | HTTP 403 |
Recommended fix: Check server or CDN logs to confirm verified search engine crawlers receive 200 for /cnn-underscored/ URLs.
2. Sitewide link to help.cnn.com returning 404
MediumFinding: 81 pages link to https://help.cnn.com/, which returned 404 during the crawl.
Why it matters: A help link in a shared component that fails sends visitors to an error page from nearly every page on the site. See broken link checker.
| URL | Broken target |
|---|---|
https://www.cnn.com/ | https://help.cnn.com/ [404] |
https://www.cnn.com/account/settings | https://help.cnn.com/ [404] |
Recommended fix: Point the help link to the current support URL.
3. Very large HTML documents
MediumFinding: 81 pages exceed 500 KB of HTML. Sizes range from 3,099 KB to 5,209 KB, with a median of 4,321 KB. The homepage is 5,171 KB.
Why it matters: Multi-megabyte documents take longer to download and parse, especially on mobile networks, and often contain inlined data that could load separately.
| URL | HTML size |
|---|---|
https://www.cnn.com/ | 5,171 KB |
https://www.cnn.com/business | 5,124 KB |
https://www.cnn.com/account/settings | 3,099 KB |
Recommended fix: Audit the shared layout for inlined JSON, repeated navigation markup and embedded SVG that can be deferred.
4. Slow server responses
MediumFinding: 46 URLs took longer than 1,000 ms to respond. The median among them was 1,580 ms and the slowest was 3,397 ms.
Why it matters: Slow first responses delay page loads and reduce how many pages crawlers fetch per visit.
| URL | Result |
|---|---|
https://www.cnn.com/business/financial-calculators | 2,999 ms |
https://www.cnn.com/account/settings | 1,674 ms |
Recommended fix: Review caching on section fronts, starting with the slowest URLs in the export.
5. Missing H1 on the homepage and section fronts
LowFinding: 6 pages have no H1, including the homepage, /business, /style and /travel.
Why it matters: The H1 identifies the main topic of a page. Section fronts are hub pages where a clear heading helps crawlers and screen readers.
| URL | Evidence |
|---|---|
https://www.cnn.com/ | No H1 |
https://www.cnn.com/business | No H1 |
https://www.cnn.com/travel | No H1 |
Recommended fix: Add a visually hidden or visible H1 to the section front template.
6. Links to /sports, which redirects
LowFinding: 83 pages link to https://www.cnn.com/sports, which 301 redirects to /sport. 25 pages also link to the non-www https://cnn.com/.
Why it matters: Navigation links that pass through a redirect add a request on every click and every crawl.
| URL | Evidence |
|---|---|
https://www.cnn.com/ | Links to /sports [301] |
https://www.cnn.com/business | Links to /sports [301] |
Recommended fix: Update the navigation to link to /sport and the www homepage directly.
7. AI crawlers blocked and HSTS missing
LowFinding: robots.txt disallows all nine AI crawlers checked, including search agents OAI-SearchBot and Claude-SearchBot. No HSTS header was returned on the 84 analyzed pages, and no llms.txt was found.
Why it matters: Blocking search agents prevents citation in those AI products, not only training. HSTS tells browsers to always use HTTPS. See AEO / GEO audit.
Recommended fix: Review whether search and citation agents should be treated separately from training agents, and add a Strict-Transport-Security header.
What's working well
- 84 analyzed pages returned 200 with a title and meta description on every one.
- 81 pages have self-referencing canonicals and hreflang annotations.
- No 5xx server errors across 100 URLs.
- Structured data is present on 81 of 84 analyzed pages.
- No mixed content or missing viewport tags.
- Google Tag Manager is present on 82 of 84 pages.
Priority fixes
- Confirm crawler access to CNN Underscored14 URLs returned 403 and are in the sitemap. Pages affected: 14
- Fix the help.cnn.com footer linkOne component, 81 pages. Pages affected: 81
- Reduce HTML size on section templatesPages run 3 to 5 MB. Pages affected: 81
- Improve response times on section fronts46 slow URLs. Pages affected: 46
- Point navigation to /sport directlyRemoves a redirect on 83 pages. Pages affected: 83
Crawl data
| Metric | Result |
|---|---|
| URLs crawled | 100 |
| HTML pages | 98 |
| 2xx responses | 84 |
| 3xx redirects | 2 |
| 4xx responses (all 403) | 14 |
| 5xx responses | 0 |
| Sitemap URLs not returning 200 | 14 |
| Missing titles | 0 |
| Missing meta descriptions | 0 |
| Missing H1s | 6 |
| Multiple H1s | 2 |
| Missing og:description | 81 |
| Self-referencing canonicals | 81 |
| HTML over 500 KB | 81 |
| Slow responses (over 1,000 ms) | 46 |
| Missing HSTS | 84 |
| AI crawlers allowed | 0 of 9 |
On-page counts (titles, descriptions, headings, canonicals) are measured across the 84 URLs that returned an HTML page with a 200 status.
Audit limitations
The crawl was limited to 100 URLs and started from the homepage, so only URLs discovered through internal links and the XML sitemap are represented. A 100-URL crawl does not represent the entire website. Findings reflect the site at the time of the crawl on October 4, 2026. SiteAuditLint analyzed the HTML returned by the server, so content added later by JavaScript and responses that differ for automated clients may not match what a browser shows. This audit evaluates technical observations, not overall business or search performance.
This is an independent SiteAuditLint analysis of publicly accessible pages. It was not requested or endorsed by CNN, and it does not imply that CNN uses SiteAuditLint.