Home›Audits›CNN

Public brand audit · News publisher

CNN SEO Audit

A 100-page technical SEO crawl of cnn.com, conducted on October 4, 2026.

100 Pages crawled47s Crawl time28 Issue typesOct 4, 2026 Audit date

Independent SiteAuditLint analysis

SiteAuditLint crawled 100 URLs on cnn.com. Most critical items trace to one source: all 14 CNN Underscored URLs returned HTTP 403, and the sitewide navigation links to them from 81 pages. Page weight is the other large pattern, with section pages between 3 and 5 MB of HTML.

MetricResult
Pages crawled100
Crawl time47s
Health score82 / 100 (Good)
Issue types detected28
Critical / warnings / opportunities / info95 / 156 / 512 / 7
Audit dateOctober 4, 2026

What stood out

14CNN Underscored URLs returned HTTP 403, all listed in the XML sitemap
81pages link to help.cnn.com, which returned 404
4.3 MBmedian HTML size across 81 pages, with the largest at 5.2 MB
46URLs responded in over 1 second, the slowest at 3,397 ms
0 / 9major AI crawlers allowed in robots.txt

Overall assessment

cnn.com's core news pages loaded with 200 status codes, self-referencing canonicals and hreflang. The critical count of 95 comes almost entirely from one section: every /cnn-underscored/ URL returned HTTP 403 to the crawler, and the global navigation links to them, so 81 pages register as linking to a broken internal page.

A 403 limited to one section points to access rules on that section rather than missing pages. It is still worth confirming, because those URLs also appear in the XML sitemap.

Two sitewide patterns deserve attention after that. A help.cnn.com link in the shared footer returned 404 on 81 pages, and section pages ship between 3.1 and 5.2 MB of HTML, which is roughly ten times the 500 KB threshold SiteAuditLint uses.

Audit methodology

SiteAuditLint started from the website's homepage and crawled up to 100 URLs by following discoverable internal links and the URLs listed in the XML sitemap. Results represent the URLs reached during the crawl and do not constitute a complete audit of the entire website.

Starting URLhttps://cnn.com/
Crawl limit100 URLs
URLs crawled100 (98 HTML)
Crawl dateOctober 4, 2026, 18:04 UTC
Crawl duration47s
Crawl methodHomepage start, internal link and sitemap discovery, robots.txt respected, raw HTML analysis
ToolSiteAuditLint Desktop SEO Crawler
ScopePublicly accessible URLs reached during the crawl

Technical SEO scorecard

Scores are the category scores SiteAuditLint calculates as part of its Site Health Score.

AreaScoreKey finding
Overall health82Good rating from SiteAuditLint
Technical8414 Underscored URLs returned 403
Content962 near-duplicate market pages, 1 thin page
Performance8846 slow responses, 81 pages over 500 KB
AEO93robots.txt blocks all 9 AI crawlers checked

Major findings

1. CNN Underscored URLs returning 403

High

14 affected URLs · SiteAuditLint severity: Critical

Finding: All 14 URLs under /cnn-underscored/ returned HTTP 403. All 14 are in the XML sitemap, and the global navigation links to them from 81 pages.

Why it matters: A 403 tells clients they may not fetch the page. If search engine crawlers receive the same response, the section cannot be crawled despite being in the sitemap. See status codes.

URLResult
https://www.cnn.com/cnn-underscoredHTTP 403
https://www.cnn.com/cnn-underscored/dealsHTTP 403
https://www.cnn.com/cnn-underscored/electronicsHTTP 403

Recommended fix: Check server or CDN logs to confirm verified search engine crawlers receive 200 for /cnn-underscored/ URLs.

2. Sitewide link to help.cnn.com returning 404

Medium

81 affected pages · SiteAuditLint severity: Warning

Finding: 81 pages link to https://help.cnn.com/, which returned 404 during the crawl.

Why it matters: A help link in a shared component that fails sends visitors to an error page from nearly every page on the site. See broken link checker.

URLBroken target
https://www.cnn.com/https://help.cnn.com/ [404]
https://www.cnn.com/account/settingshttps://help.cnn.com/ [404]

Recommended fix: Point the help link to the current support URL.

3. Very large HTML documents

Medium

81 affected pages · SiteAuditLint severity: Opportunity

Finding: 81 pages exceed 500 KB of HTML. Sizes range from 3,099 KB to 5,209 KB, with a median of 4,321 KB. The homepage is 5,171 KB.

Why it matters: Multi-megabyte documents take longer to download and parse, especially on mobile networks, and often contain inlined data that could load separately.

URLHTML size
https://www.cnn.com/5,171 KB
https://www.cnn.com/business5,124 KB
https://www.cnn.com/account/settings3,099 KB

Recommended fix: Audit the shared layout for inlined JSON, repeated navigation markup and embedded SVG that can be deferred.

4. Slow server responses

Medium

46 affected URLs · SiteAuditLint severity: Warning

Finding: 46 URLs took longer than 1,000 ms to respond. The median among them was 1,580 ms and the slowest was 3,397 ms.

Why it matters: Slow first responses delay page loads and reduce how many pages crawlers fetch per visit.

URLResult
https://www.cnn.com/business/financial-calculators2,999 ms
https://www.cnn.com/account/settings1,674 ms

Recommended fix: Review caching on section fronts, starting with the slowest URLs in the export.

5. Missing H1 on the homepage and section fronts

Low

6 affected pages · SiteAuditLint severity: Warning

Finding: 6 pages have no H1, including the homepage, /business, /style and /travel.

Why it matters: The H1 identifies the main topic of a page. Section fronts are hub pages where a clear heading helps crawlers and screen readers.

URLEvidence
https://www.cnn.com/No H1
https://www.cnn.com/businessNo H1
https://www.cnn.com/travelNo H1

Recommended fix: Add a visually hidden or visible H1 to the section front template.

6. Links to /sports, which redirects

Low

83 affected pages · SiteAuditLint severity: Opportunity

Finding: 83 pages link to https://www.cnn.com/sports, which 301 redirects to /sport. 25 pages also link to the non-www https://cnn.com/.

Why it matters: Navigation links that pass through a redirect add a request on every click and every crawl.

URLEvidence
https://www.cnn.com/Links to /sports [301]
https://www.cnn.com/businessLinks to /sports [301]

Recommended fix: Update the navigation to link to /sport and the www homepage directly.

7. AI crawlers blocked and HSTS missing

Low

9 crawlers blocked · SiteAuditLint severity: Warning

Finding: robots.txt disallows all nine AI crawlers checked, including search agents OAI-SearchBot and Claude-SearchBot. No HSTS header was returned on the 84 analyzed pages, and no llms.txt was found.

Why it matters: Blocking search agents prevents citation in those AI products, not only training. HSTS tells browsers to always use HTTPS. See AEO / GEO audit.

Recommended fix: Review whether search and citation agents should be treated separately from training agents, and add a Strict-Transport-Security header.

What's working well

  • 84 analyzed pages returned 200 with a title and meta description on every one.
  • 81 pages have self-referencing canonicals and hreflang annotations.
  • No 5xx server errors across 100 URLs.
  • Structured data is present on 81 of 84 analyzed pages.
  • No mixed content or missing viewport tags.
  • Google Tag Manager is present on 82 of 84 pages.

Priority fixes

  1. Confirm crawler access to CNN Underscored14 URLs returned 403 and are in the sitemap. Pages affected: 14
  2. Fix the help.cnn.com footer linkOne component, 81 pages. Pages affected: 81
  3. Reduce HTML size on section templatesPages run 3 to 5 MB. Pages affected: 81
  4. Improve response times on section fronts46 slow URLs. Pages affected: 46
  5. Point navigation to /sport directlyRemoves a redirect on 83 pages. Pages affected: 83

Crawl data

MetricResult
URLs crawled100
HTML pages98
2xx responses84
3xx redirects2
4xx responses (all 403)14
5xx responses0
Sitemap URLs not returning 20014
Missing titles0
Missing meta descriptions0
Missing H1s6
Multiple H1s2
Missing og:description81
Self-referencing canonicals81
HTML over 500 KB81
Slow responses (over 1,000 ms)46
Missing HSTS84
AI crawlers allowed0 of 9

On-page counts (titles, descriptions, headings, canonicals) are measured across the 84 URLs that returned an HTML page with a 200 status.

Audit limitations

Scope of this audit

The crawl was limited to 100 URLs and started from the homepage, so only URLs discovered through internal links and the XML sitemap are represented. A 100-URL crawl does not represent the entire website. Findings reflect the site at the time of the crawl on October 4, 2026. SiteAuditLint analyzed the HTML returned by the server, so content added later by JavaScript and responses that differ for automated clients may not match what a browser shows. This audit evaluates technical observations, not overall business or search performance.

This is an independent SiteAuditLint analysis of publicly accessible pages. It was not requested or endorsed by CNN, and it does not imply that CNN uses SiteAuditLint.

Audit record

Audit dateOctober 4, 2026
Crawl started18:04 UTC
Starting URLcnn.com
URLs crawled100
Crawl duration47s
Health score82 / 100