Extract and count the internal links of a page
extract all links from a page | check internal linking command line | which pages are linked | extract links from html curl | curl | links | internal-linking | seocurl -s https://example.com/ | tr '\n' ' ' | grep -oiE 'href="[^"]+"' | sed -E 's/^href="//I; s/"$//; s/#.*//' | grep -E '^(/|https?://(www\.)?example.com)' | sort | uniq -c | sort -rnPulls all href targets from the HTML, removes anchors and keeps only links that start with / or point to the site’s own domain. The frequency shows which pages are linked especially often from navigation and footer. Watch out for links to http variants, staging domains or old URLs that redirect.
Note: Relative links without a leading slash are missed, and stylesheets from <link href> are included. The list can be fed into the status code check from the sitemap audit.
Also searched as
- extract all links from a page
- check internal linking command line
- which pages are linked
- extract links from html curl