1. Semrush
  2. Blog
  3. SEO
  4. General SEO

How to get your website indexed by Google

Expert reviewed

Our subject matter experts have reviewed this article to ensure it meets the highest standard for accurate information and guidance. Learn more about our editorial standards and process.

Author:Carlos Silva

10 min read

June 2, 2026

Contributor: Christine Skopec

Getting your website indexed by Google is necessary if you want to appear in Google’s organic or AI search results.

Today, we’ll show you different ways to confirm if Google has indexed your website. We’ll also cover common indexing issues like:

After reading, you’ll know how to find and fix indexing issues and confirm whether Google has indexed your most important pages.

What is the Google index?

The Google index is a massive database of webpages that Google has crawled.

The index is a structured database that allows Google to instantly match search queries with relevant results. This means if your webpages aren’t in Google's index, they won’t appear in organic search results, AI Overviews, AI Mode, or Gemini.

Being absent from Google’s index could even impact your visibility in AI tools like ChatGPT. We know that those AI systems rely on Google at least some of the time.

The indexing process follows this sequence when no issues occur:

While Google’s own algorithms control indexing, website owners can take steps to influence the process.

How do you check if Google has indexed your site?

Check if Google has indexed your site with the "site:search" operator or using Google Search Console.

Use "site:search" operator

The "site:search" operator displays indexed pages from a particular website in search results.

Here’s how to use to to see if your own pages are indexed:

  1. Go to Google
  2. Type "site:[yourdomain.com]" in the search bar

After searching, you'll see indexed pages as search results. To see the total number, click the “Tools” drop-down to see an approximate number of results. Zero results indicate no indexed pages.

Google search results for site:backlinko.com with indexed page count highlighted in Tools menu

While the "site:search" operator works for identifying whether your pages are indexed, it doesn’t allow you to identify pages that haven’t been indexed. You’ll need to identify those pages using Google Search Console (GSC).

Use Google Search Console

Google Search Console’s "Page indexing" report shows you which pages on your site are indexed and which ones aren’t.

Open your GSC account and head to "Pages" (under "Indexing"). Click "View data about indexed pages" for a sample list of indexed pages.

Tip

Consult the Google Search Console setup guide for detailed guidance.

Google Search Console Page Indexing report with "View data about indexed pages" section highlighted

The "Indexed pages" report may not show all indexed pages if you exceed the limit of 1,000 items. Or if something was added after the most recent crawl.

Google Search Console Indexed Pages report showing 91 indexed pages and example URLs

Go back to the "Page indexing" report to view pages that aren’t indexed by scrolling down. In that table, GSC lists reasons why your pages aren’t indexed. Click a reason to see a list of affected pages.

Google Search Console report showing reasons pages aren’t indexed, including robots.txt blocks

Each status corresponds to a specific problem. The table below explains some common Google Search Console errors related to indexation and what to do about each one.

Status
What it means
What to do
Discovered – currently not indexed
Google knows the page exists but hasn't crawled it yet. This often happens when Google thinks crawling the page will overload the site.
Request indexing, strengthen internal linking to the page, or minimize duplicate/thin pages consuming crawl budget
Crawled – currently not indexed
Google visited the page but chose not to index it. This often signals a quality problem.
Improve page quality by adding original content and ensuring the page fully answers readers’ questions
Blocked by robots.txt
A robots.txt(a file that tells bots what they should and shouldn’t crawl) directive is telling Googlebot not to crawl the URL
Open your robots.txt file and check for rules telling crawlers to avoid the page. Remove or adjust the rule if the page should be indexed.
Duplicate, Google chose different canonical than user
Google found multiple versions of this page and decided a different URL is the main version
Ensure you’ve used canonical tags on all versions that point to your preferred URL
Excluded by 'noindex' tag
A <meta name="robots" content="noindex"> tag in the HTML is explicitly telling Google not to index the page
Remove the noindex tag from the page's source code if you want it indexed
Not found (404)
The URL returns a 404 error, which means the page doesn't exist at this address
Restore the page if deleted, correct the URL if wrong, or set up a 301 redirect (a permanent redirect) to the current version of the content

How do you get Google to index your site?

You don’t need to do anything aside from wait for Google to index your site, but you can speed up the process by creating and submitting a sitemap or by using the URL inspection tool in Google Search Console.

Create and submit a sitemap

Creating and submitting a sitemap — a file that includes all your important URLs and indicates how they relate to each other — helps crawlers find your priority pages more quickly.

A sitemap looks something like this:

Semrush Sitemap index file showing URLs in XML format

If you don’t know your sitemap URL, find it by reviewing your robots.txt file. Enter your "https://[yourdomain.com]/robots.txt" and look for your sitemap URL (you might have to scroll down).

Browser view of a robots.txt file with sitemap URL highlighted

If you lack a sitemap, consult our guide for creating an XML sitemap.

To submit your sitemap in GSC:

  1. Navigate to "Sitemaps" under the "Indexing" section in GSC's menu
  2. Enter your sitemap URL under "Add a new sitemap"
  3. Click "Submit"

Google Search Console Sitemaps page with sitemap_index.xml submission field highlighted

Processing typically takes a couple of days. Upon completion, you'll see your sitemap link with a green "Success" status.

Submitted sitemap report in Google Search Console showing successful sitemap status

Use the URL inspection tool

The URL inspection tool in GSC allows you to request indexation for a specific page.

Enter the URL in the top search bar in GSC and press enter. If you see “URL is on Google” near the top, it means the specified page has been indexed already. You can also see information about when Google last crawled the page, whether the page is Google’s selected canonical, and whether the page is your specified canonical.

Google Search Console URL Inspection report showing page is indexed and on Google

A "URL is not on Google" status means the URL isn't indexed and won't appear in search results. Review the provided reason and address the issue.

Google Search Console URL Inspection report showing page is crawled but not indexed

After addressing the issue listed, click the "Request Indexing" link to ask Google to prioritize crawling it. This doesn’t guarantee immediate indexing, but Google typically processes these requests within a few weeks. Periodically check the page with the URL inspection tool to confirm Google has indexed the page.

Google Search Console URL Inspection page with Request Indexing button highlighted

Common indexing issues to find and fix

Common indexing issues to find and fix include errors in your robots.txt file, lack of mobile usability, slow loading speeds, and redirect issues.

Find indexing issues specific to your site with Semrush Site Audit tool. After configuring Site Audit, click "Issues" and filter the issues by "Crawlability" to see issues that prevent search engines from crawling your site.

Click a specific error to see the affected pages, and "How to fix" for tips on resolving each error.

Semrush Site Audit report filtered for Crawlability issues with broken internal links issue details expanded

Let’s go over some of the most common indexing issues in greater detail:

Mistakes with your robots.txt file

Mistakes with your robots.txt file can tell Google to avoid crawling certain pages or even your entire site.

The robots.txt file below tells one bot to avoid crawling the entire site. If that directive targeted Googlebot instead, Google would avoid crawling the site.

Robots.txt file showing rules allowing and disallowing specific user agents from crawling the site

Find your robots.txt at “https://[yourdomain.com]/robots.txt.” Consult our robots.txt guide if you lack one and need directions on how to create one.

You can use directives to tell crawlers to avoid duplicate pages, private content, or resource files. However, if your robots.txt tells bots to avoid crawling completely, indexing is highly unlikely.

Here’s an example that tells all bots to avoid crawling the entire website:

User-agent: *Disallow: /

So, review your robots.txt to ensure no directive prevents Google from crawling pages you want indexed.

Accidental use of noindex tags

Accidentally using the "noindex" robots meta tag (an HTML tag within a page) tells crawlers not to index a page.

A noindex tag looks like this:

<meta name="robots" content="noindex">

Check which pages have noindex tags in GSC:

  1. Click "Pages" under "Indexing" in the left menu
  2. Scroll to "Why pages aren't indexed"
  3. Click "Excluded by 'noindex' tag" if present

Google Search Console report highlighting pages excluded by noindex tag

Remove the noindex tag from any pages in the list that you want to appear in Google’s index.

Site Audit warns about pages blocked via robots.txt or noindex.

Semrush Site Audit notice showing pages blocked from crawling

Site Audit also notifies you about resources that are blocked by x-robots-tag, which is typically used for non-HTML documents like PDFs.

Site Audit report showing X-Robots-Tag noindex HTTP header notice

Improper canonical tags

Improper canonical tags that point Google to the wrong URL can prevent your intended page from appearing in search results.

Find improper canonical tags within GSC's "Page indexing" report:

  1. Scroll to "Why pages aren't indexed"
  2. Click "Alternate page with proper canonical tag"

Google Search Console report showing alternate page with proper canonical tag reason

Review the affected pages list. If there’s a page you want to have indexed (meaning the canonical is used incorrectly), adjust the canonical tags on all versions of the page to point to your preferred version.

Internal link problems prevent crawlers from discovering pages, which can keep those pages out of Google's index.

Find internal linking issues in Site Audit’s “Internal Linking” thematic report. You’ll see a list of internal linking issues. Click any issue count link to see affected pages.

Semrush Internal Linking report showing broken links and crawl depth issues

These are some of the most important issues to address when it comes to crawling and indexing:

  1. Nofollow attributes in outgoing internal links: Nofollow links generally tell Google not to follow a link or pass authority to it, so Google might ignore pages on your site if you’ve used nofollow links to them internally
  2. Page Crawl Depth more than 3 clicks: If pages need more than three clicks to be reached from the homepage, there's a chance they won't be crawled and indexed. Add more internal links to these pages (and review your website architecture).
  3. Orphaned sitemap pages: Pages that have no internal links pointing to them are known as "orphaned pages." They’re rarely indexed as Google may struggle to find them. Fix this issue by linking to any orphaned pages.

When building internal links, prioritize linking to your most important pages. And also actively work to link to new pages to accelerate indexing.

404 errors

A 404 error occurs when a server can’t locate a page, and it prevents Google from finding and indexing pages.

Plus, 404 errors harm the user experience.

Find your site’s 404 errors within Site Audit’s "Issues" tab. Click the link in "# pages returned a 4XX status code."

Semrush Site Audit issues report highlighting pages returning 4XX status codes

For each "404" page, click "View broken links" to see pages linking to it.

"View broken links" highlighted

Fix 404 errors by correcting URL typos, updating links to new page locations, or replacing links with relevant substitutes if content no longer exists.

Duplicate content

Duplicate content — identical or very similar content across multiple URLs — confuses search engines and may result in undesired pages being indexed.

Click "Issues" in Site Audit and search for "duplicate." Click the hyperlink in "# pages have duplicate content issues."

Semrush Site Audit issues filtered for duplicate content problems

Fix duplicate content issues by:

Poor site quality

Poor site quality can hurt your chances of being indexed as Google prioritizes crawling and indexing sites it deems high quality.

Here are three ways to make your site appear trustworthy to Google:

Create high-quality content

Creating high-quality content that genuinely helps readers improves your chances of being indexed and shown in search results.

Follow these tips for creating quality content:

Building relevant backlinks from quality websites that are relevant to you provides more ways for Google to discover your pages and also signals authority.

Here are some link building tactics:

Use Backlink Gap to do a competitor backlink analysis. Just enter your domain and up to four competitors' domains, then click "Find prospects"

Semrush Backlink Gap tool start with 5 domains entered and arrow pointing to Find prospects button

The "Best" tab within Backlink Gap shows websites linking to all competitors but not you. These sites are often worth pitching. There’s a good chance they’ll link to you if they’re already linking to all your rivals.

Prospects for table with Referring Domain column highlighted

Prioritize E-E-A-T

Focusing on Experience, Expertise, Authority, and Trustworthiness (E-E-A-T) — the criteria Google's human quality raters use to assess page quality — helps you align with what Google defines as good content.

E-E-A-T is not a Google ranking factor, but following the E-E-A-T framework helps you create good content.

To strengthen your E-E-A-T, aim to:

Monitor your site for indexing issues

Monitor your site for indexing issues by scheduling periodic audits that let you check your site for any issues as soon as they pop up.

With Site Audit, you can schedule audits weekly or daily, so you’re alerted of new issues right away.

Semrush Site Audit settings with weekly crawl schedule dropdown open

Ready to find and fix indexing issues? Try Site Audit today.

Find and fix indexing issues

with the Site Audit tool

Sign up now

Share

Author Photo

Carlos Silva

Carlos Silva leads the editorial pipeline for the English blog—coordinating writers, editors, strategy, and building AI workflows that boost content quality and AI visibility. With 10+ years of experience writing and editing across in-house and agency roles, Carlos blends content strategy, SEO, and AI to help marketers stay ahead of an evolving search landscape.

Author Photo

Carlos Silva

Carlos Silva leads the editorial pipeline for the English blog and builds AI workflows that boost content quality and AI visibility. 10+ years writing and editing across in-house and agency roles.

Try Semrush
Everything you need to win AI visibility and drive SEO success
Start Free Trial
Unlimited access for 7 days
Boost your digital marketing efforts
Get free trial