The Moz Q&A Forum

    • Forum
    • Questions
    • My Q&A
    • Users
    • Ask the Community

    Welcome to the Q&A Forum

    Browse the forum for helpful insights and fresh discussions about all things SEO.

    1. SEO and Digital Marketing Q&A Forum
    2. Categories
    3. Intermediate & Advanced SEO
    4. XML Sitemap Index Percentage (Large Sites)

    XML Sitemap Index Percentage (Large Sites)

    Intermediate & Advanced SEO
    2 2 832
    • Oldest to Newest
    • Newest to Oldest
    • Most Votes
    Reply
    • Reply as question
    Log in to reply
    This topic has been deleted. Only users with topic management privileges can see it.
    • danng
      danng last edited by

      Hi all

      I'm wanting to find out from those who have experience dealing with large sites (10s/100s of millions of pages).

      What's a typical (or highest) percentage of indexed pages vs. submitted pages you've seen? This information can be found in webmaster tools where Google shows you the pages submitted & indexed for each of your sitemap.

      I'm trying to figure out whether,

      • The average index % out there
      • There is a ceiling (i.e. will never reach 100%)
      • It's possible to improve the indexing percentage further

      Just to give you some background, sitemap index files (according to schema.org) have been implemented to improve crawl efficiency and I'm wanting to find out other ways to improve this further.

      I've been thinking about looking at the URL parameters to exclude as there are hundreds (e-commerce site) to help Google improve crawl efficiency and utilise the daily crawl quote more effectively to discover pages that have not been discovered yet.

      However, I'm not sure yet whether this is the best path to take or I'm just flogging a dead horse if there is such a ceiling or if I'm already at the average ballpark for large sites.

      Any suggestions/insights would be appreciated. Thanks.

      1 Reply Last reply Reply Quote 0
      • Sam_McRoberts
        Sam_McRoberts last edited by

        I've worked on a site that was ~100 million pages, and I've seen indexation percentages ranging from 8% to 95%. When dealing with sites this size, there are so, so many issues at play, and there are so few sites of this size that finding an average probably won't do you much good.

        Rather than focusing on whether or not you have enough pages indexed based on averages, you should focus on two key questions: "do my sitemaps only include pages that would make great search engine entry pages" and "have I done everything possible to eliminate junk pages that are wasting crawl bandwidth."

        Of course, making sure you don't have any duplicate content, thin content, or poor on-site optimization issues should also be a focus.

        I guess what I'm trying to say is, I believe any site can have 100% of it's search entry worthy pages indexed, but sites of that size rarely have ALL of their pages indexed since sites that large often have a ton of pages that don't make great search results.

        1 Reply Last reply Reply Quote 0
        • 1 / 1
        • First post
          Last post
        • Sitemap Indexed Pages, Google Glitch or Problem With Site?
          rrhansen
          rrhansen
          0
          5
          653

        • Mobile First Index: What Could Happen To Sites w Large Desktop but Small Mobile Sites?
          mirabile
          mirabile
          0
          4
          274

        • Xml sitemap Issue... Xml sitemap generator facilitating only few pages for indexing
          Paddy_Moogan
          Paddy_Moogan
          0
          6
          153

        • Google Processing but Not Indexing XML Sitemap
          Kingof5
          Kingof5
          0
          4
          1.5k

        • XML Sitemap Indexation Rate Decrease
          DanDeceuster
          DanDeceuster
          0
          3
          210

        • Google Not Indexing XML Sitemap Images
          bogdan_c
          bogdan_c
          0
          17
          11.4k

        • How Do I Generate a Sitemap for a Large Wordpress Site?
          FedeEinhorn
          FedeEinhorn
          0
          6
          2.6k

        • Xml Sitemap for a large automobile website
          Matthew_Edgar
          Matthew_Edgar
          0
          2
          649

        Get started with Moz Pro!

        Unlock the power of advanced SEO tools and data-driven insights.

        Start my free trial
        Products
        • Moz Pro
        • Moz Local
        • Moz API
        • Moz Data
        • STAT
        • Product Updates
        Moz Solutions
        • SMB Solutions
        • Agency Solutions
        • Enterprise Solutions
        • Digital Marketers
        Free SEO Tools
        • Domain Authority Checker
        • Link Explorer
        • Keyword Explorer
        • Competitive Research
        • Brand Authority Checker
        • Local Citation Checker
        • MozBar Extension
        • MozCast
        Resources
        • Blog
        • SEO Learning Center
        • Help Hub
        • Beginner's Guide to SEO
        • How-to Guides
        • Moz Academy
        • API Docs
        About Moz
        • About
        • Team
        • Careers
        • Contact
        Why Moz
        • Case Studies
        • Testimonials
        Get Involved
        • Become an Affiliate
        • MozCon
        • Webinars
        • Practical Marketer Series
        • MozPod
        Connect with us

        Contact the Help team

        Join our newsletter
        Moz logo
        © 2021 - 2026 SEOMoz, Inc., a Ziff Davis company. All rights reserved. Moz is a registered trademark of SEOMoz, Inc.
        • Accessibility
        • Terms of Use
        • Privacy