Training archive
Submit Your Website to Search Engines
You do not need a paid submission service to make a website discoverable. Search engines usually find pages by following links, revisiting known addresses

You do not need a paid submission service to make a website discoverable. Search engines usually find pages by following links, revisiting known addresses and reading XML sitemaps. Submission gives them another route to your content, but it cannot make a blocked, broken or unsuitable page appear in results.
The practical process is simple: make each important page accessible, choose one canonical address, connect the site to a webmaster portal, submit a clean sitemap and inspect any page that remains missing. Work through the technical checks before requesting another crawl.
Understand What Submission Can Do
Discovery, crawling, indexing and ranking are separate stages. A crawler first discovers an address, then requests it. The search engine processes the response and decides whether the page belongs in its index. Only indexed pages are eligible to appear for a query, and inclusion does not guarantee a prominent position. This source provides further background on the stages.
A submission is therefore a discovery signal, not an instruction to index. A search engine may omit a submitted page because it is unavailable, marked for exclusion, substantially duplicates another page or offers too little distinct information. Repeatedly submitting the same unchanged address does not resolve those causes.
Manual submission is most useful for a new site, a newly published section, a page moved to a different address or an important revision that has not yet been revisited. For a whole site, submit a sitemap. For a small number of priority pages, use the portal's individual inspection or recrawl request after checking each page.
Prepare the Site Before Submission
Confirm the Response and Page Content
Open each priority address without signing in and confirm that it returns the intended page. An indexable page should normally return a successful response rather than an error or redirect. It should contain its useful text in the initial document or in a form that ordinary crawlers can render reliably.
Check both the page markup and the server response headers for a noindex directive. Keep that directive on private areas, internal search results, account pages and deliberate duplicates. Remove it only when the page is genuinely intended for public search results.
Review the robots file at the site's standard robots address. A broad disallow rule can prevent crawlers from reaching an entire section, while a narrower rule may accidentally cover scripts, styles or other resources needed to understand the page. Submission cannot override a crawl restriction (source).
Use One Canonical Version
Select one secure hostname and one consistent format for each page. Internal links, sitemap entries, canonical tags and redirects should all point to that version. Avoid chains in which an old address redirects through several intermediate addresses before reaching the final page.
A self-referencing canonical tag is appropriate for an original indexable page. A close duplicate should identify the preferred equivalent instead. Do not list both versions in the sitemap or alternate between them in navigation. Conflicting signals make it harder for a crawler to identify the address you want indexed.
Strengthen Internal Discovery
Every important page should be reachable through ordinary text links. Group related pages beneath clear categories, link categories from the main navigation and add contextual links where another page gives the reader a useful next step. Descriptive anchor text is more informative than repeated labels such as click here.
Find orphan pages by comparing the site's crawlable internal links with its sitemap. A sitemap may expose an orphan address, but it cannot explain the page's relationship to the rest of the site. Add a relevant internal link or reconsider whether the page deserves to remain indexable.
Create a Clean XML Sitemap
A sitemap is a machine-readable list of the canonical pages you want search engines to discover. Each entry should use an absolute address and the site's preferred secure hostname. Include only pages that return a successful response, permit indexing and identify themselves as canonical.
Exclude redirected addresses, error pages, blocked pages, deliberate duplicates, private content and pages marked noindex. If a page moves, replace the old sitemap entry with the final address after confirming that the redirect, canonical tag and internal links agree. Useful summaries of XML sitemap crawling and indexing practices explain why a sitemap supports discovery without guaranteeing inclusion.
Large sites can divide entries by content type or section and list the separate files in a sitemap index. This makes errors easier to isolate. Record a last-modified date only when the main content changed; automatically rewriting every date on every build gives crawlers little useful information.
Publish the sitemap at a stable public address. You may also declare that address in the robots file. Open the sitemap in a browser or validator before submission and check for malformed XML, non-canonical hosts, unexpected parameters and addresses outside the intended site.
Submit and Inspect the Site
Create a property in each major search engine's webmaster portal and prove that you control the site. Domain-level verification usually covers protocol and subdomain variations, while address-prefix verification covers only the stated prefix. Choose the scope that matches the canonical site and keep the verification record in place.
Submit the sitemap through the portal's sitemap report. A successful submission means the service could accept or fetch the file; it does not mean every entry has been crawled or indexed. Return later to check whether the file was read and whether any entries caused parsing, access or address errors.
Use the inspection report for the homepage, key category pages and newly published or corrected pages. Check the live response, permitted indexing state and selected canonical address. If the report refers to an older version, compare its recorded crawl with the current live test before deciding what to change.
Request another crawl only after a meaningful publication or correction. Suitable cases include a new page, removal of an accidental noindex directive, repair of a server error or a substantial content update. Minor wording and formatting changes can wait for normal recrawling. A separate website indexing guide offers more detail on how crawlers process eligible pages.
Diagnose Pages That Remain Missing
Start with the reported exclusion reason and verify it against the live page. Do not assume every excluded address is a fault: redirected pages, duplicates and intentionally private pages should stay out of the index. Concentrate on canonical pages that users can reach and that you deliberately submitted.
- Discovered but not crawled: confirm stable server performance, useful internal links and an accurate sitemap entry. Avoid creating large numbers of low-value parameter addresses.
- Crawled but not indexed: compare the page with similar pages. Consolidate genuine duplicates, clarify the purpose and add information that is specific to the page rather than repeating a template.
- Blocked: inspect robots rules, page-level directives, response headers, authentication and access controls. Change only the rule that conflicts with the page's intended visibility.
- Alternate canonical: check whether the engine selected a reasonable equivalent. If not, align internal links, redirects, canonical tags and sitemap entries around one address.
- Error or soft error: repair the response and ensure the page contains meaningful content. Remove permanently deleted addresses from the sitemap and fix links that still point to them.
After a correction, test the final address, update the sitemap if necessary and request a crawl for the most important page. Keep a short log of the address, cause, change and validation date. This prevents teams from repeating ineffective submissions without knowing what changed.
Monitor Discovery Without Chasing Daily Fluctuations
Review the webmaster reports after launches, migrations and material site changes. Compare submitted addresses with indexed canonical pages, but investigate examples rather than treating the totals as exact equivalents. A sitemap can legitimately contain pages that are later consolidated or excluded.
Watch for patterns by section. A cluster of missing pages may reveal a shared template directive, navigation gap, server fault or canonical error. One isolated page usually calls for page-level inspection. Search appearance data becomes useful only after indexing, so an absence of impressions is not by itself proof of a submission failure.
- Keep navigation and contextual links crawlable.
- Update sitemap entries when canonical pages are added, moved or removed.
- Check exclusion and error reports after technical releases.
- Inspect representative pages from each important section.
- Request recrawling only when the live page has materially changed.
