Indexing

作者 kostja948dd89c57395e無授權條款1K 個星標收錄於 2026年10月8日更新於 2026年10月8日儲存庫2 天前更新

When the user wants to fix indexing issues from Search Console, use noindex, or implement Google Indexing API. Also use when the user mentions "fix indexing," "not indexed," "Crawled - currently not indexed," "discovered - currently not indexed," "index coverage," "noindex," "noindex tag," "pages not indexed," "why not indexed," "request indexing," or "Google Indexing API." For sitemap, use xml-sitemap.

僅含說明Marketing & Sales
AI 產生的概覽

指導技術性 SEO 索引修正:noindex 用法、Search Console 收錄問題以及 Google Indexing API。

功能
此技能提供診斷與修正搜尋索引問題的指引,涵蓋 Search Console 收錄問題類型,例如已檢索但未建立索引、軟 404 和 noindex 標籤。它說明何時使用 noindex 而非 robots.txt,提供頁面移除決策框架及重新導向與狀態碼選擇,並概述 Google Indexing API 的需求。產出是依優先順序排列的行動項目與參考資料,而非產生的檔案或程式碼。
適用情境
當網站在 Search Console 出現索引問題(例如頁面被回報為未建立索引或已檢索但未建立索引)時使用。也適用於決定是否套用 noindex、移除或重新導向頁面,或設定 Google Indexing API 的情況。
執行需求
不包含指令碼,僅為說明性內容。它假定可存取 Google Search Console 進行診斷;就 Indexing API 部分而言,還需要服務帳戶、Search Console 中的擁有者權限以及 API 配額。

SEO Technical: Indexing

Guides indexing troubleshooting and fix actions. For how to find and diagnose issues in GSC, see google-search-console.

When invoking: On first use, if helpful, open with 1–2 sentences on what this skill covers and why it matters, then provide the main output. On subsequent use or when the user asks to skip, go directly to the main output.

Scope (Technical SEO)

  • Fix actions: noindex, canonical, content quality, URL Inspection; verify robots.txt does not block (see robots-txt)
  • Noindex: Page-level index control; which pages to exclude and how. Complements robots-txt (path-level crawl control) and google-search-console (Coverage diagnosis)

Initial Assessment

Project context: Read root contextus.md when present and load only the modules relevant to this task. Without Contextus, use available project material or user-provided facts and ask for missing information; do not create a parallel context system.

Identify issue from GSC (see google-search-console for Coverage report, issue types, diagnosis workflow). Then apply fix below.

Crawled - Currently Not Indexed

CauseAction
Low quality, duplicate, off-topicImprove content, fix duplicates, set correct canonical
Static assets (CSS/JS)See below
Feed, share URLs with paramsUsually OK to ignore; or noindex, canonical to main URL
Important content pagesUse URL Inspection, verify canonical/internal links/sitemap, Request indexing

Static Assets (Next.js / Vercel)

Vercel adds unique dpl= params to static assets per deploy, creating many "Crawled - currently not indexed" URLs.

DoDon't
Keep robots.txt allowing /_next/Do not block /_next/ (breaks CSS/JS loading). See robots-txt
Accept static assets in GSC as expectedDo not block /_next/static/css/ or ?dpl=
Use X-Robots-Tag for static assetsCSS/JS should not be indexed; no SEO impact

Static assets in "Crawled - currently not indexed" is normal and expected.

Other Issue Types (from GSC Coverage)

IssueFix
Excluded by «noindex» tagRemove noindex if accidental; keep if intentional
Blocked by robots.txtSee robots-txt; remove Disallow for important paths
Redirect / 404Fix URL or add redirect
Duplicate / CanonicalSet correct canonical; usually OK
Soft-404Page returns 200 but content says "not found" or empty—Google may treat as 404. Fix: return 404 status for truly missing pages; or add real content for 200 pages

Soft-404

A soft-404 occurs when a page returns HTTP 200 but the content indicates the page doesn't exist (e.g. "Page not found" message, empty state). Google may treat it as 404 and exclude from index.

FixWhen
Return 404Page truly doesn't exist; use proper 404 status
Add contentPage is intentional (e.g. empty search results); ensure substantive content or use noindex
RedirectIf URL moved, use 301 to correct destination

Noindex Usage

  • How: metadata.robots = { index: false } or <meta name="robots" content="noindex"> or X-Robots-Tag
  • Rationale: Not all site content should be indexed; noindex is a valid choice for many pages
  • Caution: Avoid noindex on important content pages
  • With robots.txt: robots.txt = path-level crawl control; noindex = page-level index control. Do not block noindex pages in robots.txt—crawlers must access the page to read the directive. Use both: robots for /admin/, /api/; noindex for /login/, /thank-you/, etc. See robots-txt for when to use which.
  • nofollow ≠ noindex: nofollow controls link equity only; it does not prevent indexing. To exclude from search, use noindex. See page-metadata for meta robots implementation.

Page Types That Typically Need Noindex

CategoryPage TypesTypical MetaReason
Auth & AccountLogin, Signup, Password reset, Account dashboardLogin: noindex,nofollow; Signup: noindex,followNo search value; login indexed = security risk; signup follow allows crawl of Privacy/Terms links
Admin & PrivateAdmin, Staging, Test pages, Internal toolsnoindex,nofollowNot for public; avoid discovery
Conversion EndpointsThank-you, Confirmation, Checkout success, Download gatenoindex,followPost-conversion; no SERP value; allow link equity
System & Utility404, Internal search results, Faceted/filter URLsnoindex,follow or noindex,nofollowThin/duplicate; 404 = error state
LegalPrivacy, Terms, Cookie Policy (optional)Often noindex,followLow-value indexed; reduces clutter
Duplicate & ThinPrinter-friendly, Parameter URLs, Near-duplicatenoindex,follow or canonicalDuplicate content; canonical preferred when possible
Low-ValueMedia kit, Feedback board (external), Thin pressnoindex or index for brand queriesCase-by-case

noindex,follow vs noindex,nofollow: Use noindex,follow for most cases—excludes from SERP but allows link equity. Use noindex,nofollow only for login (security), staging, or temporary test pages.

Page Removal Decision Framework

When intentionally removing a page from the web, choose the method based on whether a relevant alternative exists and whether the page should remain accessible:

ScenarioMethodRationale
Has a closely related replacement page301 redirectPreserves accumulated link signals and user flow
Content merged into a new page301 redirectDirect old URL to the new canonical location
Permanently deleted, no alternative410 GoneExplicitly signals permanent removal to search engines
Deleted, uncertain if permanent404 Not FoundSafe default; can reinstate later if needed
Still accessible but should not be indexednoindexPage remains available to users; excluded from SERP

Before removing: Check the URL's search traffic, backlinks, internal links, and conversion value. If the page has value, consider updating or merging rather than removing.

Common mistakes:

  • 404-ing pages that have relevant alternatives (wastes accumulated signals)
  • Redirecting all deleted pages to the homepage (breaks user intent)
  • Creating redirect chains (A → B → C) instead of direct redirects
  • Removing pages without cleaning up internal links pointing to them
  • Using robots.txt to block noindex pages (crawler must access the page to read the noindex directive)

Post-removal cleanup:

  1. Remove deleted URLs from XML sitemap; update and resubmit
  2. Update internal links to point directly to the final URL (avoid relying on redirects)
  3. For 301 redirects, ensure the target URL is in the sitemap
  4. In GSC, use URL Inspection to verify important pages; use Removals tool for temporary quick-hide (not permanent — use proper HTTP status or noindex)

Google Indexing API

TypeTypical use
JobPostingJob boards
BroadcastEventLive platforms

Requirements: Enable Indexing API, create service account, add owner in Search Console, request quota (default 200 URLs/day).

Output Format

Related Skills

  • google-search-console: Find and diagnose indexing issues in GSC
  • robots-txt: Path-level crawl control; when to use robots.txt vs noindex; do not block /_next/ or noindex pages
  • page-metadata: Meta robots implementation; noindex vs nofollow
  • xml-sitemap: Submit and maintain sitemap
  • indexnow: Faster indexing for Bing
  • canonical-tag: Resolve duplicate content

來源與署名

來源:kostja94/marketing-skills位於skills/seo/technical/indexing提交8dd89c5

授權條款: 無授權條款

內容歸原作者所有。SourceWeft 從公開儲存庫中收錄這些內容。

檢舉或申請下架