Blog Taxonomy

agricidaniel/claude-blog/skills/blog-taxonomy

作者 agricidaniel2500d4c76503MIT2.3K 个星标收录于 2026年10月8日更新于 2026年10月8日仓库昨天更新

Extract, suggest, and sync tags and categories for blog posts across all major CMS platforms. Supports WordPress REST API, Shopify GraphQL, Ghost Content API, Strapi REST/GraphQL, and Sanity GROQ. Generates tag suggestions from content analysis (keyword frequency, heading extraction, semantic grouping), enforces minimum post-count thresholds to prevent thin tag archives, and syncs taxonomy via authenticated API calls. Use when user says "tags", "categories", "taxonomy", "tag suggestions", "sync tags", "WordPress tags", "Shopify tags".

AI 生成的概览

为 WordPress、Shopify、Ghost、Strapi 和 Sanity 等 CMS 平台提取、建议、审核并同步博客标签与分类。

功能
通过解析标题层级、强调格式和词频,从博客文章中提取候选标签与分类,再经排序和去重生成精简的建议列表。它还能审核现有分类体系,找出标签过少、孤立标签、标签膨胀和重复别名等问题,并通过带认证的 API 调用把分类变更推送到 CMS。支持的平台包括 WordPress REST、Shopify GraphQL、Ghost Content API、Strapi REST/GraphQL 和 Sanity GROQ。
适用场景
适用于需要为文章生成标签或分类建议、对现有博客分类体系做健康检查,或在受支持的 CMS 中创建和分配标签的场景。适合管理主题集群和标签归档页的内容编辑流程。
运行要求
仅为说明文档,不含脚本。需要设置 CMS 相关的 shell 环境变量:CMS_TYPE、CMS_URL 和 CMS_API_KEY,WordPress 应用密码还需 CMS_USERNAME,Sanity 变更操作需 SANITY_API_VERSION,CMS_ALLOWED_HOSTS 为可选项。同步和审核操作需要能通过 HTTPS 访问 CMS API 并具备有效凭据。

Blog Taxonomy

Manage tags, categories, and topic clusters across CMS platforms.

Commands

CommandPurpose
/blog taxonomy suggest <file>Extract candidate tags and categories from content
/blog taxonomy sync <cms>Push taxonomy to CMS via authenticated API
/blog taxonomy audit [directory]Check for thin tags, orphan tags, taxonomy bloat

Tag Suggestion Workflow

Step 1: Parse Content Structure

Read the target file and extract:

  • All H2 and H3 headings (primary topic signals)
  • Bold and italic phrases (emphasis signals)
  • Existing frontmatter tags/categories if present

Step 2: Frequency Analysis

Scan the body text for high-frequency phrases:

  • 1-word terms: minimum 4 occurrences (excluding stop words)
  • 2-word phrases: minimum 3 occurrences
  • 3-word phrases: minimum 2 occurrences

Exclude common non-tag words: articles, prepositions, conjunctions, pronouns.

Step 3: Semantic Grouping

Group related candidates into clusters:

  • Merge singular/plural variants (keep the more common form)
  • Merge hyphenated and non-hyphenated forms
  • Group synonyms under the highest-frequency term

Step 4: Deduplicate and Rank

  • Fuzzy match on slugified names (Levenshtein distance <= 2)
  • Do not auto-merge short slugs under 5 characters using Levenshtein alone; require token overlap or manual review
  • Score each candidate: (frequency * 2) + (heading_presence * 5) + (emphasis * 1)
  • Return top 5-10 ranked suggestions

Output Format

## Tag Suggestions: [Post Title]
| Rank | Tag | Score | Source ||------|-----|-------|--------|| 1 | content-marketing | 18 | H2 + 6 mentions || 2 | seo-strategy | 14 | H3 + 4 mentions || 3 | keyword-research | 11 | 5 mentions + bold |
### Suggested Categories- Primary: [best-fit category]- Secondary: [optional second category]

CMS Adapters

Adapter Overview

CMSAPI TypeAuth MethodTags Model
WordPressRESTApplication Passwords (base64)First-class entities with IDs
ShopifyGraphQL (Admin API)Admin API access tokenString array on Article
GhostREST (Admin API)API key with JWT signingFirst-class entities
StrapiREST or GraphQLAPI token (Bearer)User-defined content type
SanityGROQ / MutationsProject token (Bearer)Document type

WordPress Adapter

List tags:

GET {CMS_URL}/wp-json/wp/v2/tags?per_page=100&search={keyword}Authorization: Basic {base64(username:app_password)}

Create tag:

POST {CMS_URL}/wp-json/wp/v2/tagsBody: {"name": "Tag Name", "slug": "tag-name", "description": "Optional"}

List categories (hierarchical, supports parent field):

GET {CMS_URL}/wp-json/wp/v2/categories?per_page=100

Create category:

POST {CMS_URL}/wp-json/wp/v2/categoriesBody: {"name": "Category", "slug": "category", "parent": 0}

Assign tags to post:

POST {CMS_URL}/wp-json/wp/v2/posts/{id}Body: {"tags": [1, 2, 3], "categories": [4]}

Pagination: follow X-WP-TotalPages header for full listing.

Shopify Adapter

Tags on Shopify are string arrays on the Article object, not first-class entities.

Update article tags (GraphQL Admin API):

graphql
mutation {  articleUpdate(id: "gid://shopify/Article/123", article: {    tags: ["tag-one", "tag-two", "tag-three"]  }) {    article { id tags }    userErrors { field message }  }}

List all tags in use (GraphQL):

graphql
{  articles(first: 250, after: $cursor) {    pageInfo { hasNextPage endCursor }    edges {      node { id title tags }    }  }}

Auth header: X-Shopify-Access-Token: {token}

Pagination: loop while pageInfo.hasNextPage is true, passing endCursor as the next $cursor.

Note: REST API marked legacy Oct 2024. GraphQL required for new apps since Apr 2025.

Ghost Adapter

List tags:

GET {CMS_URL}/ghost/api/admin/tags/?limit=allAuthorization: Ghost {jwt_token}

Create tag:

POST {CMS_URL}/ghost/api/admin/tags/Body: {"tags": [{"name": "Tag Name", "slug": "tag-name"}]}

JWT generation: sign with admin API key (id:secret format), iat = now, exp = 5 min, audience = /admin/.

Strapi Adapter

Endpoint auto-generated from content types. Typical setup:

GET {CMS_URL}/api/tags?pagination[pageSize]=100POST {CMS_URL}/api/tagsBody: {"data": {"name": "Tag Name", "slug": "tag-name"}}Authorization: Bearer {api_token}

Pagination: increment pagination[page] until all pages are exhausted.

Strapi v4 responses use the data wrapper with attributes; Strapi v5 uses a flatter response shape. Detect the version or normalize both shapes before deduplication. Check your content type schema for field names.

Sanity Adapter

Query tags (GROQ):

*[_type == "tag"] { _id, name, slug }

Create tag (Mutations API):

POST https://{project_id}.api.sanity.io/{SANITY_API_VERSION}/data/mutate/{dataset}Body: {"mutations": [{"create": {"_type": "tag", "name": "Tag", "slug": {"current": "tag"}}}]}Authorization: Bearer {token}

Default SANITY_API_VERSION to a current tested API date supplied by the project environment; do not hard-code it in generated requests.

Taxonomy Audit Workflow

Step 1: Inventory

Scan all posts in the target directory (or fetch from CMS). Build a map:

  • tag_name -> [list of post files/IDs using this tag]
  • category_name -> [list of post files/IDs]

Step 2: Health Checks

CheckThresholdAction
Thin tag archives< 5 posts per tagReview for merge or noindex after traffic, intent, and link checks
Orphan tags0 postsRecommend deletion
Tag bloatMore than max(50, post_count * 0.25) total tags, adjusted for taxonomy purposeRecommend consolidation
Category depth> 3 levelsRecommend flattening
Uncategorized postsNo category assignedAssign to appropriate category
Duplicate slugsSame slug, different nameMerge into canonical version

Step 3: Recommendations

Group findings by priority:

  • Critical: orphan tags creating empty archive pages (crawl waste)
  • High: thin tags with < 5 posts after traffic, intent, and link checks
  • Medium: tag bloat above the scaled threshold (diluted taxonomy, harder to navigate)
  • Low: naming inconsistencies (mixed case, hyphen vs space)

Output Format

## Taxonomy Audit: [Site/Directory]
**Total tags**: [n] | **Total categories**: [n]**Healthy**: [n] | **Thin**: [n] | **Orphan**: [n]
### Critical Issues- [orphan tags list]
### Recommendations1. Merge [tag-a] and [tag-b] (same topic, [n] combined posts)2. Delete orphan tags: [list]3. Merge or noindex tag archives with < 5 posts only after traffic, intent, and link checks

Site-Wide Guidelines

  • Aim for 5-10 main categories per site (broad topics)
  • Tags should have at least 5 posts before creating an archive page
  • Use consistent slug format: lowercase, hyphen-separated
  • Every post needs exactly 1 primary category
  • Tags per post: 3-8 recommended, never exceed 15

Environment Variables

VariablePurposeExample
CMS_TYPEPlatform identifierwordpress, shopify, ghost, strapi, sanity
CMS_URLHTTPS base URL of the CMShttps://example.com
CMS_ALLOWED_HOSTSOptional comma-separated allowlist for CMS hostsexample.com,admin.example.com
CMS_USERNAMEWordPress username when using Application Passwords[email protected]
CMS_API_KEYAuthentication credentialWordPress app password, API token, or key
SANITY_API_VERSIONSanity API date for mutationsv2026-07-01

These must be set in the shell environment. Never store credentials in files or commit them to version control. The skill reads them via $CMS_TYPE, $CMS_URL, $CMS_USERNAME, $CMS_API_KEY, and optional platform-specific variables at runtime.

Security rule for CMS calls: require HTTPS, allow only http and https parsing paths but send authenticated requests over HTTPS only, resolve DNS and block loopback/private/link-local/reserved IPs, validate redirects with the same checks or disable redirects, cap timeouts at 10 seconds, and enforce CMS_ALLOWED_HOSTS when set.

Error Handling

  • Missing environment variables: If CMS_TYPE, CMS_URL, or CMS_API_KEY is unset, or if WordPress lacks CMS_USERNAME, report which variable is missing and provide the expected format
  • Invalid credentials: If the CMS API returns 401/403, report "Authentication failed - check CMS_USERNAME/CMS_API_KEY" and do not retry
  • Connection timeouts: If the CMS endpoint is unreachable after 10 seconds, report the timeout and suggest checking CMS_URL
  • Duplicate tag slugs: If a tag already exists on the CMS, skip creation and note "Tag already exists: [name]"
  • Rate limits: If the CMS API returns 429, honor Retry-After when present; otherwise use exponential backoff and retry once. Report if the limit persists
  • Unsupported CMS: If CMS_TYPE is not one of the 5 supported platforms, list the valid options and exit

来源与署名

来源:agricidaniel/claude-blog位于skills/blog-taxonomy提交2500d4c

许可证: MIT

内容归原作者所有。SourceWeft 从公开仓库中收录这些内容。

举报或申请下架

更多来自 agricidaniel/claude-blog 的技能

Blog Cannibalization

agricidaniel

Detect keyword cannibalization across blog posts by extracting primary keywords from titles and headings, clustering semantically similar targets, and flagging posts competing for the same search intent. Supports local-only mode (grep-based) and DataForSEO API mode (Page Intersection endpoint at ~$0.01/call). Outputs severity-scored report with merge or differentiate recommendations. Use when user says "cannibalization", "keyword overlap", "competing pages", "duplicate keywords", "cannibalize".

待分类2.3K昨天更新

Blog Rewrite

agricidaniel

改写现有博客文章,以提升 Google SEO 与 AI 引用可见度,替换无来源数据并补充结构化元素。

Marketing & Sales2.3K昨天更新

Blog Translate

agricidaniel

Translate existing blog posts into one or more target languages with SEO-optimized localization. Produces native-quality translations that preserve markdown structure, frontmatter, schema JSON-LD, image and chart embeds, and citation capsules. Localizes keywords, meta tags, numbers, dates, currencies, and quote styles per locale. Flags machine-translation artifacts for review. Run BEFORE blog-localize: this handles language conversion; localize handles cultural adaptation after translation completes. Use when user says "translate blog", "blog translate", "uebersetzen", "traduire", "traducir", "translate post", "blog auf Deutsch", "blog en espanol".

待分类2.3K昨天更新

Blog Write

agricidaniel

Write new blog articles from scratch optimized for Google rankings and AI citations. Generates full articles with template selection, answer-first formatting, Key Takeaways summary box, information gain markers, citation capsules, sourced statistics, Pixabay/Unsplash images, built-in SVG chart generation, optional FAQ sections, internal linking zones, and proper heading hierarchy. Supports MDX, markdown, and HTML output. Use when user says "write blog", "new blog post", "create article", "write about", "draft blog", "generate blog post".

待分类2.3K昨天更新

Blog Schema

agricidaniel

Generate complete JSON-LD schema markup for blog posts with Article/BlogPosting, Person, Organization, BreadcrumbList, ImageObject, and optional FAQPage. Validates against Google requirements and warns about deprecated types. Use when user says "schema", "blog schema", "json-ld", "structured data", "schema markup", "generate schema".

待分类2.3K昨天更新

Blog Strategy

agricidaniel

制定博客策略,涵盖主题集群、受众画像、竞争分析和 AI 引用 SEO 规划。

Marketing & Sales2.3K昨天更新