Clickhouse Architecture Advisor

by ClickHouse2f6ec4b17a81Apache-2.0544 starsListed Oct 8, 2026Updated Oct 8, 2026Repository updated 9 days ago

MUST USE when designing ClickHouse architectures, selecting between ingestion or modeling patterns, or translating best practices into workload-specific system designs. Complements clickhouse-best-practices with decision frameworks and explicit provenance labels.

Instructions onlyData & Analytics
AI-generated overview

Guides ClickHouse architecture decisions for specific workloads, labeling each recommendation as official, derived, or field.

What it does
This skill provides workload-aware decision frameworks for designing ClickHouse architectures, covering ingestion strategy, real-time pre-aggregation, time-series partitioning, join enrichment, and late-arriving upserts. It instructs the agent to identify the workload shape, read the relevant rule files, attach official documentation links, and classify every recommendation by provenance. Output follows a structured format with workload summary, key decisions, and recommendations including confidence, sources, and validation steps.
When to use it
Use it when designing ClickHouse architectures, choosing between ingestion or modeling patterns, or translating best practices into workload-specific system designs. It suits scenarios such as observability, SIEM, product analytics, IoT telemetry, and financial market data.
Requirements
No scripts; instructions only. It references bundled rule files, mapping files, and examples, and expects access to official ClickHouse documentation.

ClickHouse Architecture Advisor

This skill adds workload-aware architecture decisioning on top of clickhouse-best-practices.

Official docs remain the source of truth. This skill must always prefer official ClickHouse documentation when available.

Required behavior

Before producing recommendations:

  1. Identify the workload shape
    • observability
    • security / SIEM
    • product analytics
    • IoT / telemetry
    • market data / financial services
    • mixed OLAP with point-lookups
  2. Read the relevant decision rule files in rules/
  3. Use mappings/doc_links.yaml to attach official documentation
  4. Classify every recommendation as:
    • official
    • derived
    • field
  5. Never present field guidance as official guidance
  6. If a recommendation is uncertain, say so explicitly

Provenance rules

official

Use this when the recommendation is directly backed by official docs.

derived

Use this when the recommendation is not stated verbatim in docs but follows logically from documented ClickHouse behavior.

field

Use this only for experience-based guidance that may be situational. When using field, include:

  • a disclaimer that the advice is heuristic
  • a relevant official doc if one partially applies
  • the reason the advice depends on workload context

Read these rule files by scenario

Real-time ingestion design

  1. rules/decision-ingestion-strategy.md
  2. rules/decision-real-time-preaggregation.md
  3. Relevant best-practices insert rules

Time-series and retention design

  1. rules/decision-partitioning-timeseries.md
  2. Relevant best-practices schema partition rules

Enrichment and dimension lookups

  1. rules/decision-join-enrichment.md
  2. Relevant best-practices query join rules

Mutable state / late-arriving events

  1. rules/decision-late-arriving-upserts.md
  2. Relevant best-practices mutation avoidance rules

Output format

Structure responses like this:

markdown
## Workload Summary- workload:- latency target:- data shape:- primary query patterns:- operational constraints:
## Key Decisions- ...- ...
## Recommendations
### <Recommendation title>
**What**...
**Why**...
**How**...
**Category**official | derived | field
**Confidence**high | medium | heuristic
**Source**- doc link(s)
**Validation**- concrete SQL, metric, or smoke test

Architecture-specific guidance

Prefer decision frameworks over generic advice. Good responses should:

  • explain tradeoffs
  • identify the likely operating bottleneck
  • separate immediate actions from structural redesign
  • provide target architecture patterns, not just isolated settings

Full reference

See AGENTS.md for the compiled version and examples/ for sample outputs.

Source and attribution

Source:ClickHouse/agent-skillsinskills/clickhouse-architecture-advisorat commit2f6ec4b

License: Apache-2.0

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal

More from ClickHouse/agent-skills

Clickhouse Managed Postgres Rca

ClickHouse

Evidence-based root cause analysis workflow for performance issues on ClickHouse-managed Postgres instances.

DevOps & Cloud544updated 9 days ago

Clickhouse Js Node Troubleshooting

ClickHouse

Troubleshoots common errors and configuration issues with the ClickHouse Node.js client (@clickhouse/client).

Software Development544updated 9 days ago

Clickhouse Best Practices

ClickHouse

Reviews ClickHouse schemas, queries, and ingestion strategies against 31 documented best-practice rules.

Data & Analytics544updated 9 days ago

Chdb Sql

ClickHouse

Use when the user wants to run SQL — especially analytical SQL — on local files (parquet/csv/json), URLs, S3 paths, or remote databases (Postgres, MySQL, MongoDB, ClickHouse Cloud, Iceberg, Delta Lake) without setting up a server. Provides chDB — embedded ClickHouse SQL in Python with 1000+ functions, Session for stateful multi-step pipelines, parametrized queries, and cross-source joins via `s3()`, `mysql()`, `postgresql()`, `iceberg()`, `deltaLake()`, `remoteSecure()` table functions. TRIGGER when: user wants SQL on parquet/csv/files or across remote analytical sources; uses ClickHouse SQL features (window functions, windowFunnel, geoToH3, JSON path ops, Session, parametrized queries); imports `chdb` or calls `chdb.query()`. SKIP this skill for pandas-style DataFrame method-chaining (use chdb-datastore instead) or ClickHouse server administration.

Includes scripts
Awaiting classification544updated 9 days ago

Chdb Datastore

ClickHouse

Use chdb DataStore as a drop-in, ClickHouse-backed replacement for pandas to query, join and aggregate tabular data.

Includes scripts
Data & Analytics544updated 9 days ago