
Chdb Datastore
by ClickHouse2f6ec4b17a81Apache-2.0544 starsListed Oct 8, 2026Updated Oct 8, 2026Repository updated 9 days ago
Use when the user has tabular data (pandas DataFrame, parquet, csv, Arrow, json) and wants to filter, group, aggregate, join, or speed up slow pandas. Provides chDB DataStore — same pandas API, ClickHouse engine underneath. Also handles reading from S3, MySQL, PostgreSQL, MongoDB, ClickHouse Cloud, Iceberg, Delta Lake as DataFrames and joining across sources. TRIGGER when: user mentions DataFrame, parquet, csv, "fast pandas", "speed up pandas", or cross-source DataFrame joins; user imports `chdb.datastore` or `from datastore import DataStore`. SKIP this skill for raw SQL syntax (use chdb-sql instead), ClickHouse server administration, or non-Python DataStore API work.
Only the file list is public. File contents are available once the skill is installed in a workspace.
| Path | Size | Type |
|---|---|---|
| examples/examples.md | 10.2 KB | text/markdown |
| metadata.json | 464 B | application/json |
| README.md | 1.1 KB | text/markdown |
| references/api-reference.md | 8.6 KB | text/markdown |
| references/connectors.md | 7.2 KB | text/markdown |
| scripts/verify_install.py | 2.9 KB | text/plain |
| SKILL.md | 5.4 KB | text/markdown |
Source and attribution
Source:ClickHouse/agent-skillsinskills/chdb-datastoreat commit2f6ec4b
License: Apache-2.0
Content belongs to its original authors. SourceWeft indexes it from a public repository.
More from ClickHouse/agent-skills

Clickhouse Managed Postgres Rca
ClickHouse
Evidence-based root cause analysis workflow for performance issues on ClickHouse-managed Postgres instances.

Clickhouse Js Node Troubleshooting
ClickHouse
Troubleshoots common errors and configuration issues with the ClickHouse Node.js client (@clickhouse/client).

Clickhouse Best Practices
ClickHouse
Reviews ClickHouse schemas, queries, and ingestion strategies against 31 documented best-practice rules.

Clickhouse Architecture Advisor
ClickHouse
Guides ClickHouse architecture decisions for specific workloads, labeling each recommendation as official, derived, or field.

Chdb Sql
ClickHouse
Use when the user wants to run SQL — especially analytical SQL — on local files (parquet/csv/json), URLs, S3 paths, or remote databases (Postgres, MySQL, MongoDB, ClickHouse Cloud, Iceberg, Delta Lake) without setting up a server. Provides chDB — embedded ClickHouse SQL in Python with 1000+ functions, Session for stateful multi-step pipelines, parametrized queries, and cross-source joins via `s3()`, `mysql()`, `postgresql()`, `iceberg()`, `deltaLake()`, `remoteSecure()` table functions. TRIGGER when: user wants SQL on parquet/csv/files or across remote analytical sources; uses ClickHouse SQL features (window functions, windowFunnel, geoToH3, JSON path ops, Session, parametrized queries); imports `chdb` or calls `chdb.query()`. SKIP this skill for pandas-style DataFrame method-chaining (use chdb-datastore instead) or ClickHouse server administration.
More in Data & Analytics

Spanner Basics
Guides Google Cloud Spanner administration, schema design, querying and performance diagnosis.

Gke Cost Analysis
Answers natural-language questions about GKE cluster and workload costs using BigQuery billing exports and live cluster metrics.

Datalineage Bigquery Asset Impact Analysis
Guides an agent through downstream impact (blast radius) analysis for a BigQuery table or view using Data Lineage.

Bigquery Troubleshooting
Diagnoses failing, slow, or unexpectedly expensive BigQuery jobs through structured root-cause workflows.

Bigquery Optimization
Guides BigQuery cost and performance optimization across capacity editions, storage layout, and SQL queries.

Bigquery Bigframes
Guides writing Python code with BigQuery DataFrames (BigFrames) for data processing, analysis, and machine learning.