Chdb Datastore

by ClickHouse2f6ec4b17a81Apache-2.0544 starsListed Oct 8, 2026Updated Oct 8, 2026Repository updated 9 days ago

Use when the user has tabular data (pandas DataFrame, parquet, csv, Arrow, json) and wants to filter, group, aggregate, join, or speed up slow pandas. Provides chDB DataStore — same pandas API, ClickHouse engine underneath. Also handles reading from S3, MySQL, PostgreSQL, MongoDB, ClickHouse Cloud, Iceberg, Delta Lake as DataFrames and joining across sources. TRIGGER when: user mentions DataFrame, parquet, csv, "fast pandas", "speed up pandas", or cross-source DataFrame joins; user imports `chdb.datastore` or `from datastore import DataStore`. SKIP this skill for raw SQL syntax (use chdb-sql instead), ClickHouse server administration, or non-Python DataStore API work.

Includes scriptsData & Analytics

Add to a SourceWeft workspace

  1. Open the skill in your dashboard and add it to a workspace.
  2. Enable it for the chats that should use it.

This skill includes scripts. They run in your workspace sandbox when the skill is used — review the file list and source before adding it.

Add to SourceWeft

You will be asked to sign in first, then taken straight to this skill.

Ask your agent to install it

Paste this prompt into Claude Code, Codex, Cursor or another agent that can run commands — or into SourceWeft chat. The agent reads this skill's install guide, shows you its source, license and scripts, and installs it with the SourceWeft CLI once you agree.

Read https://sourceweft.com/skills/gh-clickhouse-agent-skills-chdb-datastore/install.md and install the skill it describes. Before installing, show me its source, license and whether it ships scripts, and wait for my OK. Ask me before changing anything else on my machine.

Read the install guide the agent follows

Install it yourself from a terminal

For Claude Code, Codex, Cursor and other local agents. The SourceWeft CLI fetches the skill from its source repository at the commit scanned here, and verifies every file against the hashes recorded when the skill was scanned. If anything differs, nothing is written.

npx @sourceweft/cli skills install gh-clickhouse-agent-skills-chdb-datastore

Add --agent claude-code, codex, cursor or universal to choose which agent gets it (Claude Code by default).

Installed locally, this skill's scripts run on your machine, not in a sandbox. Read them first — the CLI asks before installing.

Upstream installer — not verified by SourceWeft

The open-source skills installer fetches the same pinned commit, but does not check the files against the hashes SourceWeft recorded.

npx skills add https://github.com/ClickHouse/agent-skills/tree/2f6ec4b17a81a435dd116f9ac19d7b45d44dbd61/skills/chdb-datastore

Source and attribution

Source:ClickHouse/agent-skillsinskills/chdb-datastoreat commit2f6ec4b

License: Apache-2.0

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal

More from ClickHouse/agent-skills

Clickhouse Managed Postgres Rca

ClickHouse

Evidence-based root cause analysis workflow for performance issues on ClickHouse-managed Postgres instances.

DevOps & Cloud544updated 9 days ago

Clickhouse Js Node Troubleshooting

ClickHouse

Troubleshoots common errors and configuration issues with the ClickHouse Node.js client (@clickhouse/client).

Software Development544updated 9 days ago

Clickhouse Best Practices

ClickHouse

Reviews ClickHouse schemas, queries, and ingestion strategies against 31 documented best-practice rules.

Data & Analytics544updated 9 days ago

Clickhouse Architecture Advisor

ClickHouse

Guides ClickHouse architecture decisions for specific workloads, labeling each recommendation as official, derived, or field.

Data & Analytics544updated 9 days ago

Chdb Sql

ClickHouse

Use when the user wants to run SQL — especially analytical SQL — on local files (parquet/csv/json), URLs, S3 paths, or remote databases (Postgres, MySQL, MongoDB, ClickHouse Cloud, Iceberg, Delta Lake) without setting up a server. Provides chDB — embedded ClickHouse SQL in Python with 1000+ functions, Session for stateful multi-step pipelines, parametrized queries, and cross-source joins via `s3()`, `mysql()`, `postgresql()`, `iceberg()`, `deltaLake()`, `remoteSecure()` table functions. TRIGGER when: user wants SQL on parquet/csv/files or across remote analytical sources; uses ClickHouse SQL features (window functions, windowFunnel, geoToH3, JSON path ops, Session, parametrized queries); imports `chdb` or calls `chdb.query()`. SKIP this skill for pandas-style DataFrame method-chaining (use chdb-datastore instead) or ClickHouse server administration.

Includes scripts
Awaiting classification544updated 9 days ago