Aidp Migration

作者 oracle-samples90b42d6c24d4无许可证收录于 2026年10月8日更新于 2026年10月8日

Guide a migration of notebooks/jobs from another platform (e.g. Databricks) into AIDP. Use when the user wants to port Databricks notebooks/jobs to AIDP, move workloads onto the AIDP lakehouse, or plan a migration. Orchestration-only — it composes the other self-contained aidp-* skills; it adds no new API surface.

仅含说明DevOps & Cloud
AI 生成的概览

指导将 Databricks 等平台的笔记本和作业迁移到 AIDP,需人工确认。

功能
按顺序编排 AIDP 迁移:盘点源笔记本、作业、表和库,落地数据,移植笔记本,重建作业 DAG 与调度,然后验证并切换。它组合其他 aidp-* 技能,不新增 API 接口。该技能仅做编排,并明确不承诺批量自动转换。
适用场景
当用户希望把 Databricks 笔记本或作业移植到 AIDP、将工作负载迁移到 AIDP 湖仓,或规划此类迁移时使用。
运行要求
通过代理运行,需要 OCI CLI 访问权限以使用 oci raw-request 调用控制平面,并使用随附的 scripts/aidp_sql.py 辅助脚本执行交互式 Spark-SQL 和单元格执行;还依赖所组合的 aidp-* 技能和参考文档。无需 MCP 服务器。该技能本身不附带脚本。

aidp-migration — guided migration into AIDP

Plan and execute a migration of notebooks/jobs onto AIDP by composing the other skills. Adds no new API surface — it sequences ingestion, notebooks, pipelines, and validation. Like every skill in this plugin it is self-contained: control-plane ops run via oci raw-request and interactive Spark-SQL/cell execution runs via the bundled scripts/aidp_sql.py. No MCP server or ai-data-engineer-agent repo is required.

When to use

  • "Migrate these Databricks notebooks/jobs to AIDP", "move this workload onto AIDP", "plan a migration".

Workflow

  1. Inventory the source assets (notebooks, jobs/schedules, tables, libraries) — list what must move.
  2. Land data: ingest source tables/files (aidp-ingest-file-to-table; external sources via the spark-connectors plugin + aidp-federate).
  3. Port notebooks: recreate notebooks in the workspace (aidp-notebooks / aidp-workspace-files), adapting platform-specifics (paths, compute:/// defaultFS caveats, cluster/session APIs, Delta vs other formats). Validate cells run with the bundled helper (python "$PLUGIN_DIR/scripts/aidp_sql.py" … --code …).
  4. Recreate jobs: build the task DAG + schedule (aidp-pipelines), heeding the clusterName-UUID pitfall and NOTEBOOK_TASK/dependsOn shape.
  5. Validate: profile + quality-check migrated tables (aidp-profiling-tables, aidp-data-quality); compare row counts/aggregates against the source; dry-run the job and inspect output.
  6. Cut over: only after validation; keep the source as fallback until confirmed.

Engines (inherited from the composed skills)

  • Control-plane (workspaces, catalogs, tables, clusters, jobs, files) → oci raw-request against the AIDP REST API — see references/oci-raw-request.md and references/no-mcp-rest-map.md.
  • Interactive Spark-SQL / cell execution (validate ported cells, compare counts/aggregates) → python "$PLUGIN_DIR/scripts/aidp_sql.py" --region <r> --datalake <OCID> --workspace <ws> --cluster <key> --code <…>.

Notes

  • Common AIDP gotchas to apply during porting: compute:/// defaultFS (executors can't write the driver FS; size APIs return 0 — measure via oci://), manifest commit semantics for external tables, and the clusterName-UUID pitfall when wiring jobs.
  • Keep scope to AIDP-native migration. OAC and OCI networking are out of scope.
  • This is a guided, human-confirmed process — no bulk automated conversion claims.

References

  • composes aidp-ingest-file-to-table, aidp-notebooks, aidp-workspace-files, aidp-pipelines, aidp-profiling-tables, aidp-data-quality, aidp-federate
  • references/oci-raw-request.md · references/no-mcp-rest-map.md · scripts/aidp_sql.py

来源与署名

来源:oracle-samples/oracle-aidp-samples位于ai/claude-code-plugins/oracle-ai-data-platform-workbench-engineer-agent/skills/aidp-migration提交90b42d6

许可证: 无许可证

内容归原作者所有。SourceWeft 从公开仓库中收录这些内容。

举报或申请下架

更多来自 oracle-samples/oracle-aidp-samples 的技能

Aidp Workspace Admin

oracle-samples

Provision and inspect AIDP DataLake instances and workspaces, including private-network workspaces attached to a customer VCN/subnet. Use when the user wants to create/list/get a workspace or DataLake instance, set up a new (e.g. private) AIDP environment, or replicate a customer setup. Create/delete are guarded — confirm before any provisioning.

待分类2026年10月8日

Aidp Volumes

oracle-samples

Work with AIDP volumes — list volumes, browse files inside a volume, upload/download via the PAR flow, and create directories. Use when the user mentions volumes, needs to stage large/binary files, or move data in/out of a volume (distinct from the workspace filesystem). Control-plane via the official `aidp` CLI.

待分类2026年10月8日

Aidp Verified Queries

oracle-samples

维护经过验证的问题到 Spark SQL 配对库,让智能体在生成新 SQL 前优先复用可信查询。

Data & Analytics2026年10月8日

Aidp User Settings

oracle-samples

通过 aidp CLI 或 oci raw-request 备用方式管理 AIDP DataLake 用户设置与偏好。

Productivity & Workflow2026年10月8日

Aidp Spark Optimization

oracle-samples

指导 Apache Spark 3.5.0 性能调优:分区、shuffle、连接、倾斜、内存、文件布局、AQE 与 Delta Lake。

Data & Analytics2026年10月8日

Aidp Semantic Model

oracle-samples

维护 .aidp/semantic.md 业务语义层,定义指标、连接、同义词和值字典,为自然语言转 SQL 提供依据。

Data & Analytics2026年10月8日