Aidp Migration

by oracle-samples90b42d6c24d4No licenseListed Oct 8, 2026Updated Oct 8, 2026

Guide a migration of notebooks/jobs from another platform (e.g. Databricks) into AIDP. Use when the user wants to port Databricks notebooks/jobs to AIDP, move workloads onto the AIDP lakehouse, or plan a migration. Orchestration-only — it composes the other self-contained aidp-* skills; it adds no new API surface.

Instructions onlyDevOps & Cloud
AI-generated overview

Guides a human-confirmed migration of notebooks and jobs from platforms like Databricks onto AIDP.

What it does
Sequences an AIDP migration: inventorying source notebooks, jobs, tables and libraries, landing data, porting notebooks, recreating job DAGs and schedules, then validating and cutting over. It composes other aidp-* skills and adds no new API surface. It is orchestration-only and explicitly makes no bulk automated conversion claims.
When to use it
Use when a user wants to port Databricks notebooks or jobs to AIDP, move workloads onto the AIDP lakehouse, or plan such a migration.
Requirements
Runs via the agent with OCI CLI access for control-plane calls through oci raw-request, plus the bundled scripts/aidp_sql.py helper for interactive Spark-SQL and cell execution; it also depends on the composed aidp-* skills and reference documents. No MCP server is required. The skill ships no scripts of its own.

aidp-migration — guided migration into AIDP

Plan and execute a migration of notebooks/jobs onto AIDP by composing the other skills. Adds no new API surface — it sequences ingestion, notebooks, pipelines, and validation. Like every skill in this plugin it is self-contained: control-plane ops run via oci raw-request and interactive Spark-SQL/cell execution runs via the bundled scripts/aidp_sql.py. No MCP server or ai-data-engineer-agent repo is required.

When to use

  • "Migrate these Databricks notebooks/jobs to AIDP", "move this workload onto AIDP", "plan a migration".

Workflow

  1. Inventory the source assets (notebooks, jobs/schedules, tables, libraries) — list what must move.
  2. Land data: ingest source tables/files (aidp-ingest-file-to-table; external sources via the spark-connectors plugin + aidp-federate).
  3. Port notebooks: recreate notebooks in the workspace (aidp-notebooks / aidp-workspace-files), adapting platform-specifics (paths, compute:/// defaultFS caveats, cluster/session APIs, Delta vs other formats). Validate cells run with the bundled helper (python "$PLUGIN_DIR/scripts/aidp_sql.py" … --code …).
  4. Recreate jobs: build the task DAG + schedule (aidp-pipelines), heeding the clusterName-UUID pitfall and NOTEBOOK_TASK/dependsOn shape.
  5. Validate: profile + quality-check migrated tables (aidp-profiling-tables, aidp-data-quality); compare row counts/aggregates against the source; dry-run the job and inspect output.
  6. Cut over: only after validation; keep the source as fallback until confirmed.

Engines (inherited from the composed skills)

  • Control-plane (workspaces, catalogs, tables, clusters, jobs, files) → oci raw-request against the AIDP REST API — see references/oci-raw-request.md and references/no-mcp-rest-map.md.
  • Interactive Spark-SQL / cell execution (validate ported cells, compare counts/aggregates) → python "$PLUGIN_DIR/scripts/aidp_sql.py" --region <r> --datalake <OCID> --workspace <ws> --cluster <key> --code <…>.

Notes

  • Common AIDP gotchas to apply during porting: compute:/// defaultFS (executors can't write the driver FS; size APIs return 0 — measure via oci://), manifest commit semantics for external tables, and the clusterName-UUID pitfall when wiring jobs.
  • Keep scope to AIDP-native migration. OAC and OCI networking are out of scope.
  • This is a guided, human-confirmed process — no bulk automated conversion claims.

References

  • composes aidp-ingest-file-to-table, aidp-notebooks, aidp-workspace-files, aidp-pipelines, aidp-profiling-tables, aidp-data-quality, aidp-federate
  • references/oci-raw-request.md · references/no-mcp-rest-map.md · scripts/aidp_sql.py

Source and attribution

Source:oracle-samples/oracle-aidp-samplesinai/claude-code-plugins/oracle-ai-data-platform-workbench-engineer-agent/skills/aidp-migrationat commit90b42d6

License: No license

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal

More from oracle-samples/oracle-aidp-samples

Aidp Workspace Admin

oracle-samples

Provision and inspect AIDP DataLake instances and workspaces, including private-network workspaces attached to a customer VCN/subnet. Use when the user wants to create/list/get a workspace or DataLake instance, set up a new (e.g. private) AIDP environment, or replicate a customer setup. Create/delete are guarded — confirm before any provisioning.

Awaiting classificationOct 8, 2026

Aidp Volumes

oracle-samples

Work with AIDP volumes — list volumes, browse files inside a volume, upload/download via the PAR flow, and create directories. Use when the user mentions volumes, needs to stage large/binary files, or move data in/out of a volume (distinct from the workspace filesystem). Control-plane via the official `aidp` CLI.

Awaiting classificationOct 8, 2026

Aidp Verified Queries

oracle-samples

Maintains a repository of validated question-to-Spark-SQL pairs so an agent reuses trusted SQL before writing new queries.

Data & AnalyticsOct 8, 2026

Aidp User Settings

oracle-samples

Manage AIDP DataLake user settings and preferences via the aidp CLI or oci raw-request fallback.

Productivity & WorkflowOct 8, 2026

Aidp Spark Optimization

oracle-samples

Guides Apache Spark 3.5.0 performance tuning: partitions, shuffle, joins, skew, memory, file layout, AQE and Delta Lake.

Data & AnalyticsOct 8, 2026

Aidp Semantic Model

oracle-samples

Maintains a .aidp/semantic.md business-meaning layer defining metrics, joins, synonyms and value dictionaries for NL-to-SQL grounding.

Data & AnalyticsOct 8, 2026