Skip to main content

Migration

Choose a migration path based on whether you need to replace a Hive table, keep a separate copy, or expose historical Paimon data to existing Hive queries.

Migration moves Hive files into Paimon, clone copies data while keeping the source, and tag-to-partition exposes Paimon snapshots as Hive partitions.

Choose a Guide​

GoalGuideEffect on the sourceResult
Convert a Hive table or database to PaimonMigrate from HiveMoves data files; removes the original Hive table by defaultPaimon append tables in a Hive catalog
Copy tables while keeping the source availableClone to PaimonKeeps source tables and dataSeparate Paimon tables; Hive sources become append tables
Query daily views of an updating Paimon table using Hive partition filtersExpose Tags as Hive PartitionsKeeps the Paimon table and its write pathHive partition values that select tags or preview snapshots

Migration and clone import existing data. Tag-to-partition changes how Hive reads a Paimon table; it does not migrate Hive files or physically repartition the Paimon table.

Before Moving Data​

  1. Choose the target table model. Hive migration and Hive clone create append tables. If the target needs primary-key updates or a different schema or partition layout, plan a data rewrite into a table with that model.
  2. Prepare the runtime and catalogs. Configure access to the Hive metastore and storage. Use the engine setup and connector requirements linked from each guide.
  3. Define the validation scope. Record source schemas, partitions, row counts, and key aggregates before the operation. For a consistent comparison, keep source data stable while it is being moved or copied.
  4. Plan the switch. In-place migration is not atomic and requires a backup. Clone lets you validate a separate target before switching readers and writers.