BladePipe 1.9.0: New data pipelines, faster Oracle writes, and improved stability.
跳到主要内容

PolarDB-X 到 PostgreSQL

选择对端数据库:

数据链路

基本功能

功能说明
Schema Migration

If the target schema does not exist, BladePipe will automatically generate and execute CREATE statements based on the source metadata and the mapping rule.

Full Data Migration

Migrate data by sequentially scanning data in tables and writing it in batches to the target database.

Incremental Data Sync

Sync of common DML like INSERT, UPDATE, DELETE is supported.

Data Verification and Correction

Verify all existing data. Optionally, you can correct the inconsistent data based on verification results. Scheduled DataTasks are supported.
For more information, see Create Verification and Correction DataJob.

Subscription Modification

Add, delete, or modify the subscribed tables with support for historical data migration. For more information, see Modify Subscription.

Position Resetting

Reset positions by file position or timestamp. Allow re-consumption of incremental data logs in a past period or since a specific Binlog file and position.

Table Name Mapping

Support the mapping rules, namely, keeping the name the same as that in Source, converting the text to lowercase, converting the text to uppercase, truncating the name by "_digit" suffix.

Metadata Retrieval

Retrieve the target metadata with filtering conditions or target primary keys set from the source table.

高级功能

功能说明
0 值时间处理

支持将 0 值时间设置成不同类型的值,防止写入对端报错

Custom Code

For more information, see Custom Code Processing, Debug Custom Code and Logging in Custom Code.

Data Filtering Conditions

Support data filtering using WHERE conditions, with SQL-92 as the SQL language. For more information, see Data Filtering.

限制和注意点

限制项说明
版本支持

只支持 PolarDB-X 2.0 版本

PolarDB-X 2.0 字符集

支持 utf8, utf8mb4, latin1,其他编码暂未测试

源库限制

若源端 PolarDB-X 2.0 待同步的表名中含大写字母,则不支持增量同步

主键冲突处理

PostgreSQL <= 9.4 或 Greenplum <= 6, 因不支持冲突掠过或覆盖,当大量主键冲突场景下,性能较低


源端数据源

前置条件

条件说明
账号权限

云数据库为读写权限账号
自建数据库权限:

  • GRANT SELECT ON . TO 'user'@'host'
  • GRANT REPLICATION CLIENT ON . TO 'user'@'host'
  • GRANT REPLICATION SLAVE ON . TO 'user'@'host'

任务参数

参数名称说明
parseBinlogParallel

增量解析 Binlog 的并发数

parseBinlogBufferSize

用于增量解析 Binlog 的环形队列大小

maxTransactionSize

单事务最大数据条数,超过则分段刷出

limitThroughputMb

限制增量 Binlog 流量

needJsonEscape

将 json 中特殊字符进行转义,以写入到对端

Tips: 通用参数配置请参考 通用参数及功能


目标端数据源

前置条件

条件说明
账号权限

具备 SELECT, INSERT, DELETE, UPDATE 常见 DDL 权限
阿里云 AnalyticDB for PostgreSQL 初始账号,或有 SELECT, INSERT, DELETE, UPDATE, 常见 DDL 权限

网络准备

迁移同步节点(sidecar)可连接 PostgreSQL / Greenplum / AnalyticDB for PostgreSQL / PolarDB for PostgreSQL 标准交互接口(如 5432)

任务参数

参数名称说明
keyConflictStrategy

增量写入遇到主键冲突策略:

  • IGNORE 冲突忽略(默认)
  • REPLACE 冲突替换(可选)

dstWholeReplace

将 INSERT 和 UPDATE 操作变成对端的整行覆盖

enableTimeZoneProcess

是否对时间字段进行时区转换

timezone

目标端时区,例如 +08:00, Asia/Shanghai, America/New_York

defaultZeroDate

在遇到'0000-00-00 00:00:00' / '0000-00-00' 值时用于替换的默认值,可选参数有:

  • null (空值)
  • 时间 (14:23:33)
  • 日期 (1970-01-01)
  • 时间日期 (1970-01-01 00:00:00),
  • 时区时间 (14:23:33+08:00 或 1970-01-01 00:00:00+08:00)
caseSensitive

对端写入SQL语句表名大小写策略,包含

  • UpperCase (转大写)
  • LowerCase (转小写)
  • Sensitive (添加限定符)
  • NoSpecified (不转换/不加限定符)
writeStrategy

对端写入策略,包含

  • ROW (单条)
  • MULTI_SQL (多语句)
  • BATCH (批量,默认选项)
  • COPY (PostgreSQL COPY 指令)
defaultGisSRID

设置 GIS 数据类型的 SRID

Tips: 通用参数配置请参考 通用参数及功能