This is an automated email from the ASF dual-hosted git repository.
corgy-w pushed a commit to branch dev
in repository https://gitbox.apache.org/repos/asf/seatunnel.git
The following commit(s) were added to refs/heads/dev by this push:
new 7e7738d5ad [Docs][Connector-V2] Improve Hive Elasticsearch Mysql and
SqlServer connector docs (#11719)
7e7738d5ad is described below
commit 7e7738d5addc467bd2cd3d15b9730490115ae354
Author: Daniel Carter <[email protected]>
AuthorDate: Sun Aug 16 13:10:50 2026 +0800
[Docs][Connector-V2] Improve Hive Elasticsearch Mysql and SqlServer
connector docs (#11719)
---
docs/en/connectors/sink/Elasticsearch.md | 62 +++++++++++++----------
docs/en/connectors/sink/Hive.md | 48 ++++++++++--------
docs/en/connectors/sink/Mysql.md | 12 ++---
docs/en/connectors/source/Hive.md | 87 ++++++++++++++------------------
docs/en/connectors/source/SqlServer.md | 5 +-
docs/zh/connectors/sink/Elasticsearch.md | 4 +-
docs/zh/connectors/source/Hive.md | 27 +++-------
7 files changed, 117 insertions(+), 128 deletions(-)
diff --git a/docs/en/connectors/sink/Elasticsearch.md
b/docs/en/connectors/sink/Elasticsearch.md
index 1502ffc476..a414dde0ab 100644
--- a/docs/en/connectors/sink/Elasticsearch.md
+++ b/docs/en/connectors/sink/Elasticsearch.md
@@ -2,9 +2,15 @@ import ChangeLog from
'../changelog/connector-elasticsearch.md';
# Elasticsearch
+## Support Those Engines
+
+> Spark<br/>
+> Flink<br/>
+> SeaTunnel Zeta<br/>
+
## Description
-Output data to `Elasticsearch`.
+Output data to Elasticsearch or OpenSearch-compatible clusters. The connector
uses the Bulk API to buffer documents and flush them in batches. Document IDs
are derived from the primary key columns, which makes the sink suitable for CDC
workloads that need update and delete semantics. Elasticsearch `2.x` through
`8.x` is supported.
## Key features
@@ -23,33 +29,33 @@ Engine Supported
## Options
-| name | type | required | default value |
-|-------------------------|---------|----------|------------------------------|
-| hosts | array | yes | - |
-| index | string | yes | - |
-| schema_save_mode | string | yes | CREATE_SCHEMA_WHEN_NOT_EXIST |
-| data_save_mode | string | yes | APPEND_DATA |
-| index_type | string | no | |
-| primary_keys | list | no | |
-| key_delimiter | string | no | `_` |
-| auth_type | string | no | basic |
-| username | string | no | |
-| password | string | no | |
-| auth.api_key_id | string | no | - |
-| auth.api_key | string | no | - |
-| auth.api_key_encoded | string | no | - |
-| max_retry_count | int | no | 3 |
-| max_batch_size | int | no | 10 |
-| tls_verify_certificate | boolean | no | true |
-| tls_verify_hostname | boolean | no | true |
-| tls_keystore_path | string | no | - |
-| tls_keystore_password | string | no | - |
-| tls_truststore_path | string | no | - |
-| tls_truststore_password | string | no | - |
-| common-options | | no | - |
-| vectorization_fields | array | no | - |
-| vector_dimensions | int | no | 0 |
-| multi_table_sink_replica | int | no | 1
|
+| name | type | required | default value
| description |
+|-------------------------|---------|----------|------------------------------|-------------|
+| hosts | array | yes | -
| Cluster HTTP addresses in `host:port` form. Multiple hosts are allowed, e.g.
`["host1:9200", "host2:9200"]`. |
+| index | string | yes | -
| Target index name. May contain field placeholders such as `seatunnel_${age}`;
the referenced field must exist in the upstream row. Set `schema_save_mode =
"IGNORE"` when using placeholder indices. |
+| schema_save_mode | string | no | CREATE_SCHEMA_WHEN_NOT_EXIST
| How to handle the target index schema before writing: `RECREATE_SCHEMA`,
`CREATE_SCHEMA_WHEN_NOT_EXIST`, `ERROR_WHEN_SCHEMA_NOT_EXIST`, `IGNORE`. |
+| data_save_mode | string | no | APPEND_DATA
| How to handle existing documents before writing: `DROP_DATA`, `APPEND_DATA`,
`ERROR_WHEN_DATA_EXISTS`. The Elasticsearch sink restricts this option to a
`singleChoice` and explicitly excludes `CUSTOM_PROCESSING`. |
+| index_type | string | no | -
| Deprecated. Maps to Elasticsearch `_type` for clusters that still require it.
Leave unset for modern clusters. |
+| primary_keys | list | no | -
| Primary key fields used to generate the document `_id`. Required for CDC
sources that produce update / delete events. |
+| key_delimiter | string | no | `_`
| Delimiter joining composite keys into `_id` (default `_`). Use a different
character to avoid clashes with field values. |
+| auth_type | string | no | basic
| Authentication mode: `basic` (HTTP Basic with `username`/`password`) or
`api_key` (Elasticsearch API key). |
+| username | string | no | -
| Username for `basic` auth. |
+| password | string | no | -
| Password for `basic` auth. |
+| auth.api_key_id | string | no | -
| API key id for `api_key` auth. |
+| auth.api_key | string | no | -
| API key secret for `api_key` auth. |
+| auth.api_key_encoded | string | no | -
| Base64-encoded `id:secret` API key, alternative to `auth.api_key_id` +
`auth.api_key`. |
+| max_retry_count | int | no | 3
| Maximum retry attempts for a single bulk request. |
+| max_batch_size | int | no | 10
| Maximum number of documents buffered in one bulk request before flushing. |
+| tls_verify_certificate | boolean | no | true
| Validate the server certificate when using HTTPS. |
+| tls_verify_hostname | boolean | no | true |
Validate the server hostname against the certificate. |
+| tls_keystore_path | string | no | -
| Path to a PEM or JKS keystore for client-side mTLS. |
+| tls_keystore_password | string | no | -
| Password for the keystore. |
+| tls_truststore_path | string | no | -
| Path to a PEM or JKS truststore. |
+| tls_truststore_password | string | no | -
| Password for the truststore. |
+| common-options | | no | -
| Sink plugin common parameters. See [Sink Common
Options](../common-options/sink-common-options.md). |
+| vectorization_fields | array | no | -
| Field names whose values should be stored as dense vectors. |
+| vector_dimensions | int | no | 0
| Dimensionality of the dense vectors stored under `vectorization_fields`. Set
together with `vectorization_fields`. |
+| multi_table_sink_replica | int | no | 1
| Number of sink writer replicas when writing multiple tables. |
### hosts [array]
diff --git a/docs/en/connectors/sink/Hive.md b/docs/en/connectors/sink/Hive.md
index 0fdcb5c230..d4ee708d2c 100644
--- a/docs/en/connectors/sink/Hive.md
+++ b/docs/en/connectors/sink/Hive.md
@@ -4,9 +4,15 @@ import ChangeLog from '../changelog/connector-hive.md';
> Hive sink connector
+## Support Those Engines
+
+> Spark<br/>
+> Flink<br/>
+> SeaTunnel Zeta<br/>
+
## Description
-Write data to Hive.
+Write data to Apache Hive tables. The connector uses Hive Metastore for table
management and writes data files (text, CSV, parquet, ORC, JSON) to HDFS (or
S3/OSS when configured). By default it uses two-phase commit so each checkpoint
either commits the whole batch or rolls it back.
:::tip
@@ -33,26 +39,26 @@ By default, we use 2PC commit to ensure `exactly-once`
## Options
-| name | type | required | default value |
-|---------------------------------------|---------|----------|----------------|
-| table_name | string | yes | - |
-| metastore_uri | string | yes | - |
-| compress_codec | string | no | none |
-| hdfs_site_path | string | no | - |
-| hive_site_path | string | no | - |
-| hive.hadoop.conf | Map | no | - |
-| hive.hadoop.conf-path | string | no | - |
-| remote_user | string | no | - |
-| krb5_path | string | no | /etc/krb5.conf |
-| kerberos_principal | string | no | - |
-| kerberos_keytab_path | string | no | - |
-| abort_drop_partition_metadata | boolean | no | false |
-| parquet_avro_write_timestamp_as_int96 | boolean | no | false |
-| overwrite | boolean | no | false |
-| data_save_mode | enum | no | APPEND_DATA |
-| schema_save_mode | enum | no |
CREATE_SCHEMA_WHEN_NOT_EXIST |
-| save_mode_create_template | string | no | - |
-| common-options | | no | - |
+| name | type | required | default value
| description |
+|---------------------------------------|---------|----------|----------------|-------------|
+| table_name | string | yes | -
| Target Hive table name, e.g. `db1.table1`. For multi-table mode, you can use
`${database_name}.${table_name}` and SeaTunnel will substitute the upstream
values. |
+| metastore_uri | string | yes | -
| Hive metastore URI. Comma-separated values enable HA failover. |
+| compress_codec | string | no | none
| Output compression codec. `lzo` and `none` are supported. Parquet / ORC
auto-detect compression. |
+| hdfs_site_path | string | no | -
| Local path of `hdfs-site.xml`. Deprecated for new jobs — prefer
`hive.hadoop.conf` or `hive.hadoop.conf-path`. |
+| hive_site_path | string | no | -
| Local path of `hive-site.xml`. |
+| hive.hadoop.conf | Map | no | -
| Inline Hadoop configuration properties. |
+| hive.hadoop.conf-path | string | no | -
| Directory that contains `core-site.xml`, `hdfs-site.xml`, and
`hive-site.xml`. |
+| remote_user | string | no | -
| Hadoop remote user name used when connecting without Kerberos. |
+| krb5_path | string | no | /etc/krb5.conf
| Path to `krb5.conf` for Kerberos authentication. |
+| kerberos_principal | string | no | -
| Principal for Kerberos authentication. |
+| kerberos_keytab_path | string | no | -
| Path to the keytab file paired with `kerberos_principal`. |
+| abort_drop_partition_metadata | boolean | no | false
| If true, drop partition metadata from Hive Metastore on abort. The data files
in the partition are still removed. |
+| parquet_avro_write_timestamp_as_int96 | boolean | no | false
| Write Parquet `INT96` from a timestamp. Only valid for parquet output. |
+| overwrite | boolean | no | false
| Replace existing data before commit. Equivalent to `data_save_mode =
"DROP_DATA"`. |
+| data_save_mode | enum | no | APPEND_DATA
| How to handle existing data before writing. `APPEND_DATA` (default),
`DROP_DATA`, `CUSTOM_PROCESSING`, `ERROR_WHEN_DATA_EXISTS`. |
+| schema_save_mode | enum | no |
CREATE_SCHEMA_WHEN_NOT_EXIST | How to handle the target table schema before
writing. `RECREATE_SCHEMA`, `CREATE_SCHEMA_WHEN_NOT_EXIST`,
`ERROR_WHEN_SCHEMA_NOT_EXIST`, `IGNORE`. |
+| save_mode_create_template | string | no | -
| Custom DDL template used when auto-creating the target Hive table. Variables:
`${database}`, `${table}`, `${rowtype_fields}`, `${rowtype_partition_fields}`,
`${table_location}`. |
+| common-options | | no | -
| Sink plugin common parameters. See [Sink Common
Options](../common-options/sink-common-options.md). |
### table_name [string]
diff --git a/docs/en/connectors/sink/Mysql.md b/docs/en/connectors/sink/Mysql.md
index ccf3c5c216..c58c758dba 100644
--- a/docs/en/connectors/sink/Mysql.md
+++ b/docs/en/connectors/sink/Mysql.md
@@ -106,7 +106,7 @@ semantics (using XA transaction guarantee).
> This example defines a SeaTunnel synchronization task that automatically
> generates data through FakeSource and sends it to JDBC Sink. FakeSource
> generates a total of 16 rows of data (row.num=16), with each row having two
> fields, name (string type) and age (int type). The final target table is
> test_table will also be 16 rows of data in the table. Before run this job,
> you need create database test and table test_table in your mysql. And if you
> have not yet installed and deployed SeaTunne [...]
-```
+```hocon
# Defining the runtime environment
env {
parallelism = 1
@@ -152,7 +152,7 @@ sink {
> This example not need to write complex sql statements, you can configure
> the database name table name to automatically generate add statements for you
-```
+```hocon
sink {
jdbc {
url =
"jdbc:mysql://localhost:3306/test?useUnicode=true&characterEncoding=UTF-8&rewriteBatchedStatements=true"
@@ -171,7 +171,7 @@ sink {
> For accurate write scene we guarantee accurate once
-```
+```hocon
sink {
jdbc {
url =
"jdbc:mysql://localhost:3306/test?useUnicode=true&characterEncoding=UTF-8&rewriteBatchedStatements=true"
@@ -190,7 +190,7 @@ sink {
> CDC change data is also supported by us In this case, you need config
> database, table and primary_keys.
-```
+```hocon
sink {
jdbc {
url =
"jdbc:mysql://localhost:3306/test?useUnicode=true&characterEncoding=UTF-8&rewriteBatchedStatements=true"
@@ -215,7 +215,7 @@ sink {
> Sync multiple tables from MySQL CDC to target MySQL database, using
> placeholders for dynamic table name mapping
-```
+```hocon
env {
parallelism = 1
job.mode = "STREAMING"
@@ -252,7 +252,7 @@ sink {
> Batch sync multiple tables from MySQL using JDBC Source to another MySQL
> database
-```
+```hocon
env {
parallelism = 1
job.mode = "BATCH"
diff --git a/docs/en/connectors/source/Hive.md
b/docs/en/connectors/source/Hive.md
index 50c07f3273..985d383cd5 100644
--- a/docs/en/connectors/source/Hive.md
+++ b/docs/en/connectors/source/Hive.md
@@ -4,9 +4,15 @@ import ChangeLog from '../changelog/connector-hive.md';
> Hive source connector
+## Support Those Engines
+
+> Spark<br/>
+> Flink<br/>
+> SeaTunnel Zeta<br/>
+
## Description
-Read data from Hive.
+Read data from Apache Hive tables. The connector talks to Hive Metastore for
schema discovery and reads the underlying files from HDFS (or S3/OSS when
configured). The supported file formats include text, CSV, parquet, ORC, JSON,
and markdown. Each table can be read as one batch split, with parallelism and
snapshot/offset resume supported through the checkpoint mechanism.
When using markdown format, SeaTunnel can parse markdown files stored in Hive
tables and extract structured data with elements like headings, paragraphs,
lists, code blocks, and tables. Each extracted element is converted to a
document-element row with the following schema:
- `element_id`: Unique identifier for the element
@@ -48,39 +54,39 @@ Read all the data in a split in a pollNext call. What
splits are read will be sa
## Options
-| name | type | required | default value |
-|-----------------------|--------|----------|----------------|
-| table_name | string | no | Required for single-table mode |
-| table_list | array | no | - |
-| tables_configs | array | no | Deprecated, use `table_list`
instead |
-| use_regex | boolean| no | false |
-| metastore_uri | string | no | Required for single-table mode |
-| krb5_path | string | no | /etc/krb5.conf |
-| kerberos_principal | string | no | - |
-| kerberos_keytab_path | string | no | - |
-| hdfs_site_path | string | no | - |
-| hive_site_path | string | no | - |
-| hive.hadoop.conf | Map | no | - |
-| hive.hadoop.conf-path | string | no | - |
-| remote_user | string | no | - |
-| read_partitions | list | no | - |
-| read_columns | list | no | - |
-| compress_codec | string | no | none |
-| common-options | | no | - |
+| name | type | required | default value | description |
+|-----------------------|--------|----------|----------------|-------------|
+| table_name | string | no | Required for single-table mode |
Target Hive table name in the form `db1.table1`. When `use_regex = true`, this
field uses `databasePattern.tablePattern` to match multiple tables. |
+| table_list | array | no | Deprecated, use `tables_configs`
instead | Deprecated multi-table configuration list. New jobs should use
`tables_configs`. Kept for backward compatibility; will be removed in a future
release. |
+| tables_configs | array | no | - | List of Hive
table configurations for multi-table reading. Each item can override any of the
root-level options. |
+| use_regex | boolean| no | false | Treat
`table_name` as a regular expression that matches multiple tables. Works at the
root level and inside each `table_list` / `tables_configs` entry. |
+| metastore_uri | string | no | Required for single-table mode |
Hive metastore URI. Comma-separated values enable HA failover; whitespace is
ignored. |
+| krb5_path | string | no | /etc/krb5.conf | Path of the
`krb5.conf` file used for Kerberos authentication. |
+| kerberos_principal | string | no | - | Principal for
Kerberos authentication against Hive Metastore / HDFS. |
+| kerberos_keytab_path | string | no | - | Path to the
keytab file paired with `kerberos_principal`. |
+| hdfs_site_path | string | no | - | Local path of
`hdfs-site.xml`. Used to load HDFS HA configuration. Deprecated for new jobs —
prefer `hive.hadoop.conf` or `hive.hadoop.conf-path`. |
+| hive_site_path | string | no | - | Local path of
`hive-site.xml`. |
+| hive.hadoop.conf | Map | no | - | Inline Hadoop
configuration properties (equivalent to entries from `core-site.xml` /
`hdfs-site.xml` / `hive-site.xml`). |
+| hive.hadoop.conf-path | string | no | - | Directory that
contains `core-site.xml`, `hdfs-site.xml`, and `hive-site.xml`. |
+| remote_user | string | no | - | Hadoop remote
user name used when connecting to HDFS / Hive storage without Kerberos. |
+| read_partitions | list | no | - | Restrict the
read to a subset of partitions. All entries must have the same directory depth.
|
+| read_columns | list | no | - | Column
projection list. Only the listed columns are read from the source. |
+| compress_codec | string | no | none | Compression
codec for text / CSV / JSON outputs. `lzo` and `none` are supported. Parquet /
ORC auto-detect compression. |
+| common-options | | no | - | Source plugin
common parameters. See [Source Common
Options](../common-options/source-common-options.md). |
### table_name [string]
Target Hive table name eg: `db1.table1`. When `use_regex = true`, this field
uses `databasePattern.tablePattern` (Hive has no schema) to match multiple
tables from Hive metastore.
-For a single-table source, configure `table_name` and `metastore_uri` at the
root level. For multi-table reading, configure `table_list`. `tables_configs`
is still accepted for compatibility, but `table_list` is preferred.
+For a single-table source, configure `table_name` and `metastore_uri` at the
root level. For multi-table reading, configure `tables_configs`. `table_list`
is still accepted for backward compatibility, but `tables_configs` is the
current option.
### table_list [array]
-List of Hive table configurations for multi-table reading. Each item can
contain `table_name`, `metastore_uri`, `use_regex`, `read_partitions`,
`read_columns`, and the same authentication/Hadoop options as the root
connector block.
+Deprecated multi-table configuration list. Kept for backward compatibility;
new jobs should use `tables_configs`.
### tables_configs [array]
-Deprecated multi-table configuration list. New jobs should use `table_list`.
+List of Hive table configurations for multi-table reading. Each item can
contain `table_name`, `metastore_uri`, `use_regex`, `read_partitions`,
`read_columns`, and the same authentication/Hadoop options as the root
connector block.
### use_regex [boolean]
@@ -155,7 +161,7 @@ Source plugin common parameters, please refer to [Source
Common Options](../comm
### Example 1: Single table
-```bash
+```hocon
Hive {
table_name = "default.seatunnel_orc"
@@ -166,7 +172,7 @@ Source plugin common parameters, please refer to [Source
Common Options](../comm
### Example 2: Metastore URI failover
-```bash
+```hocon
Hive {
table_name = "default.seatunnel_orc"
metastore_uri = "thrift://metastore-1:9083,thrift://metastore-2:9083"
@@ -174,27 +180,10 @@ Source plugin common parameters, please refer to [Source
Common Options](../comm
```
### Example 3: Multiple tables
-> Note: Hive is a structured data source and should be use 'table_list', and
'tables_configs' will be removed in the future.
+> Note: Hive is a structured data source and should use 'tables_configs'; the
older 'table_list' key is deprecated and will be removed in a future release.
> You can also set `use_regex = true` in each table config to match multiple
> tables.
-```bash
-
- Hive {
- table_list = [
- {
- table_name = "default.seatunnel_orc_1"
- metastore_uri = "thrift://namenode001:9083"
- },
- {
- table_name = "default.seatunnel_orc_2"
- metastore_uri = "thrift://namenode001:9083"
- }
- ]
- }
-
-```
-
-```bash
+```hocon
Hive {
tables_configs = [
@@ -213,7 +202,7 @@ Source plugin common parameters, please refer to [Source
Common Options](../comm
### Example 4: Regex matching (whole database / subset)
-```bash
+```hocon
Hive {
metastore_uri = "thrift://namenode001:9083"
@@ -223,7 +212,7 @@ Source plugin common parameters, please refer to [Source
Common Options](../comm
}
```
-```bash
+```hocon
Hive {
metastore_uri = "thrift://namenode001:9083"
@@ -233,7 +222,7 @@ Source plugin common parameters, please refer to [Source
Common Options](../comm
}
```
-```bash
+```hocon
Hive {
metastore_uri = "thrift://namenode001:9083"
@@ -246,7 +235,7 @@ Source plugin common parameters, please refer to [Source
Common Options](../comm
### Example 5: Kerberos
-```bash
+```hocon
source {
Hive {
table_name = "default.test_hive_sink_on_hdfs_with_kerberos"
@@ -270,7 +259,7 @@ Description:
Run the case:
-```bash
+```hocon
env {
parallelism = 1
job.mode = "BATCH"
diff --git a/docs/en/connectors/source/SqlServer.md
b/docs/en/connectors/source/SqlServer.md
index 3ea567eb65..59257c6ab4 100644
--- a/docs/en/connectors/source/SqlServer.md
+++ b/docs/en/connectors/source/SqlServer.md
@@ -38,7 +38,10 @@ import ChangeLog from '../changelog/connector-jdbc.md';
## Description
-Read external data source data through JDBC.
+Read data from Microsoft SQL Server through the official `mssql-jdbc` driver.
The connector supports
+single-table reads via `query` or `table_path`, partitioned reads via
`partition_column`/`partition_lower_bound`/
+`partition_upper_bound`/`partition_num`, and multi-table reads via
`table_list`. Each partition runs as one
+parallel SeaTunnel split.
## Supported DataSource Info
diff --git a/docs/zh/connectors/sink/Elasticsearch.md
b/docs/zh/connectors/sink/Elasticsearch.md
index 3ed7c4a3e1..bc38bccb95 100644
--- a/docs/zh/connectors/sink/Elasticsearch.md
+++ b/docs/zh/connectors/sink/Elasticsearch.md
@@ -27,8 +27,8 @@ import ChangeLog from
'../changelog/connector-elasticsearch.md';
|------------------------|---------|------|------------------------------|
| hosts | array | 是 | - |
| index | string | 是 | - |
-| schema_save_mode | string | 是 | CREATE_SCHEMA_WHEN_NOT_EXIST |
-| data_save_mode | string | 是 | APPEND_DATA |
+| schema_save_mode | string | 否 | CREATE_SCHEMA_WHEN_NOT_EXIST |
+| data_save_mode | string | 否 | APPEND_DATA |
| index_type | string | 否 | |
| primary_keys | list | 否 | |
| key_delimiter | string | 否 | `_` |
diff --git a/docs/zh/connectors/source/Hive.md
b/docs/zh/connectors/source/Hive.md
index fb59b04f11..e864d8bb34 100644
--- a/docs/zh/connectors/source/Hive.md
+++ b/docs/zh/connectors/source/Hive.md
@@ -51,8 +51,8 @@ import ChangeLog from '../changelog/connector-hive.md';
| 名称 | 类型 | 必需 | 默认值 |
|-----------------------|--------|------|---------|
| table_name | string | 否 | 单表模式必填 |
-| table_list | array | 否 | - |
-| tables_configs | array | 否 | 已废弃,请使用 `table_list` |
+| table_list | array | 否 | 已废弃,请使用 `tables_configs` |
+| tables_configs | array | 否 | 多表读取时使用的 Hive 表配置列表,每项可覆盖根配置中的任意选项。 |
| use_regex | boolean| 否 | false |
| metastore_uri | string | 否 | 单表模式必填 |
| krb5_path | string | 否 | /etc/krb5.conf |
@@ -72,15 +72,15 @@ import ChangeLog from '../changelog/connector-hive.md';
目标 Hive 表名,例如:`db1.table1`。当 `use_regex = true` 时,该字段支持 `数据库正则.表正则`(Hive 没有
schema)来匹配 Hive 元存储中的多张表。
-单表读取时,在根配置中填写 `table_name` 和 `metastore_uri`。多表读取时,建议使用
`table_list`。`tables_configs` 仍兼容旧配置,但新作业建议使用 `table_list`。
+单表读取时,在根配置中填写 `table_name` 和 `metastore_uri`。多表读取时,建议使用
`tables_configs`。`table_list` 仍可作为向后兼容的旧配置,但新作业建议使用 `tables_configs`。
### table_list [array]
-Hive 多表读取配置列表。每个元素可以包含
`table_name`、`metastore_uri`、`use_regex`、`read_partitions`、`read_columns`,以及与根配置相同的认证和
Hadoop 配置。
+已废弃的多表读取配置列表,仅为向后兼容保留。新作业请使用 `tables_configs`。
### tables_configs [array]
-已废弃的多表配置列表。新作业请使用 `table_list`。
+Hive 多表读取配置列表。每个元素可以包含
`table_name`、`metastore_uri`、`use_regex`、`read_partitions`、`read_columns`,以及与根配置相同的认证和
Hadoop 配置。
### use_regex [boolean]
@@ -172,24 +172,9 @@ Kerberos 认证的 keytab 文件路径
```
### 示例 3:多表
-> 注意:Hive 是结构化数据源,应使用 `table_list`,`tables_configs` 将在未来移除。
+> 注意:Hive 是结构化数据源,应使用 `tables_configs`,`table_list` 已在新的 API 中废弃,并将在未来移除。
> 也支持在每个表配置中设置 `use_regex = true` 来按正则匹配多表。
-```bash
- Hive {
- table_list = [
- {
- table_name = "default.seatunnel_orc_1"
- metastore_uri = "thrift://namenode001:9083"
- },
- {
- table_name = "default.seatunnel_orc_2"
- metastore_uri = "thrift://namenode001:9083"
- }
- ]
- }
-```
-
```bash
Hive {
tables_configs = [