tHiveConfiguration properties for Apache Spark Structured Streaming | Talend Components for Jobs Help
Skip to main content Skip to complementary content

tHiveConfiguration properties for Apache Spark Structured Streaming

Last updated: 9/30/2026

These properties are used to configure tHiveConfiguration running in the Spark Structured Streaming Job framework.

The Spark Structured Streaming tHiveConfiguration component belongs to the Databases family.

The streaming version of this component is available in Talend Real-Time Big Data Platform and in Talend Data Fabric.

Basic settings

Properties Description

Property Type

Select the way the connection details will be set.

  • Built-In: The connection details will be set locally for this component. You need to specify the values for all related connection properties manually.

  • Repository: The connection details stored centrally in Repository > Metadata will be reused by this component.

    You need to click the [...] button next to it and in the pop-up Repository Content dialog box, select the connection details to be reused, and all related connection properties will be automatically filled in.

Hive thrift metastore

Enter the location of the metastore of the Hive system to be used by specifying the name of its Host and the number of its listening Port. If HA metastore has been defined for this Hive system, select the Enable high availability check box and in the field that is displayed, enter the URIs of the multiple remote metastore services, each being separated with a comma(,).

Use kerberos authentication

If you are accessing a Hive metastore running with Kerberos security, select this checkbox.

Then you need to enter the Hive principal that should have been defined in the hive-site.xml file of the cluster to be used.

Hive principal uses the value of hive.metastore.kerberos.principal. This is the service principal of the Hive metastore.

Spark catalog

Select the Spark implementation to use.
  • In-memory: select this value if you set the Hive thrift metastore to a Hive metastore that is not an external metastore.
  • Hive: select this value if you set the Hive thrift metastore to an external Hive metastore that exists outside of your cluster.

Usage

Usage guidance Description
Usage rule

This component is used standalone. Add it to the same Job as the Hive-related subJob so that the configuration is available to the whole Job at runtime.

This component, along with the Spark Structured Streaming component Palette it belongs to, appears only when you are creating a Spark Structured Streaming Job.

Spark Connection

In the Structured Streaming Configuration tab of the Run view, define the connection to the Spark cluster for the whole Job. Specify the directory to which the Job's dependent JAR files are transferred so Spark can access them.

This connection is effective on a per-Job basis.

Did this page help you?

If you find any issues with this page or its content – a typo, a missing step, or a technical error – please let us know!