tKafkaOutput properties for Apache Spark Structured Streaming | Talend Components for Jobs Help
Skip to main content Skip to complementary content

tKafkaOutput properties for Apache Spark Structured Streaming

Last updated: 9/30/2026

Use these properties to configure tKafkaOutput running in the Spark Structured Streaming Job framework.

The Spark Structured Streaming tKafkaOutput component belongs to the Messaging family.

This component is available in Talend Real-Time Big Data Platform and Talend Data Fabric.

Information noteImportant: For Spark Streaming Jobs, Talend Studio does not support a specific version of Kafka but relies on a Kafka broker version compatibility provided by Spark. The Kafka broker version supported depends on the Spark version you use. For each Spark version, Talend Studio supports the targeted Kafka broker version provided by Spark. As of today Talend Studio relies on Spark compatibility statement, and therefore supports Kafka broker version 0.10.0 onwards. For more information, see Spark Streaming + Kafka Integration Guide on the official Spark documentation.

Basic settings

Properties Description
Configuration component

Select the tKafkaConfiguration component to use for the Kafka connection settings.

Topic name

Enter the name of the topic you want to publish messages to. This topic must already exist.

Partition

Enter the partition number to be used from the topic.

Information noteNote: If you leave this field blank, Kafka selects the partition according to its producer partitioning behavior. If you enter a value greater than 0, the component verifies that the Kafka topic has a partition range that includes the specified value.
Key

Enter the message key to assign to the Kafka records.

Information noteNote: If you leave this field blank, the key defaults to null.
Compress the data

Select the Compress the data check box to compress the output data.

Set checkpoint location

Enter the path to the directory where Spark stores the checkpoint data for this streaming query.

Checkpointing enables fault tolerance and allows a failed query to resume from where it stopped.

Set trigger Select this check box to configure how often the streaming query processes data. The available trigger types are:
  • Continuous: processes data continuously.
  • Available now: processes all available data in a single batch, then stops the query.
  • Fixed interval: runs the query at a fixed time interval. Enter the interval duration in the field that appears.
Output mode Select the output mode from the drop-down list:
  • Append: adds new rows without modifying existing rows.
  • Update: writes only the rows that were updated since the last trigger.

Advanced settings

Properties Description
Kafka properties

Add the Kafka new producer properties you need to customize to this table.

For more information about the new producer properties you can define in this table, see the section describing the new producer configuration from the official Kafka documentation.

Usage

Usage guidance Description
Usage rule

This component is used as an end component and requires an input link.

This component, along with the Spark Structured Streaming component Palette it belongs to, appears only when you are creating a Spark Structured Streaming Job.

Note that in this documentation, unless otherwise explicitly stated, a scenario presents only Standard Jobs, that is to say traditional Talend data integration Jobs.

Spark Connection

You need to use the Structured Streaming Configuration tab in the Run view to define the connection to a Spark cluster for the whole Job.

This connection is effective on a per-Job basis.

Did this page help you?

If you find any issues with this page or its content – a typo, a missing step, or a technical error – please let us know!