tPubSubOutput properties for Apache Spark Structured Streaming
Last updated: 9/30/2026These properties are used to configure tPubSubOutput running in the Spark Structured Streaming Job framework.
The Spark Structured Streaming tPubSubOutput component belongs to the Messaging family.
The streaming version of this component is available in Talend Real-Time Big Data Platform and in Talend Data Fabric.
Basic settings
| Properties | Description |
|---|---|
| Schema and Edit schema |
A schema is a row description. It defines the fields (columns) processed by the component. When you create a Spark Job, avoid the reserved word line when naming fields.
Click Edit schema to modify the schema. If you modify a Repository schema, the available options include:
|
| Google Cloud configuration | Enter the name of the subscription to use or create when the selected Topic operation manages a subscription. Pub/Sub publishes messages to the topic; the subscription receives messages from that topic. |
| Pub/Sub topic | Enter the name of the Pub/Sub topic to which you want to publish messages. |
| Pub/Sub subscription | Enter the name of the Pub/Sub subscription associated with the topic to which you want to publish messages. |
| Topic operation | Select how to handle the topic and its subscription:
|
| Set trigger | Select this check box to configure how often the streaming query
processes data. When selected, choose one of the following trigger types:
|
| Output mode | Select the output mode from the drop-down list:
|
| Use ordering key | Select this check box to enable Pub/Sub message ordering. When enabled, the partitionKey column from the input data is used as the message ordering key. |
| Checkpoint location | Enter the path to the directory where Spark stores checkpoint data for the streaming query. |
Usage
| Usage guidance | Description |
|---|---|
| Usage rule |
This component is used as an end component and requires an input link. This component publishes serialized messages to a Google Cloud Pub/Sub topic. This component, along with the Spark Structured Streaming component Palette it belongs to, appears only when you are creating a Spark Structured Streaming Job. |
| Spark Connection |
You need to use the Structured Streaming Configuration tab in the Run view to define the connection to a Spark cluster for the whole Job. This connection is effective on a per-Job basis. |