tPatternUnmasking properties for Apache Spark Structured Streaming
Last updated: 9/30/2026These properties are used to configure tPatternUnmasking running in the Spark Structured Streaming Job framework.
The Spark Structured Streaming tPatternUnmasking component belongs to the Data Quality family.
This component is supported on Local Spark 3.5.x and Databricks/EMR with Spark 3.x.
This component is available in Talend Real-Time Big Data Platform and Talend Data Fabric.
Basic settings
| Properties | Description |
|---|---|
|
Schema and Edit Schema |
|
|
Modifications |
Define in the table what fields to unmask and how to unmask them: Use the same settings for the Field type, Values, Path, Range and Date Range columns as the ones used for masking the input data with the tPatternMasking component. Column to unmask: Select the column from the input flow that contains the data to be unmasked. Each column is processed sequentially, meaning that data unmasking operations will be performed on the data from the first column, the second column, and so on. In a colum, each data field is a fixed length field, except the last data field. For fixed length fields, each value must contain the same number of characters, for example: "30001,30002,30003" or "FR,EN". In a column, the last Enumeration or Enumeration from file data field is a variable length field. For variable length fields, each value might not always contain the same number of characters, for example: "30001,300023,30003" or "FR,ENG".
Field type: Select
the field type the data belongs to.
In the Values, Path, Range and Date Range, values must be enclosed in double quotes. When the input data is invalid, meaning that a value does not match the pattern defined in the component, the generated value is null. |
Usage
| Usage guidance | Description |
|---|---|
| Usage rule |
This component is used as an intermediate step. This component, along with the Spark Structured Streaming component Palette it belongs to, appears only when you are creating a Spark Structured Streaming Job. |
| Spark Connection |
You need to use the Structured Streaming Configuration tab in the Run view to define the connection to a Spark cluster for the whole Job. This connection is effective on a per-Job basis. |