0% found this document useful (0 votes)
14 views9 pages

Informatica Transformations Interview Guide

The document summarizes a YouTube video discussing Informatica PowerCenter interview questions, focusing on transformations. Key concepts such as mappings, active vs. passive transformations, and various types of transformations like Router, Filter, and Lookup are defined and explained. The video provides detailed answers to interview questions, covering the functionality and differences between different transformation types in Informatica.

Uploaded by

upscstudent0
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
14 views9 pages

Informatica Transformations Interview Guide

The document summarizes a YouTube video discussing Informatica PowerCenter interview questions, focusing on transformations. Key concepts such as mappings, active vs. passive transformations, and various types of transformations like Router, Filter, and Lookup are defined and explained. The video provides detailed answers to interview questions, covering the functionality and differences between different transformation types in Informatica.

Uploaded by

upscstudent0
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd

‭ ased on the analysis of the provided YouTube video transcript and our previous conversation history, the‬

B
‭video is the second part of a series discussing important Informatica PowerCenter interview questions,‬
‭presented by Almas. This segment specifically focuses on‬‭Transformations‬‭in Informatica.‬

‭The video first defines key concepts:‬

‭●‬ A ‭ ‬‭Mapping‬‭is described as a collection of sources and targets linked together by a series of‬
‭transformations. It must contain at least a source and a target.‬
‭●‬ ‭A‬‭Transformation‬‭is defined as a set of rules or instructions applied to data to change and move‬
‭it from source to target.‬

‭ ollowing these definitions, the video presents several interview questions and provides detailed answers,‬
F
‭explaining various Informatica transformations and concepts. Here are the questions and answers‬
‭covered:‬

‭●‬ E
‭ xplain the difference between active and passive transformations.‬‭(This was also covered in‬
‭the previous video).‬

‭○‬ A ‭ n‬‭Active Transformation‬‭changes the‬‭number of input rows‬‭that pass through it. The‬
‭number of incoming rows is different from the number of outgoing rows.‬
‭○‬ ‭A‬‭Passive Transformation‬‭does‬‭not change the number of input rows‬‭. The number of‬
‭incoming rows is equal to the number of outgoing rows.‬
‭‬ D
● ‭ ifferences between router and filter.‬

‭‬ B
○ ‭ oth work on‬‭conditions‬‭to filter data.‬
‭○‬ ‭Router‬‭: Works on‬‭multiple conditions‬‭. It has‬‭three groups‬‭: input, output, and a‬‭default‬
‭group‬‭for rows not matching any condition. Records not matching conditions go to the‬
‭default group and are‬‭not lost‬‭. It works like‬‭IF...ONLY IF‬‭or‬‭ CASE‬‭statements in‬
‭SQL.‬
‭○‬ ‭Filter‬‭: Works on‬‭only one condition‬‭. It‬‭does not have any groups‬‭. Records not‬
WHERE‬‭clause in SQL.‬
‭matching the condition are‬‭lost‬‭. It works like a‬‭
‭‬ W
● ‭ hy Sorter is an active transformation?‬

‭○‬ W ‭ hile sorting itself doesn't reduce rows, the Sorter transformation is active because it has‬
‭a‬‭distinct option‬‭.‬
‭○‬ ‭If the distinct option is selected, it retrieves only distinct rows, which changes the number‬
‭of input rows relative to the output rows.‬
‭‬ W
● ‭ hat is the use of Source Qualifier?‬

‭○‬ I‭ t's a transformation that is‬‭automatically created‬‭when you drag a source into a‬
‭mapping.‬
‭○‬ ‭It represents the‬‭source data‬‭.‬
‭○‬ ‭It can perform various functions like‬‭joining‬‭,‬‭filtering rows‬‭,‬‭sorting input‬‭, selecting‬
‭distinct records‬‭, and allowing a‬‭custom SQL query‬‭.‬
‭●‬ ‭What are the different ways in which we can filter the data in Informatica?‬

‭○‬ D
‭ ata can be filtered using the‬‭Source Qualifier‬‭,‬‭Router‬‭,‬‭Filter‬‭, and‬‭Joiner‬
‭transformations.‬
‭‬ W
● ‭ hat are the different types of groups in Router?‬

‭○‬ T‭ he Router has an‬‭input group‬‭, an‬‭output group‬‭, and a‬‭default group‬‭. Records that do‬
‭not match any defined condition are directed to the default group.‬
‭‬ W
● ‭ hat is Expression Transformation?‬

‭ ‬ I‭ t is a‬‭passive transformation‬‭.‬

‭○‬ ‭It is used to perform‬‭calculations‬‭or‬‭manipulate data‬‭by applying formulas.‬
‭○‬ ‭It processes‬‭one record at a time‬‭, and therefore, it does not change the number of input‬
‭rows.‬
‭‬ W
● ‭ hat is a Lookup transformation?‬

‭ ‬ ‭Used to‬‭look up data‬‭from a particular table, view, or flat file.‬



‭○‬ ‭Based on the data found (or not found), actions are decided for the input record.‬
‭○‬ ‭It consists of‬‭input, output, lookup, and return ports‬‭.‬
‭ ‬ ‭Different types of Lookup transformations.‬

‭○‬ C ‭ onnected Lookup‬‭: Integrated directly into the data flow of a mapping. Can return‬
‭more than one value‬‭.‬
‭○‬ ‭Unconnected Lookup‬‭: Not part of the main data flow. It is called using‬‭lkp functions‬‭. It‬
‭can only return‬‭one value‬‭.‬
‭‬ A
● ‭ n unconnected lookup can have how many input parameters?‬

‭○‬ A
‭ n unconnected lookup can have‬‭any number of input parameters‬‭but will have‬‭only‬
‭one output parameter‬‭.‬
‭‬ N
● ‭ ame different types of lookup caches.‬

‭ ‬ ‭The different types are‬‭Static‬‭,‬‭Dynamic‬‭,‬‭Persistent‬‭, and‬‭Shared‬‭.‬



‭ ‬ ‭How do you differentiate static and dynamic caches?‬

‭○‬ D ‭ ynamic Cache‬‭: Refreshes automatically as soon as there are‬‭changes (insert, update,‬
‭delete)‬‭in the lookup table during a mapping run. It must be‬‭explicitly configured‬‭as‬
‭dynamic.‬
‭○‬ ‭Static Cache‬‭: Does‬‭not change‬‭within a single mapping run, even if the lookup table is‬
‭updated. It only reflects changes in subsequent runs. By default, caches are static.‬
‭‬ W
● ‭ hat is Persistent Cache?‬

‭○‬ A
‭ cache type used for values that are‬‭expected to remain constant‬‭or change very‬
‭infrequently (like the capital of a country).‬
‭○‬ I‭ t stores lookup data persistently across mapping runs to‬‭reduce lookup time‬‭. This‬
‭requires enabling the‬‭lookup cache persistent property‬‭.‬
‭‬ H
● ‭ ow can we delete duplicate rows from a flat file?‬

‭‬ T
○ ‭ he Source Qualifier for flat files disables options like distinct selection or custom SQL.‬
‭○‬ ‭Duplicates from a flat file can be removed using a‬‭Sorter transformation‬‭by selecting‬
‭the‬‭distinct option‬‭within it.‬
‭‬ D
● ‭ ifferences between SQL override and a Lookup override.‬

‭○‬ S ‭ QL Override‬‭: An explicitly written SQL query in the‬‭Source Qualifier‬‭. Primarily used‬
‭to‬‭limit the number of rows‬‭read from the source into the mapping. Requires explicit‬
‭inclusion of an‬‭ORDER BY‬‭clause if needed. Can return multiple matching records.‬
‭○‬ ‭Lookup Override‬‭: An explicitly written query within the‬‭Lookup transformation‬‭.‬
‭Used to‬‭limit the number of rows‬‭cached or looked up from the lookup source to‬
‭improve performance. Uses an‬‭ ORDER BY‬‭clause by default‬‭. It is designed to return‬
‭only one record‬‭even if multiple records match the condition.‬
‭‬ W
● ‭ hat is the Rank transformation?‬

‭ ‬ I‭ t is an‬‭active transformation‬‭.‬

‭○‬ ‭It is used to‬‭rank data‬‭, allowing the selection of the top or bottom records.‬
‭○‬ ‭It can identify records with the‬‭largest or smallest numeric values‬‭based on a specified‬
‭port.‬
‭‬ W
● ‭ hat is Joiner transformation?‬

‭○‬ A ‭ significant transformation used to perform‬‭joins‬‭similar to those in SQL (like inner,‬


‭outer, etc.).‬
‭○‬ ‭Informatica's Joiner supports normal join, full outer join, master outer, and detail outer‬
‭joins.‬
‭‬ W
● ‭ hat is Aggregator transformation?‬

‭ ‬ I‭ t is an‬‭active transformation‬‭.‬

‭○‬ ‭Used to perform‬‭aggregate functions‬‭(like sum, min, max, average) on data.‬
GROUP BY‬‭clause‬‭in SQL, allowing data to be grouped‬
‭○‬ ‭It functions similarly to the‬‭
‭before applying aggregate functions.‬
‭‬ W
● ‭ hat is a Sequence Generator?‬

‭‬ T
○ ‭ his transformation‬‭generates a series of numbers‬‭.‬
‭○‬ ‭It is commonly used when a source lacks a‬‭unique primary key‬‭or any primary key, to‬
‭help uniquely identify records in the target database.‬
‭‬ W
● ‭ hat is a Union transformation and what are its restrictions?‬

‭ ‬ I‭ t is an‬‭active transformation‬‭.‬

‭○‬ ‭It‬‭merges data from multiple input sources‬‭into a single output stream.‬
‭○‬ I‭ t works like‬‭ UNION ALL‬‭in SQL, meaning it‬‭does not automatically remove‬
‭duplicate rows‬‭.‬
‭○‬ ‭Restrictions‬‭:‬
‭■‬ ‭All input sources must have the‬‭same number of ports‬‭, and the data type of‬
‭corresponding ports must be the‬‭same‬‭.‬
‭■‬ ‭It‬‭does not remove duplicates‬‭.‬
‭■‬ ‭You‬‭cannot use Sequence Generator or Update Strategy‬‭transformations with‬
‭the Union transformation.‬

‭1.‬ ‭What is actually a router transformation?‬

‭‬ A
○ ‭ ‬‭router transformation‬‭is used to‬‭filter rows in a mapping‬‭.‬
‭○‬ ‭It allows you to specify‬‭more than one condition‬‭, unlike a Filter transformation‬
‭which specifies only one condition.‬
‭○‬ ‭A Router is an‬‭active transformation‬‭.‬
‭2.‬ W
‭ hat are the different groups in router transformation?‬

‭‬ T
○ ‭ here are two types of general groups:‬‭input‬‭and‬‭output groups‬‭.‬
‭○‬ ‭Within the output groups, there are two types:‬‭user-defined‬‭and the‬‭default‬
‭group‬‭.‬
‭○‬ ‭Rows that meet the conditions flow into the‬‭user-defined groups‬‭.‬
‭○‬ ‭Rows that‬‭do not meet the conditions‬‭are sent to the‬‭default group‬‭.‬
‭3.‬ C
‭ an you connect ports of two output groups from router to a single target?‬

‭‬ T
○ ‭ he answer is‬‭no‬‭.‬
‭○‬ ‭You‬‭cannot connect more than one output port‬‭from a Router to a single‬
‭target.‬
‭○‬ ‭One output group will go to only one particular target.‬

‭●‬ ‭What is actually Expression transformation in Informatica?‬

‭○‬ E ‭ xpression transformation is used to‬‭manipulate row-wise data‬‭through the‬


‭mapping.‬
‭○‬ ‭It means you can perform‬‭calculations on the data row-wise‬‭, on each row.‬
‭○‬ ‭It is a‬‭passive transformation‬‭.‬
‭○‬ ‭It can be used to perform any‬‭non-aggregate calculations‬‭, like adding‬
‭something to a value, subtracting something, or concatenating values.‬
‭‬ H
● ‭ ow many types of ports are there in an expression transformation?‬

‭○‬ T
‭ here are‬‭three types of ports‬‭in an expression transformation:‬‭input‬‭,‬‭output‬‭,‬
‭and a‬‭variable port‬‭.‬
‭ ‬ ‭Complex calculations are performed using the‬‭variable port‬‭.‬

‭ ‬ ‭What is the execution order of ports in an expression transformation?‬

‭‬ A
○ ‭ ll the ports are executed‬‭from top to bottom serially‬‭.‬
‭○‬ ‭The execution is done in the following groups:‬
‭■‬ ‭First, all‬‭input ports‬‭are given the values.‬
‭■‬ ‭Then, all‬‭variable ports‬‭are executed (calculated based on inputs).‬
‭■‬ ‭Lastly, all the‬‭output expressions‬‭are executed so values can be sent to‬
‭the output port.‬
‭‬ W
● ‭ hat is the use of variable port in an expression transformation?‬

‭○‬ T ‭ he use of variable ports is to‬‭temporarily store the data‬‭while processing is‬
‭going on.‬
‭○‬ ‭Variable ports simplify‬‭complex calculations‬‭.‬
‭○‬ ‭For example, they can be used in scenarios where you need to‬‭extract the‬
‭month part from a date‬‭.‬
‭‬ H
● ‭ ow to generate sequence numbers using expression transformation?‬

‭○‬ T ‭ o generate sequence numbers using expression transformation, you can use a‬
‭variable port‬‭.‬
‭○‬ ‭Create a variable port in the expression transformation and‬‭increment it by one‬
‭for every new row‬‭.‬
‭○‬ ‭Then, assign this variable to an‬‭output port‬‭.‬
‭○‬ ‭Every time a new row comes, the value will increment by one.‬

‭ ased on the provided YouTube video transcripts (,), here are the questions asked about the‬
B
‭Informatica Lookup Transformation and their answers:‬

‭1.‬ ‭What is actually a lookup in ETL?‬

‭‬ A
○ ‭ lookup is a‬‭common ETL operation‬‭used to‬‭calculate a field's value‬‭.‬
‭○‬ ‭It works by providing input parameters to the lookup transformation.‬
‭○‬ ‭Based on these parameters, the lookup‬‭queries a database or other data‬
‭source‬‭to return a value or a data set (list of values).‬
‭○‬ ‭The lookup performs this by‬‭joining data‬‭in the input columns with columns in‬
‭the referenced data set.‬
‭○‬ ‭A lookup transformation can look up data in a‬‭flat file, relational table, view, or‬
‭synonym‬‭.‬
‭2.‬ W
‭ hat are the tasks of lookup transformation?‬

‭○‬ T ‭ o‬‭get a related value‬‭of a field based on input parameters by referencing a data‬
‭source.‬
‭○‬ ‭To‬‭perform calculations‬‭using the value retrieved from the lookup table.‬
‭○‬ T
‭ o‬‭update slowly changing dimensions (SCD)‬‭; the lookup is used to check if a‬
‭record from the source already exists in the target table based on source data,‬
‭and then records are flagged as insert, update, or delete.‬
‭3.‬ W
‭ hat is connected and unconnected lookup transformation?‬

‭○‬ A ‭ ‬‭connected lookup‬‭is‬‭connected in the mapping pipeline‬‭like other‬


‭transformations. It receives data from the source, performs the lookup, and‬
‭returns data to the pipeline. It behaves similarly to other connected‬
‭transformations.‬
‭○‬ ‭An‬‭unconnected lookup‬‭is‬‭not connected‬‭to other transformations in the‬
‭mapping pipeline. Other transformations that want to use it do so with the use of‬
lkp‬‭function‬‭, passing the required input parameters.‬
‭the‬‭
‭4.‬ W
‭ hat are the differences between connected and unconnected lookup?‬

‭○‬ D ‭ ata Flow:‬‭Connected lookup participates in data flow and receives input directly‬
‭from the pipeline. Unconnected lookup receives values using the‬‭ lkp‬‭function.‬
‭○‬ ‭Cache Type:‬‭Connected lookup can use both‬‭dynamic and static cache‬‭.‬
‭Unconnected lookup can only use‬‭static cache‬‭.‬
‭○‬ ‭Return Values:‬‭Connected lookup can return‬‭more than one value‬‭(a data set).‬
‭Unconnected lookup can return‬‭only one value‬‭.‬
‭○‬ ‭Caching of Ports:‬‭Connected lookup caches all lookup columns. Unconnected‬
‭lookup caches only the lookup output ports in the lookup conditions and the‬
‭return port.‬
‭○‬ ‭Default Values:‬‭Connected lookup supports user-defined default values to return‬
‭when lookup conditions are not satisfied. Unconnected lookup does not support‬
‭user-defined default values.‬
‭5.‬ H
‭ ow do you handle multiple matches in lookup transformation?‬

‭○‬ Y ‭ ou use the‬‭"lookup policy on multiple match"‬‭option in the lookup‬


‭transformation.‬
‭○‬ ‭This option determines which row the lookup transformation returns if it finds‬
‭multiple rows that match the lookup conditions.‬
‭○‬ ‭You can configure it to return‬‭any row‬‭(first, last, or any in between rows‬
‭matching) or to‬‭report an error‬‭.‬
‭6.‬ W
‭ hat are the options available to configure a lookup cache?‬

‭○‬ T
‭ he available options are:‬‭persistent cache, static cache, dynamic cache,‬
‭shared cache, re-cache from lookup source, and pre-build lookup cache‬‭.‬
‭7.‬ W
‭ hat is a cached lookup transformation?‬

‭○‬ W ‭ hen a lookup is cached, the integration service‬‭builds a cache memory‬‭when‬


‭it processes the first row of data.‬
‭○‬ ‭It stores the‬‭condition values in the index cache‬‭and‬‭output values in the‬
‭data cache‬‭.‬
‭○‬ T
‭ he integration service will‬‭query the cache‬‭for each new row that enters the‬
‭transformation, and then the lookup source if needed.‬
‭8.‬ W
‭ hat is uncached lookup?‬

‭‬ F
○ ‭ or an uncached lookup, the integration service‬‭will not build any cache‬‭.‬
‭○‬ ‭For each new row that enters the lookup transformation, the integration service‬
‭will‬‭directly query the lookup source‬‭and return a value.‬
‭9.‬ W
‭ hat is a dynamic cache?‬

‭‬ T
○ ‭ he dynamic cache‬‭represents the data in the target‬‭.‬
‭○‬ ‭The integration service builds the cache when it processes the first lookup‬
‭request.‬
‭○‬ ‭It queries the cache based on the lookup condition for each row.‬
‭○‬ ‭The integration service‬‭updates the lookup cache‬‭as it passes rows to the‬
‭target by either inserting a new row, updating an existing row, or making no‬
‭change.‬
‭10.‬‭What is a persistent cache?‬

‭○‬ I‭f the data in the lookup source does‬‭not change between session runs‬‭, a‬
‭persistent cache can be used to improve performance.‬
‭○‬ ‭When a session runs for the first time, the integration service creates a cache‬
‭file.‬
‭○‬ ‭Instead of deleting the file after the session completes, it is‬‭saved to the disk‬‭.‬
‭○‬ ‭The next time the session runs, the integration service will‬‭build memory from‬
‭the saved cache file‬‭.‬
‭11.‬‭What is a shared cache?‬

‭○‬ I‭n a shared cache,‬‭multiple lookup transformations‬‭in the same mapping can‬
‭be configured to share a‬‭single lookup cache‬‭.‬
‭○‬ ‭The integration service builds the cache when it processes the first lookup‬
‭transformation that shares the cache, and this cache is then used by subsequent‬
‭lookup transformations configured to share it.‬

‭●‬ ‭What is actually aggregator transformation?‬

‭‬ A
○ ‭ n aggregator is an‬‭active‬‭and‬‭connected transformation‬‭.‬
‭○‬ ‭It‬‭performs aggregate calculations‬‭like min, max, average, count, sum, and‬
‭many others.‬
‭‬ W
● ‭ hat is actually a joiner transformation?‬

‭○‬ A ‭ joiner is a transformation that‬‭joins the data between two heterogeneous‬


‭sources‬‭.‬
‭○‬ ‭It can also join the data from the same source.‬
‭○‬ A ‭ heterogeneous source refers to different types of sources within a single‬
‭mapping, such as an Oracle table, a flat file, and XML. Homogeneous sources, in‬
‭contrast, consist of only one type of source, like solely Oracle tables or flat files.‬
‭○‬ ‭A joiner will join the sources only if there is at least one matching column‬
‭between them, and the data type of the matching column should be the same.‬
‭○‬ ‭The Joiner is an‬‭active and connected transformation‬‭.‬
‭‬ H
● ‭ ow many joiners are required to join n number of sources?‬

n‬‭number of sources, you need‬‭


‭ ‬ ‭To join‬‭
○ n‬‭minus 1 joiners‬‭.‬
‭○‬ ‭For instance, if you have 10 sources, you would need 10 - 1 = 9 joiners.‬
‭ ‬ ‭What are the different types of joints?‬

‭○‬ ‭There are‬‭four types of joints in Informatica‬‭:‬


‭■‬ ‭Normal join‬‭: The integration service discards all rows from both the‬
‭master and detailed sources that do not match the join condition. Only the‬
‭rows that match the condition proceed.‬
‭■‬ ‭Master outer join‬‭: This type keeps all the rows from the detailed source‬
‭and only the matching rows from the master source.‬
‭■‬ ‭Detail outer join‬‭: This type keeps all the rows from the master source‬
‭and only the matching rows from the detail source. Any unmatched rows‬
‭from the detail source are discarded.‬
‭■‬ ‭Full outer join‬‭: In this type, rows from both the master and detail sources‬
‭are kept; nothing is discarded.‬
‭‬ W
● ‭ hich table should be selected as the master table in Joiner transformation?‬

‭○‬ I‭n the properties of the Joiner transformation, you can select which source is the‬
‭master and which is the detail.‬
‭○‬ ‭During execution, the‬‭master source is cached into memory‬‭for the joining‬
‭process.‬
‭○‬ ‭Therefore, to minimize memory usage, the source with the‬‭less number of‬
‭records should always be selected as the master source‬‭.‬
‭‬ H
● ‭ ow to improve the performance of Joiner transformation?‬

‭○‬ J ‭ oin sorted data whenever possible‬‭. It is recommended to always sort the data‬
‭before joining.‬
‭○‬ ‭For an‬‭unsorted Joiner‬‭transformation, designate the source with‬‭fewer rows‬
‭as the master source‬‭.‬
‭○‬ ‭If the data is‬‭sorted‬‭, then designate the source with‬‭fewer duplicate keys as‬
‭the master source‬‭.‬
‭‬ W
● ‭ hich one is better Joiner or a lookup?‬

‭○‬ A
‭ ‬‭Joiner is a better choice if you want to join heterogeneous sources‬‭, such‬
‭as flat files with tables.‬
‭○‬ A
‭ ‬‭Lookup is a better option if you want to join homogeneous sources‬‭,‬
‭meaning if both tables are from the same type, like both being Oracle tables.‬
‭‬ W
● ‭ hat is the difference between a joiner and a lookup transformation?‬

‭○‬ O ‭ ne difference is that in‬‭Lookup, the default join type is a left outer join‬‭. In‬
‭Joiner, you can explicitly define the join type‬‭as a left outer, right outer, full‬
‭outer, or normal join.‬
‭○‬ ‭Another difference is that‬‭Lookup allows you to use an SQL override‬‭, which‬
‭means you can write an SQL query within the Lookup transformation. This‬‭is not‬
‭possible in a Joiner‬‭transformation.‬

‭ ased on the provided YouTube video transcript, here are the questions asked about the‬
B
‭Informatica Union and Rank Transformations and their corresponding answers:‬

‭●‬ ‭What is actually a union transformation?‬

‭ ‬ ‭The Union transformation is used to‬‭merge the data from multiple sources‬‭.‬

‭○‬ ‭It works similar to the‬‭union all in SQL statement‬‭.‬
‭ ‬ ‭Now Union gives Union all output but how will you get Union output to achieve the‬

‭union output from Union transformation?‬

‭○‬ O ‭ ne way is to‬‭pass the output of the Union transformation to a sorter‬


‭transformation‬‭. In the properties of the sorter, you‬‭check the option "select‬
‭distinct"‬‭. This will give the union output.‬
‭○‬ ‭Another way is to‬‭pass the output of Union transformation to aggregator‬
‭transformation‬‭and in aggregator transformation‬‭specify all ports as Group by‬
‭ports‬‭. These are the two ways you can achieve Union all output as Union output.‬
‭‬ W
● ‭ hy Union transformation is an active transformation?‬

‭○‬ A ‭ union transformation is active because it‬‭combines two or more data‬


‭streams into one data stream‬‭.‬
‭○‬ ‭Although the total number of rows passing into the union is the same as the total‬
‭number of rows passing out, and the sequence of rows from any given input‬
‭stream is preserved in the output, the‬‭positions of the rows are not preserved‬‭.‬
‭○‬ ‭For example, row number one from input stream one might not be row number‬
‭one in the output stream.‬
‭○‬ ‭Union also‬‭does not guarantee that the output is repeatable‬‭. Because of‬
‭these reasons, Union is considered an active transformation.‬
‭‬ W
● ‭ hat is rank transformation?‬

‭‬
○ ‭ rank transformation is used to select‬‭top or bottom rank of data‬‭.‬
A
‭○‬ ‭It is used to select the largest or smallest numeric value in a port or a group.‬
‭○‬ ‭Rank is an‬‭active and connected transformation‬‭.‬
‭○‬ ‭It can be used in cases like wanting 5 records of employees having the highest‬
‭salary or getting the top 5 salaried employees Department wise.‬

You might also like