souce file having the columns like
name company
krish IBM
pooja TCS
nandini WIPRO
krish IBM
pooja TCS
if first row will be repeat i want the result like this
name company count
krish IBM 1
pooja TCS 1
nandini WIPRO 1
krish IBM 2
pooja TCS 2
Answer Posted / ankit gosain
Hi ALL,
Job Design:
SourceSeqFile--->SortStage--->Transformer--->TgtSeqFile
1. In Sort Stage, take two key, name & company and then go
to options and create a keyChange column.
2. In transformer stage, create a stage variable of integer
type (say Var1) and write in it's derivation:
if keyChange=1 then 1 else Var1+1
3. Now create a new column in tgt (say count) and in
transformer, assign that Var1 to the derivation of count.
4. Goto o/p tab of transformer and there sort the data on
count column.
You'll get the desired output.
If you have more queries, you can mail me on
ankitgosain@gmail.com
Cheers,
Ankit :)
Is This Answer Correct ? | 3 Yes | 0 No |
Post New Answer View All Answers
What is the purpose of interprocessor stage in server jobs?
How many Key we can define in remove duplicate stage?
what should be ensure to run the sequence job so that if its get aborted in 10th job before 9job should get succeeded?
What is the differentiate between data file and descriptor file?
What is the difference between operational data stage (ods) and data warehouse?
Explain the situation where you have applied SCD in your project?
What is difference between symmetric multiprocessing and massive parallel processing?
What is staging variable?
Is possible to create skid in dim,fact tables?
Can you define merge?
If you want to use a same piece of code in different jobs, how will you achieve this?
Which warehouse using in your datawarehouse
What are the types of containers in datastage?
Field,NVL,INDEX,REPLACE,TRANSLATE,COLESC
Highlight the main features of datastage?