Difference in the implementation of lookup and join
stages,in joining two tables?
Answers were Sorted based on User's Feedback
Answer / kiran
Hai This is Kiran...
If u want to join more than one table ,u can use join,lookup
and merge also.
Join: it is used join more than one table based one key
column .it can perform 4 join as inner join,left outer
join,right join and full outer join.
Lokk-UP:it is used join more than one table,but not necesary
to join based ont he key column but it need data-type.it
give reference link and single out put link and give reject
link also.it can perform inner join and left outer jojn.
main difference: if the huge amount of the data contain in
reference table refer to join else look-up.
Is This Answer Correct ? | 14 Yes | 2 No |
Answer / krishna
Generally we are using lookup for comparision purpose,
based on the reference table size we r using join or lookup.
If the reference table size is less than the main table
then use lookup stage other wise use join stage.
In lookup we have two types:
Normal lookup - reference lookup is less no of rows
compared to main table
sparse lookup - reference lookup is more no of rows
compared to main table.
But the peformance wise use join if the reference table has
more data.
Regards,
Krishna
Is This Answer Correct ? | 8 Yes | 0 No |
Answer / zulfi123786
Hi This is Zulfi
Basically Join is used when you have large amount of data
about in millions and it performs inner join,left
outer,right outer and full outer joins
The join stage requires the incomming data to be hash
partitioned and sorted on the joining keys
The look up is used when the reference records are fewer in
number about less than one lakh and it doesnot require the
incomming source data to be sorted, instead the refrence
link should be in Entire partition mode.
In look up there are two types
Normal and Sparse
Sparse is available only when the reference is a database.
usually Normal has to be used unless when the refrence to
source rows ratio is 100:1
Is This Answer Correct ? | 8 Yes | 0 No |
Answer / sadanand
HI All,
I would like to add one more point to JOIN.
To achieve full outer join the number of inputs need would
be only two.
The Primay table need to be sorted.
Memory used is very less compared to Lookup.
Regards,
Sadanand.
Is This Answer Correct ? | 2 Yes | 0 No |
Answer / indian
Hi Zulfi..you are answer is more explained one and clear
we will go for Merge if we want the rejected data for every
update link
Is This Answer Correct ? | 0 Yes | 0 No |
how do u capture duplicates through sort & transformer
Define oconv () and iconv () functions in datastage?
WHAT are unix quentios in datastage
Have you have ever worked in unix environment and why it is useful in datastage?
i have 3 diffrent tables. 1) US rate data 2)CANADA rate data and 3)MEXICO rate data. All 3 tables have 6 collumns each. 4 collumns are commun to all tables and 2 are diffrent. Now at target i want single table say Country rate which will have (4+2+2+2+1 flag) 11 collumns. I will add a flag collumn which will indicate country and will put nullable collumns which are not common to other. How i can implement this in datastage?
Explain usage analysis in datastage?
if a column contains data like abc,aaa,xyz,pwe,xok,abc,xyz,abc,pwe,abc,pwe,xok,xyz,xxx,abc, roy,pwe,aaa,xxx,xyz,roy,xok.... how to send the unique data to one source and remaining data to another source????
Can you explain kafka connector?
What is process model?
Where the datastage stored his repository?
What is a range lookup?
How the ipc stage work?