Why Informatica sequences & Sybase/SQL server identity columns, should not be used?

Informatica provides the sequence object to create surrogate keys during load. These object is also sharable within various mappings hence can be used in parallel. But I will recommend never using it.Â
Hereâ€™s the Reason whyâ€¦.

1. MIGRATION ISSUE:Â The sequence is stored in the Informatica repository. That means it is disconnected from the target database. So during migration of code between environments the sequence has to be reset manually. Bigger problem arise when the tables with data is brought from production environment to QA environment; the mapping cannot be immediately run because the sequences are out of sync.

2. Sequences belong to the target schema and the database and do not belong to the processes as the table object should define & hold values for the attributes and not something external to the system.

3. At first it might seem that Sybase/SQL server does not support sequences but checkout the post & that will solve the problem.

4. Logically it seems that Informatica calls to oracle procedure will slowdown the ETL process but in real world, I have found it not to be true. Additionally this process is only called when a new reference data/Dimension is added. New reference data is not as volatile as transactions; so the adverse effect is nullified.

5. If Identity columns are used instead, it causes more problems as you loose programmatic control on it. Example any type II dimension changes are become a nightmare to manage.

This entry was posted on Sunday, May 7th, 2006 at 3:51 pm and is filed under Informatica, Oracle, SQL Server, Sybase. You can follow any responses to this entry through the RSS 2.0 feed. You can leave a response, or trackback from your own site.

ETLGuru

Why Informatica sequences & Sybase/SQL server identity columns, should not be used?

Leave a Reply