ETL Developer
Key Responsibilities
-
Design, develop, test, and deploy ETL workflows using Ab Initio , Informatica PowerCenter , or IBM DataStage .
-
Develop and optimize complex SQL queries , stored procedures, and database objects.
-
Build and maintain data integration processes for large-scale enterprise data warehouses.
-
Work extensively with Teradata databases for data extraction, transformation, and loading activities.
-
Develop and maintain PySpark applications for big data processing and data transformation.
-
Analyze source-to-target mappings and implement data transformation logic.
-
Perform data validation, reconciliation, and quality checks to ensure data accuracy and consistency.
-
Troubleshoot ETL job failures, performance bottlenecks, and data issues.
-
Collaborate with business analysts, data architects, and stakeholders to gather and understand requirements.
-
Follow best practices for ETL development, coding standards, and documentation.
-
Participate in code reviews, testing, deployment, and production support activities.
-
Monitor ETL processes and ensure timely execution of data loads
Technical Skills
-
Strong experience with at least one ETL tool:
- Ab Initio (Preferred)
- Informatica PowerCenter
- IBM DataStage
-
Advanced SQL development and query optimization skills.
-
Hands-on experience with Teradata databases.
-
Strong knowledge of PySpark and distributed data processing concepts.
-
Experience with data warehousing concepts, dimensional modeling, and ETL methodologies.
- Knowledge of Unix/Linux shell scripting.
-
Experience with data validation, data quality, and reconciliation processes.
-
Understanding of performance tuning and optimization techniques.
Requirements
-
Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
-
5+ years of experience in ETL/Data Engineering development.
-
Strong analytical, problem-solving, and troubleshooting skills.
-
Excellent communication and collaboration abilities.
Originally posted on Himalayas