- Contribute to modernization initiatives by migrating legacy data systems to modern platforms such as PySpark and cloud-based architectures
- Design, develop, and optimize data pipelines using PySpark, SQL, and BigQuery
- Perform data migration activities including data validation, reconciliation, and gap analysis
- Ensure high data quality through rigorous validation and issue resolution processes
- Work closely with onshore and offshore teams to drive project delivery and ensure alignment
- Participate in client discussions, validation reviews, and requirement clarifications
- Support production activities including troubleshooting, monitoring, and resolving issues within SLA
- Analyze and troubleshoot batch job failures and system issues in mainframe environments
- Contribute to automation initiatives to reduce manual efforts and improve efficiency
- Provide data analysis and reporting support for business stakeholders
Required Skills & Experience
Core Technical Skills:
- Strong experience in PySpark, SQL, and BigQuery
- Experience working on data pipeline development and ETL processes
- Hands-on experience with Mainframe technologies (COBOL, JCL, CICS)
- Knowledge of Hive and large-scale data processing frameworks
- Experience with data migration and system modernization projects