Job Responsibilities:
- Develop and maintain ETL/data ingestion pipelines using PySpark and AWS services.
- Design and implement Lambda-based event-driven workflows.
- Optimize Spark jobs for performance, scalability, and cost efficiency.
- Perform data transformation, cleansing, and enrichment.
- Monitor and troubleshoot ETL workflows to ensure reliability and high data quality.
Technical Skills:
- Python, PySpark
- AWS EMR / AWS Glue / AWS Lambda
- AWS IAM, ECS, S3
- Athena, Redshift, CloudFormation (Good to Have)
Preferred:
Experience in AWS-based Data Engineering projects.
BFSI domain experience is an added advantage.
Role
Data Scientist
Qualifications
BACHELOR OF ENGINEERING, BACHELOR OF SCIENCE (B.Sc), BACHELOR OF TECHNOLOGY, MASTER OF ENGINEERING, Master of Information Technology