Data Engineer - PySpark
On-sitePune, Maharashtra, India
Job Summary
Build and maintain data architectures, pipelines, and warehouses to ensure accurate, accessible, and secure data transfer. Design scalable storage solutions using Snowflake and AWS tools like Glue, S3, and Lake Formation, while developing ELT pipelines with DBT. Lead or support the delivery of Location Strategy projects, managing end-to-end operational processing and risk controls. Collaborate with data scientists to deploy machine learning models and influence decision-making within the organization. Work in Pune to drive innovation in digital offerings, adhering to strict governance standards and Barclays values.
Required Qualifications
- Hands on experience in pyspark
- strong knowledge on Dataframes, RDD and SparkSQL
- Hands on Experience in developing, testing and maintaining applications on AWS Cloud
- Strong hold on AWS Data Analytics Technology Stack (Glue, S3, Lambda, Lake formation, Athena)
- Design and implement scalable and efficient data transformation/storage solutions using Snowflake
- Experience in Data ingestion to Snowflake for different storage format such Parquet, Iceberg, JSON, CSV etc
- Experience in using DBT (Data Build Tool) with snowflake for ELT pipeline development
- Experience in Writing advanced SQL and PL SQL programs
- Hands On Experience for building reusable components using Snowflake and AWS Tools/Technology
- Should have worked at least on two major project implementations
- Ability to engage with Stakeholders, elicit requirements/ user stories and translate requirements into ETL components
- Ability to understand the infrastructure setup and be able to provide solutions either individually or working with teams
- Good knowledge of Data Marts and Data Warehousing concepts
- Implement Cloud based Enterprise data warehouse with multiple data platform along with Snowflake and NoSQL environment to build data movement strategy
- Knowledge on Abinitio ETL tool
- Exposure to data governance or lineage tools such as Immuta and Alation
- Experience in using Orchestration tools such as Apache Airflow or Snowflake Tasks
Desired Qualifications
- Knowledge on Abinitio ETL tool is a plus
- Some other highly valued skills may include: Ability to engage with Stakeholders, elicit requirements/ user stories and translate requirements into ETL components
- Ability to understand the infrastructure setup and be able to provide solutions either individually or working with teams
- Good knowledge of Data Marts and Data Warehousing concepts
- Implement Cloud based Enterprise data warehouse with multiple data platform along with Snowflake and NoSQL environment to build data movement strategy
- You may be assessed on key critical skills relevant for success in role, such as risk and controls, change and transformation, business acumen, strategic thinking and digital and technology, as well as job-specific technical skills
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.