Data Engineer
RemoteUnited States
Job Summary
Design and build robust data pipelines and ETL processes to support data ingestion, transformation, and integration from various sources into cloud-based data warehouses. Implement data validation, cleansing, and quality assurance processes to ensure data integrity for analytics and reporting. Collaborate with data scientists and analysts to deliver solutions that enhance analytical capabilities. Optimize storage and processing solutions in AWS Cloud using Glue, Redshift, S3, and Apache Spark. Develop documentation for workflows and architectures while monitoring pipeline performance to resolve issues. Requires 5+ years of federal data engineering experience, proficiency in Python, SQL, and Scala, and a Bachelor's degree. U.S. citizenship and security clearance eligibility are mandatory.
Required Qualifications
- Minimum of 5 years of experience in data engineering or related fields in a public sector or federal government setting, with a strong focus on ETL processes and data pipeline development
- Proficiency in programming languages such as Python, SQL, and Scala, with practical experience in data processing frameworks
- Hands-on experience with cloud platforms, specifically AWS, and tools like AWS Glue, Apache Spark, DataStage, and Redshift
- Strong understanding of data governance and compliance standards, particularly in federal environments
- Excellent analytical and problem-solving skills, with the ability to work independently and manage multiple priorities
- Effective communication skills, with the ability to collaborate with technical and non-technical stakeholders
- Bachelor's degree in Computer Science, Information Technology, Data Science, or a related field
- U.S. Citizenship
- Ability and willingness to obtain security clearance
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.