Python Pyspark
$2,000,000–$2,000,000 year
RemoteUnited States
United StatesRemoteContract$2,000,000–$2,000,000 yearSmall
ContractSmall
Job Summary
Design and develop Hadoop applications using PySpark with Python or Scala, leveraging core Java, MapReduce, and Hive programming expertise. Manage source code through Git repositories and build scripts with Maven or Cradle while integrating Jenkins for CI/CD. Gain exposure to the AWS ecosystem including EC2 and S3, and apply SQL programming and agile methodology to deliver software solutions. This remote role prioritizes hands-on development of jobs and performance optimization within the Hadoop and Spark environments.
Required Qualifications
- Design and develop on Hadoop applications
- Hands-on in developing Jobs in pySpark with Python
- Experience on Core Java
- Experience on Map Reduce programs
- Hive programming
- Hive queries performance concepts
- Experience on source code management with Git repositories
Desired Qualifications
- Hands-on in developing Jobs in pySpark with SCALA
- Java/ SCALA
- Exposure to AWS Ecosystem with hands-on knowledge of ec2, S3 and services
- Basic SQL programming
- Knowledge of agile methodology for delivering software solutions
- Build scripting with Maven / Cradle
- Exposure to Jenkins
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.