Data Scientist
$75,000–$150,000 year
HybridVirginia, United States or Herndon, Virginia, United States
Job Summary
Build data pipelines, optimize machine learning models, and enable data-driven outputs for public sector and commercial clients across Intelligence, Defense, and AI/ML fields. Develop, deploy, and maintain ML models using Python, SQL, Apache Spark, and orchestration tools like Airflow or DBT to translate business requirements into tangible insights. Collaborate with engineering teams and non-technical stakeholders to present data-driven solutions while ensuring data quality, governance, and bias mitigation. This TS/SCI-cleared role supports end-to-end data science workflows for global, multi-national corporations and federal missions.
Required Qualifications
- TS/SCI eligibility with willingness/ability to obtain CI polygraph
- Bachelor's Degree
- at least 2 years experience in an analytical role as a data scientist, analyst or machine learning engineer
- ability to work with data using SQL and tools for statistical analysis (eg. Python, R)
- Solid knowledge of statistics and probability
- demonstrated experience applying these techniques to commercial use cases such as ML models and A/B tests
- demonstrated experience of developing, deploying and maintaining ML models and data pipelines within a commercial setting
- demonstrated experience of translating business and/or project requirements into insights, models and data products that generated tangible business value
- ability to work effectively in a team setting
- familiarity with commercial collaboration tools such as JIRA and Confluence
- strong communicator
- Curiosity and attention to detail with regards to data quality, distributions, biases and governance
- ability to work effectively within a fast paced and dynamic environment
- awareness and keen interest in the latest trends / techniques within data, machine learning and AI
- Fluent in Python/Scala/R as well as SQL/NOSQL
- demonstrable experience" with Fluent in Python/Scala/R as well as SQL/NOSQL, "Experience with Apache Spark
- other Big Data tools like Apache Flink/MapReduce
- Familiar with Big Data tools such as EMR/ Dataproc /Cloudera/Hortonworks
- Familiar with orchestration tools such as Airflow/DBT/NiFi
Desired Qualifications
- Experience with cloud solution providers such as GCP, AWS and Azure
- Prior experience building robustable and reusable code
- Knowledge of NLP and Image processing
- Prior experience with REST APIs
- An understanding of version control, modularity and maintaining code within a production setting
- An understanding of the web analytics and digital marketing ecosystem (eg. Adobe, Salesforce, Google)
- Prior experience working with DMPs and CDPs
- Prior consulting experience
- Experience with LLM's and GPT models a plus
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.