Research Engineer
$250,000–$350,000 year
On-siteNew York City, New York, United States or New York, United States
Job Summary
Train and evaluate task-specific models for OCR, layout, and document extraction; optimize inference across diverse hardware setups; and integrate trained models into our API and product stack. Source, design, and clean datasets for supervised and synthetic training while maintaining reproducible pipelines. Run ablations, track metrics, and publish findings to inform model design and internal research direction. Occasionally engage with users and partners to understand customer needs and inform work priorities. Balance experimental rigor with shipping velocity to turn research insights into production-grade systems used by frontier AI labs and Fortune 500 enterprises.
Required Qualifications
- 3+ years experience training, fine-tuning, and evaluating deep learning models
- Trained at least one production-grade model or system used in real-world applications
- Deep expertise in PyTorch and Python
- strong fundamentals in deep learning (optimization, evaluation, architecture design)
- Comfortable with data engineering, benchmarking, and performance profiling across hardware setups
- Comfortable with an early stage startup
Desired Qualifications
- Have experience with OCR, document AI, or structured extraction
- Have published work, whether that's a paper, a benchmark report, or a deep technical blog post
- Have been a major contributor to open-source projects, especially in ML, vision, or NLP
- Enjoy writing about your work and sharing learnings with the community
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.