REMOTE (INDIA): AI Engineer- SaaS Platform
RemoteIndia
Job Summary
Maintain and optimize LLM- and VLM-powered services for content generation, compliance scoring, and campaign testing while managing Flask/FastAPI microservices to ensure high uptime and low latency. Deploy, monitor, and debug Uvicorn/Gunicorn-based hosting in production environments, integrating with OpenRouter tools to balance cost, latency, and quality. Build feedback pipelines for AI model evaluation and expose secure, versioned REST APIs for AI services. Collaborate with backend/frontend teams to keep microservice architecture aligned and track token consumption, latency, and error rates. Requires 3–5 years of Python experience with production-grade codebases, microservice architecture, and CI/CD pipelines.
Required Qualifications
- Strong in Python, with experience in production-grade codebases
- Flask (for APIs)
- Dramatiq (or Celery/RQ equivalent) for background jobs
- Hands-on with LLMs and VLMs, including prompt engineering, fine-tuning, and evaluation
- Familiar with OpenRouter or equivalent LLM/VLM routing & fallback tools
- Experience designing and maintaining microservice architectures
- Strong experience with REST API design (auth, rate limiting, documentation)
- Dockerized deployments, CI/CD pipelines, logging/monitoring, error handling
- Building structured evaluation/feedback systems for AI model performance
- 3–5 years as an AI Engineer or Python Backend Engineer working with production systems
- Demonstrated ability to maintain AI pipelines in production, not just prototypes
Desired Qualifications
- FastAPI (optional)
- Uvicorn/Gunicorn for async hosting
- AWS/GCP experience preferred (deployment, monitoring, scaling)
- Prior work with SaaS platforms, LLM/VLM integrations, or AI-first products is highly valued
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.