MLOps Engineer (AI Infrastructure)

Mumbai | Remote (India)Full-timeMidAI/Data
Run the infrastructure AI systems depend on: serving, monitoring, cost and model updates.

About the Role

Own the platform our AI runs on. You will make deployment boring, cost visible, and model updates something we test rather than discover.
What You'll Do
What You'll Do
Check Icon
Build and operate model serving infrastructure
Check Icon
Implement monitoring for quality drift, latency and spend
Check Icon
Build CI pipelines that run evaluations before a model change ships
Check Icon
Manage model registries and rollback paths
Check Icon
Optimise inference cost through routing, caching and batching
Check Icon
Support private and self-hosted deployments for regulated clients
What You'll Bring
Check Icon
2+ years in DevOps, platform or ML infrastructure
Check Icon
Strong Docker and cloud experience on AWS or GCP
Check Icon
Python and comfort with CI/CD tooling
Check Icon
Understanding of GPU cost and capacity planning
Check Icon
Monitoring and alerting instincts
What You'll Bring
Nice to Have
Nice to Have
Check Icon
vLLM, Ray or similar serving stacks
Check Icon
Kubernetes in production
Check Icon
Experience self-hosting open-weight models
What You'd Build
These aren't hypothetical projects, they're live products you can try before your first interview.
Check Icon
The StackBinary MarTech Suite, The full product line, every system we ship, most with live demos.
Why Join StackBinary™?
Check Icon
Flexible working hours
Check Icon
Remote-friendly culture
Check Icon
Learning & development budget
Check Icon
High-ownership projects
Check Icon
Pragmatic engineering culture
Check Icon
Work with cutting-edge tech
Why Join StackBinary

Ready to Apply?

Join our team of builders who love shipping quality software.
We connect with shortlisted candidates through LinkedIn or our official email IDs.
Follow StackBinary on LinkedIn →
Questions about this role?
contact@stackbinary.io