We're hiring a Senior MLOps Engineer to own the reliability and scale of our GPU compute platform. Vecura runs 300+ scientific AI tools — protein structure prediction, molecular dynamics, docking, and more — across a wide range of GPU types, on both serverless and self host spanning cloud and on-prem. You'll own the platform layer between infrastructure and models: how GPU jobs are scheduled, queued, isolated, observed, and recovered. You think in SLOs and design for repeatability, building standard systems that scale across hundreds of models rather than one-off deployments. You'll work alongside our DevOps engineer (infra/cluster) and our AI engineers (model onboarding), owning the orchestration and reliability surface that connects them.
Must have:
We provide a dynamic, fast-paced, and collaborative environment where problem-solving and agility are at the heart of what we do. Along with a competitive salary, we foster a culture that values ambition, confidence, and humility, consistently pushing the boundaries of innovation. If you're excited about working in a young, talented tech company and want to explore the world of AI and pharmaceuticals, we encourage you to apply.