About SailPoint SailPoint is the leader in identity security for the cloud enterprise. Our identity security solutions secure and enable thousands of companies worldwide, giving customers unmatched visibility into their digital workforce and ensuring workers have the right access — no more, no less. Built on a foundation of AI and machine learning, our Identity Security Cloud (Atlas) platform delivers the right level of access to the right identities and resources at the right time — matching the scale, velocity, and changing needs of today’s cloud-oriented enterprise. SailPoint is all-in on the AI revolution, and our engineers have access to the latest frontier models and agentic frameworks. About the Role As a Senior Staff DevOps Engineer on the Infrastructure Platform team, you will be a technical leader responsible for designing, building, and operating SailPoint’s global Identity Security Cloud infrastructure on AWS, Azure, and GCP. You will partner with engineering teams across India, the US, EMEA, and APAC to deliver a resilient, secure platform. You will be a key contributor to our AI-assisted and agentic automation and tooling. You will serve as a Kubernetes platform leader for the team — guiding engineers on EKS and AKS operations, service mesh, and cloud-native deployment patterns across our production environments. You will also play a key role in supporting SailPoint’s PCI compliance initiative, helping ensure our platform infrastructure meets PCI DSS requirements. This is a fully remote position for candidates based in India. The role carries significant technical influence across the organization — setting architecture direction, mentoring engineers, and driving operational excellence without requiring people management. The ideal candidate is a self-starter who thrives in complex, fast-paced SaaS environments, enjoys solving hard infrastructure problems, brings expert-level Kubernetes, AWS platform, and modern AI knowledge, values quality and reliability, and is excited to thrive on an AI enabled team. About the Team The Infrastructure Platform Team designs and operates the foundational cloud infrastructure that powers SailPoint’s Identity Security Cloud. Our platform runs on a mature, cloud-native, event-driven microservices architecture on AWS, with Kubernetes, GitOps, and Infrastructure-as-Code at the core. We also support smaller environments in Azure and GCP. Engineers in India work closely with global peers in the US, the UK, Israel, the UAE, and Mexico. The team owns all SailPoint’s cloud infrastructure, compute capabilities, and operational practices for a 24×7 SaaS environment serving large enterprises worldwide. We focus on GitOps and Infrastructure as Code, agentic automation, and managing web platforms that serves millions of concurrent requests. Key Responsibilities Multi-Account AWS Operations Design, build, and maintain a resilient, secure, and cost-efficient multi-account AWS SaaS platform that meets established SLAs and uptime targets. Design and evolve multi-account AWS architectures, guardrails, and IaC patterns. Kubernetes Platform Leadership Serve as a technical lead on Kubernetes for the Infrastructure Platform team — our platform is heavily Kubernetes-based, and this role requires deep, hands-on expertise. Design, operate, and optimize production Kubernetes clusters at scale on AWS (EKS) and Azure (AKS), including cluster architecture, upgrades, node management, networking, storage, and multi-tenant isolation patterns. Design, implement, and operate a service mesh (e.g., Istio, Linkerd, or AWS App Mesh) for microservices communication — including mTLS, traffic management, observability, resilience patterns, and progressive delivery across environments. Define and drive Kubernetes and service mesh standards and best practices across teams — workload design, resource management, security hardening (RBAC, Pod Security Standards, network policies), Helm/chart conventions, and deployment patterns. Guide and mentor engineers on Kubernetes and service mesh troubleshooting, performance tuning, capacity planning, and production incident resolution for containerized microservices. Platform Architecture Design and scale infrastructure to meet rapidly increasing customer demand, data sovereignty requirements, and regional expansion. Define standards for AWS and Azure networking, IAM, secrets management, AWS Organizations governance, and cross-account connectivity. Automate deployment, monitoring, incident response, and capacity management using GitOps and CI/CD best practices. Develop and improve operational practices, runbooks, and platform engineering standards. Collaborate with development teams to bring new features and services into production safely and efficiently. Proactively meet information security and compliance standards (e.g.,PCI DSS), including supporting SailPoint’s PCI compliance initiative through secure platform design, controls implementation, and audit readiness. Participate in and help improve the on-call rotation; drive post-incident reviews and systemic fixes. AI & Agentic Workflow Innovation Champion AI-assisted and agentic workflows across DevOps, SRE, and platform engineering — integrating LLM-based tooling, AI code assistants, and autonomous agents into day-to-day operations. Design and operationalize agentic automation on AWS using services such as Amazon Bedrock, including model selection, prompt/orchestration patterns, and integration with operational tooling for tasks such as infrastructure code generation/validation, incident triage, root-cause analysis, deployment troubleshooting, compliance checks, and repetitive operational runbooks. Establish guardrails, observability, cost controls, and human-in-the-loop practices for safe adoption of Bedrock-powered and AI-driven automation in production environments. Partner with engineering leadership to define a roadmap for AI-native platform operations on AWS and measure productivity, reliability, and cost impact. Technical Leadership & Cross-Functional Influence Develop and document cutting-edge techniques, patterns, and best practices; mentor Staff, Senior, and mid-level engineers. Manage cross-functional requirements with Engineering, Product, Services, and Security teams across time zones. Influence engineers and teams who are not direct reports — setting process, raising the bar, and driving consensus on technical direction. Serve as a subject-matter expert during critical escalations and architecture reviews. Background & Experience Required Strong interpersonal and teaming skills — ability to set and enforce process and influence engineers across teams and geographies. Ability to operate effectively in an agile, entrepreneurial environment with global stakeholders. Prior experience as a technical lead or Staff+ IC in a global engineering organization. 12+ years of experience in 24×7 production operations, supporting highly available SaaS or cloud service environments. 12+ years of experience with containerization, virtualization, and configuration management technologies. 5+ years of experience with multi-account AWS cloud infrastructure. 5+ years of hands-on experience with Kubernetes in production at scale. 5+ years of experience with Terraform (IaC), managing infrastructure across multiple AWS accounts and regions. 5+ years of experience designing and implementing CI/CD pipelines, especially for Terraform, Kubernetes, and microservices. 5+ years of experience with scripting/programming languages (Python, Go, or similar) and strong shell scripting proficiency. Strong understanding of Linux, networking, distributed systems, and production troubleshooting. Hands-on experience implementing and operating a service mesh in production Kubernetes environments (Istio, Linkerd, or AWS App Mesh). Experience designing and operating multi-account AWS environments with AWS Account Factory, and organizational guardrails (SCPs, preventive/detective controls, landing-zone standards). Experience applying AI tools on AWS (e.g., Amazon Bedrock or equivalent) to improve engineering productivity, automation, or operational efficiency. Experience with monitoring and logging stacks (e.g., Prometheus, Grafana, OpenSearch or equivalent). Experience with GitOps tooling (e.g., ArgoCD, Kargo, Flux). Strong understanding of SRE principles — SLIs/SLOs, error budgets, incident management, and observability. Preferred Hands-on experience implementing Account Factory for Terraform (AFT) or similar automated account vending and baseline provisioning at scale. Experience designing or operating agentic workflows on AWS using Amazon Bedrock, LLM agents, AI-driven runbook automation, or intelligent incident response systems in enterprise production contexts. Experience supporting regional/sovereign AWS deployments (e.g., UAE, EU, India data residency requirements). Familiarity with compliance frameworks in regulated enterprise SaaS, including PCI DSS and FedRAMP-adjacent practices — with experience implementing or operating platforms subject to PCI compliance requirements preferred. What Success Looks Like First 30 Days Onboard into the role; learn Identity Security Cloud architecture, platform tooling, and team processes. Build relationships with peers and stakeholders across India and global DevOps/SRE teams. Join team ceremonies, contribute to in-flight projects, and begin participating in on-call shadowing. First 90 Days Own significant platform initiatives; contribute to architecture decisions and operational improvements. Contribute to or improve multi-account guardrails and Account Factory patterns within the AWS organization. Integrate AI-assisted tooling into at least one team workflow (e.g., IaC review, incident triage, deployment validation). Join the on-call rotation; become proficient in core services, escalation paths, and runbooks. Mentor engineers and document at least one new platform pattern or operational standard. First 6 Months Lead cross-functional projects to improve platform reliability, deployment velocity, or multi-account AWS consistency and standardization. Deliver measurable improvements in automation, incident response time, or operational toil reduction — including through agentic workflow pilots. Become a recognized SME on platform services; independently handle complex production escalations. Influence roadmap and engineering practices across the broader Infrastructure organization. First 12 Months Drive a strategic initiative that advances SailPoint’s multi-account, AI-enabled platform operations vision. Establish durable standards, guardrails, and documentation adopted by multiple teams. Lead or significantly advance service mesh adoption and standardization — including security controls, observability integration, and documented patterns adopted by engineering teams. Demonstrate sustained impact on uptime, deployment success rate, cost efficiency, or engineering productivity. Work Model This is a fully remote position for candidates based in India. Candidates must be eligible to work in India and available for collaboration with global teams across US, EMEA, and APAC time zones. Participation in an on-call rotation is required. Education Bachelor’s and/or Master’s degree in Computer Science or equivalent technical experience. SailPoint is an equal opportunity employer and we welcome all qualified candidates to apply to join our team. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, protected veteran status, or any other category protected by applicable law. Alternative methods of applying for employment are available to individuals unable to submit an application through this site because of a disability. Contact [email protected] or mail to 11120 Four Points Dr, Suite 100, Austin, TX 78726, to discuss reasonable accommodations. NOTE: Any unsolicited resumes sent by candidates or agencies to this email will not be considered for current openings at SailPoint.
Linux DevOps Engineer
Kyndryl
DevOps Engineer
Xtglobal
DevOps Engineer
Weekday AI
Assistant Manager | DevOps | Bengaluru | Engineering as a Service/ Operate
Southasiacareers
Senior Associate | DevOps | Bengaluru | Engineering as a Service/ Operate
Southasiacareers
GCP DevOps Engineer
EXL Talent Acquisition Team