REQUIRED QUALIFICATIONS: 5+ years of experience in Site Reliability Engineering, Platform Engineering, or related roles supporting production systems Strong experience with Kubernetes-based platforms; OpenShift experience preferred Proven experience designing and operating distributed systems at scale Experience implementing monitoring, alerting, and incident response practices Strong scripting or programming skills (Python, Go, or similar) with a focus on automation Experience with CI/CD pipelines and GitOps workflows (GitHub, Jenkins, or similar) Experience with observability tooling (Prometheus, Grafana, ThousandEyes, or similar) Strong Linux and troubleshooting skills across complex systems Excellent communication skills with the ability to collaborate across technical and non-technical teams PREFERRED QUALIFICATIONS: Experience supporting internal developer platforms or platform engineering organizations Familiarity with shared services such as databases, messaging systems, and caching layers (MongoDB, PostgreSQL, Redis, RabbitMQ) Experience implementing SLOs and error budget frameworks Exposure to service mesh, traffic management, or advanced Kubernetes networking Experience mentoring engineers and influencing technical decision-making Preferred Location: MO Certain states and localities require employers to post a reasonable estimate of salary range. A reasonable estimate of the current base pay range for this position is $108,400 to $135,500 annually. Actual salary will be based on a variety of factors, including shift, location, experience, skill set, performance, licensure and certification, and business needs. The range for this position in other geographic locations may differ. Certain positions may also be eligible for variable incentive compensation, such as bonuses or commissions, that is not included in the base pay. The well-being of WWT employees is essential. So, when it comes to our benefits package, WWT has one of the best. We offer the following benefits to all full-time employees: Health and Wellbeing: Health (Medical & Prescription), Dental, and Vision Care, Onsite Health Centers (MO & IL), Employee Assistance Program, Wellness program Financial Benefits: Competitive Pay, Profit Sharing, 401k Plan with Company Matching, Life and Disability Insurance, Flexible Spending Accounts, Tuition Reimbursement Paid Time Off: PTO & Holidays, Parental Leave, Medical Leave, Military Leave, Bereavement, Day of Caring Additional Perks: Family Planning Benefits, Nursing Mothers Benefits, Voluntary Legal, Voluntary Supplemental Accident/Illness/Hospital, Voluntary ID Theft, Pet Insurance, Employee Discount Program Note: This is not an all-encompassing list and should not be used as a complete description of the plan’s benefits. For more information, see our US Benefits Website We strive to create an environment where all employees are empowered to succeed based on their skills, performance, and dedication. Our goal is to cultivate a culture of belonging that encourages innovation, collaboration, and respect for all team members, ensuring that WWT remains a great place to work for All! If you require accessibility accommodation(s) or adjustment during any stage of the hiring process, please let your WWT Recruiter know. The recruiter will work with you to understand your needs and help ensure an accessible experience throughout the interview process. If you have any questions or concerns about this posting, please email [email protected] . #LI-SSJ1 #LI-REMOTE PLEASE NOTE: This position requires permanent U.S. work authorization. Candidates requiring current or future visa sponsorship, including those on OPT, CPT or H1/H4, are not eligible for this role. This role is not open for staffing partners or corp‑to‑corp candidates. Why WWT? World Wide Technology (WWT) strives to make a new world happen. WWT's work benefits clients and partners as much as it does its people and community across the globe. Founded in 1990, WWT brings together strategy, deep technical expertise and world-class partnerships to help public and private sector organizations design, build and scale intelligent AI, digital, cybersecurity, cloud and infrastructure solutions. Through its Advanced Technology Center (ATC)—a collaborative ecosystem featuring state-of-the-art hardware and software—WWT enables clients and partners to conceptualize, test and validate innovative technology and then deploy solutions at scale using its global integration and distribution capabilities. With more than 14,000 team members and over 60 locations globally, WWT's culture—grounded in core values and leadership philosophies—has been recognized by Fortune® and Great Place to Work for its commitment to innovation, trust and creating a great place to work for all. WWT provides products and services to large enterprise, global service provider and public sector clients in up to 130 countries across six continents. Softchoice, a World Wide Technology company, supports commercial and SMB markets in the U.S. and Canada. Want to work with highly motivated individuals on high-performance teams? Join WWT today! Why Join the Internal IT Team? The Internal WWT IT team serves as the backbone of the company’s technology ecosystem. We design, build, and operate the platforms that power application development across the enterprise. Our work enables secure, scalable, and highly automated delivery of software and services that directly support WWT’s strategic objectives. Joining this team means working on foundational technology that matters—shared platforms, developer tooling, and automation that WWT depends on every day. You’ll be part of a collaborative, forward-looking engineering culture where ownership, innovation, and continuous improvement are expected and encouraged . What is the INFOPS Platform Engineering Team and why join? The INFOPS Application Development team focuses on enabling developers at scale. We build and evolve internal platforms, automation, and self-service capabilities that allow application teams to deploy and operate software safely and efficiently. Our work emphasizes: Platform Engineering best practices Kubernetes and OpenShift-based platforms GitOps-driven CI/CD automation Shared data and messaging platforms (MongoDB, PostgreSQL, RabbitMQ, etc.) AI-assisted and automation-driven engineering workflows As the organization matures, we are evolving toward a model that blends Platform Engineering with Site Reliability Engineering (SRE) to ensure our platforms are not only scalable—but also highly reliable and resilient. RESPONSIBILITIES: Improve platform availability, latency, and scalability across Kubernetes and supporting services, establishing and evolving reliability standards across shared platform capabilities, and measuring outage impact as a health signal to track progress in preventing, detecting, and recovering from failures Design and implement monitoring, alerting, and observability frameworks across platform services, standardizing telemetry (metrics, logs, traces) for consistent visibility across environments, and proactively identifying risks and performance bottlenecks before they impact users Lead triage and resolution of platform-related incidents, driving root cause analysis, long-term fixes, and blameless postmortem practices Collaborate with Support Engineers on the incident queue to surface recurring failure patterns, develop long-term remediation plans, reduce operational toil through automation, and drive sustainable resolution of systemic issues Design and build automation to reduce manual intervention, improve incident response and recovery time, and enable self-healing platform behaviors; develop reusable tools, runbooks, and operational patterns that scale across teams Collaborate with INFOPS Engineering and the Support Automation team to hand off support work where appropriate Partner with internal customers, application teams, and platform/DevOps engineering functions to deliver scalable, repeatable solutions, address shared reliability challenges, and promote a culture of operational excellence and continuous improvement
Lead Site Reliability & Security Engineer
Clera
Senior Site Reliability Engineer (SRE)
LeoLabs, Inc.
Principal Site Reliability Engineer
Navy Federal Credit Union
Site Reliability Engineer - Public Sector
Blitzy
Senior Site Reliability Engineer
Synapse Health
Site Reliability Engineer (SRE)
Thinking Machines Lab