Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Akuity logo

Senior Site Reliability Engineer

Akuity
Posted 4 weeks ago
🇺🇸United States🏠Remote📁Engineering & Development
Is this job info correct?

About Akuity With the move to the cloud, Kubernetes has become widely adopted by DevOps and Platform Engineering teams, but it has also added complexity. While scaling Kubernetes at Intuit, the Akuity founders started building Argo CD in order to streamline the adoption of Kubernetes. Argo CD helps developers own, understand and deploy their K8s deployments via GitOps. Today, Argo CD is the third most popular project in the CNCF (Cloud Native Computing Foundation) and is used by 70% of companies who are using Kubernetes in production. The list of Argo CD users includes companies like Intuit, BlackRock, Tesla, Major League Baseball, Peloton, and many more. The team founded Akuity in 2021 to enable enterprises to ship software faster and more reliably with modern GitOps best practices. The Akuity Platform enables teams to manage the development and deployment across hundreds – if not thousands – of Kubernetes clusters from a single control plane. Trusted by top companies around the globe, the Akuity Platform provides the only end-to-end GitOps platform for the enterprises. Our mission is to simplify the software delivery process so that DevOps and Platform Engineering teams can move fast, and deploy code effortlessly without the fear of breaking things. The Role We are looking for a Senior SRE to help us keep the Akuity platform running at the level our enterprise customers expect. This is a high-ownership role; you won't just respond to incidents, you'll shape how we define and defend reliability across the entire platform. You'll work closely with engineering, infrastructure, and product to build the systems and culture that let us scale with confidence. What You'll Own Platform Reliability & SLAs Own SLI/SLO/SLA definitions for the Akuity SaaS platform and drive continuous improvement against them Design, instrument, and maintain observability systems (metrics, logs, traces) across multi-region AWS infrastructure Identify reliability gaps, lead blameless post-mortems, and close the loop with permanent fixes Partner with engineering teams to build reliability into new features before they ship to production On-Call & Incident Response Participate in an on-call rotation and act as incident commander for high-severity production events Build and maintain runbooks, escalation paths, and incident playbooks that keep mean time to resolution low Drive improvements to alerting fidelity; reduce noise, increase signal, eliminate toil Lead post-incident reviews with clear timelines, root cause analysis, and follow-through on action items What We're Looking For Required 5+ years of SRE, platform engineering, or production operations experience in a SaaS environment Deep hands-on Kubernetes expertise; you understand the scheduler, networking, storage, and autoscaling at a level where you can debug anything Strong AWS fundamentals across compute (EC2, EKS), networking (VPC, NLB, Route53), storage (S3, RDS), and IAM Experience defining and operating against SLOs in production; you've written error budgets, not just read about them Proficiency with observability tooling (Prometheus, Grafana, OpenTelemetry, Datadog, or equivalent) Solid scripting and automation skills; Go, Python, Bash, or similar; you automate what you touch Strong written communication: clear runbooks, sharp incident reports, thoughtful post-mortems Live within US time zones (Pacific through Eastern), including Canada and other regions Strong Advantage Experience with Argo CD, Kargo, or GitOps-based delivery workflows Familiarity with multi-region, multi-cluster Kubernetes deployments Experience with compliance-adjacent infrastructure (SOC 2, ISO 27001, HIPAA, or PCI DSS) Background operating infrastructure for other platform or developer tooling companies Our Stack Kubernetes (EKS): multi-region, enterprise-grade clusters serving Argo CD and Kargo workloads AWS: primary cloud provider across all production and DR environments Argo CD & Kargo: GitOps delivery tools we build and run ourselves Prometheus, Grafana, and OpenTelemetry for observability Terraform and GitOps-driven infrastructure management What We Offer Competitive compensation, commensurate with experience Equity participation in a well-funded, growing company Fully remote: work from anywhere within US time zones (Pacific through Eastern), including Canada and other regions Home office stipend and equipment budget Flexible time off and a culture that respects it Work directly with the engineers who built Argo CD and Kargo; you'll learn a lot here US-based employees receive full benefits, including comprehensive health, dental, and vision coverage. Candidates based outside the US will be engaged as contractors.

Similar jobs

Similar jobs

Empower logo

Senior Data Reliability Engineer AWS

Empower

🇺🇸United States17 hours ago
KO

Senior Reliability Engineer (Remote)

Kohls

🇺🇸United States17 hours ago
DD

Site Reliability Engineer

Ddcdine

🇺🇸United States17 hours ago
itD Tech logo

Site Reliability Engineer (6266)

itD Tech

🇺🇸United States17 hours ago
Salas O'Brien logo

Reliability Engineer

Salas O'Brien

🇺🇸United States17 hours ago
I8

Mid-Senior Site Reliability Engineer – Kubernetes Platform

I8Is

🇺🇸United States17 hours ago