Hatch Innovations Canada logo

Senior DevOps Engineer - Electronic Arts [EADV261006]

Salary
CA$115K–CA$135K
Hiring from
Canada
Work type
Remote
Posted
Is this job info correct?
Show job description

About Us

Hatch Studio has been building and running software since 2011. Our engineers have worked on products used by millions of people, alongside teams at companies like EA, Epic Games and Krafton. We've lasted 15 years by delivering great work and treating people well.

We believe AI is changing how great software gets built, and we want our people at the front of that change. At Hatch, you'll be encouraged to explore AI tools in your daily work, share what you learn, and help shape how we build. We want engineers who are curious, care about quality, and take ownership of what they ship.

We're based in Vancouver, with team members across Canada.

About The Role

We are looking for an experienced Senior DevOps Engineer to lead the foundation, reliability, and deployment strategy for high-concurrency cloud environments supporting an innovative game by Electronic Arts (EA).

In this role, you will work hand-in-hand with our Go backend engineering team to ensure high-throughput microservices run seamlessly across cloud providers. You will design resilient Infrastructure as Code (IaC), build automated CI/CD deployment pipelines, optimize cloud spend, and establish deep observability across backend microservices and data streams.

You will be working in an engineering-oriented, fast-paced environment with minimal management and detailed task definition. You need to be a self-starter who excels at making your own decisions, driving reliability best practices, and organizing your work to support our product deployment goals.

This is a remote position with working hours that must align with Pacific Time (PT) business hours.

You will

  • Infrastructure as Code & Cloud Architecture: Design, provision, and maintain production-grade, highly available cloud infrastructure on AWS (and GCP) using Terraform and modern cloud-native patterns.
  • Developer Platform & CI/CD: Architect and maintain fast, secure, and resilient automated CI/CD pipelines to support continuous deployment of containerized Go microservices and event processing engines.
  • Scalability & Reliability Engineering: Collaborate with backend engineers to support high-concurrency systems (high RPS, streaming data, and low-latency client-server flows). Define reliability metrics (SLOs/SLAs) and build autoscaling strategies.
  • Observability & Monitoring: Implement robust monitoring, tracing, and logging stacks using tools like Prometheus, Grafana, OpenTelemetry, and Datadog to ensure system visibility and proactive incident detection.
  • On-Call & Incident Management: Establish, refine, and participate in blameless on-call rotations alongside backend engineers for infrastructure and platform services, driving post-mortem reviews and automated incident remediation.
  • Database & Messaging Support: Support zero-downtime PostgreSQL migrations, backup/disaster recovery strategies, and streaming infrastructure (AWS Kinesis, SNS/SQS, Kafka, NATS).
  • Technical Communication: Write clear Architecture Decision Records (ADRs), infrastructure RFCs, post-mortems, and deliver actionable infrastructure code reviews.

You have

  • 8+ years of professional experience operating production infrastructure, platform engineering, or DevOps at scale.
  • Production Cloud & Kubernetes Experience: Deep expertise provisioning and running containerized microservices on AWS (using EKS, ECS, or similar container orchestration systems).
  • Strong Infrastructure as Code (IaC): Advanced hands-on proficiency with Terraform for complex, multi-environment cloud infrastructure.
  • CI/CD & Automation: Experience building automated build and release pipelines (GitHub Actions, GitLab CI, ArgoCD, or similar).
  • Observability Mastery: Demonstrated experience setting up and managing end-to-end telemetry (OpenTelemetry, Prometheus, Grafana, Datadog) for high-volume microservices.
  • Backend Systems Familiarity: Solid understanding of backend execution environments, containerization (Docker), networking protocols (gRPC, HTTP/REST), and relational database operations (PostgreSQL backup/restore, replication, zero-downtime migrations).
  • Modern Tooling: Daily use of AI coding and automation tools (Claude Code, Cursor, Copilot, etc.) with a highly critical lens—you carefully review, test, and reject substandard generated code or configurations rather than blindly applying them.
  • Canadian Work Eligibility: Must be a resident of Canada and eligible to work in Canada.

Nice-to-Haves

  • Multi-Cloud Expertise: Hands-on experience managing Google Cloud Platform (GCP) infrastructure alongside AWS.
  • Event & Messaging Infrastructure: Experience running or operating event-driven platforms (AWS Kinesis, SNS/SQS, Apache Kafka, or NATS).
  • Systems & Scripting Languages: Proficiency in Go, Python, or Shell scripting for writing custom automation scripts or Kubernetes operators.
  • Gaming / Low-Latency Domain: Prior experience supporting live-service online games, studio environments, agency/consultancy client projects, or high-throughput real-time APIs.

Time Zone Requirements:

  • Ability to attend regular syncs with EU team members before 9 am Eastern Time Zone.

How to apply:

  • To apply, please send your PDF resume and Github profile.
  • Note: A background check will be required for employment in this role.

Job Types:

  • Permanent, Full-time
  • Schedule: Monday to Friday

Pay:

$115K to $135K CAD per year

We offer:

  • Health Spending Account
  • Disability insurance
  • Life insurance
  • Paid time off
  • Work from home

Accommodation

Hatch Studio is committed to an inclusive and accessible hiring process. If you need an accommodation at any stage of recruitment, tell us when you apply or at any point after, and we'll work with you to meet your needs.

Similar jobs

Apply for this job