Semarchy logo

Junior Site Reliability Engineer

Hiring from
United States
Work type
Remote
Posted
Sep 25, 2026
Is this job info correct?
Semarchy is looking for a Junior Site Reliability Engineer to help keep our SaaS products and infrastructure reliable and scalable. Reporting to the VP of IT Operations and Security, you'll work alongside senior engineers, software developers and support teams to keep our systems running well and available.

This is a great role for someone with a few years of hands-on cloud or DevOps experience who wants to grow their SRE skills in a production SaaS environment.

  • 2–3 years in an SRE, DevOps, Cloud Infrastructure or similar role.
  • Hands-on experience with AWS (e.g., EC2, EKS, RDS, S3).
  • Working knowledge of Kubernetes and containers (Docker, Helm a plus).
  • Some experience with an IaC tool such as Terraform, CloudFormation or Pulumi.
  • Familiarity with monitoring, alerting and logging tools (e.g., Datadog, Prometheus, Grafana, CloudWatch).
  • Scripting in Python, Go or Bash.
  • Good troubleshooting skills and eagerness to learn from incidents.

Nice to Have
  • Bachelor's degree in Computer Science, Information Technology or a related field (or equivalent experience).
  • Experience with CI/CD pipelines (e.g., GitHub Actions, GitLab CI, ArgoCD).
  • Experience with cloud security monitoring tools (e.g., AWS GuardDuty, Wiz).
  • Familiarity with Agile, DevOps and DevSecOps practices.
​​​​
1. Reliability and Availability
  • Help build and maintain systems that support high availability, fault tolerance and disaster recovery across our cloud environments.
  • Track SLOs and SLIs for critical services, and help improve them over time.
2. Infrastructure and Automation
  • Write and maintain automation scripts and Infrastructure as Code (e.g., Terraform) to cut down on manual operational work.
  • Set up and tune monitoring, alerting and logging so issues get caught early.
  • Follow security and change-management practices that support our SOC 2 and ISO 27001 commitments.
  • Help remediate vulnerabilities and keep infrastructure patched and hardened.
3. Incident Response and Post-Mortems
  • Take part in on-call rotations, with support from senior team members, to respond to incidents and outages.
  • Contribute to post-incident reviews and follow through on the actions that prevent repeat issues.
  • Write and keep up-to-date documentation and runbooks.
4. Performance and Cost
  • Monitor system performance, capacity and cloud spend, and flag areas to optimize.
  • Help carry out capacity upgrades and performance improvements.
We are committed to making Semarchy an even better place to work, and this means creating a high performing environment where everyone can do their Best Work, where they can Grow in their career path, and Be Accountable – to each other and Semarchy.

Semarchy is proud to be an equal opportunity employer (EEO) that celebrates difference and diversity. We are committed to equal employment opportunities regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We are committed to building an inclusive work environment where all employees feel a sense of belonging and respect. If there is anything we can do to ensure you have a comfortable and positive interview experience, please let us know.

Similar jobs

Apply for this job