Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Oowlish Technology logo

Senior Site Reliability Engineer (SRE)

Oowlish Technology
Posted Jul 2, 2026, 7:05 AM UTC
🌍Brazil, Colombia, Mexico🏠Remote📁Engineering & Development
Is this job info correct?

Join Our Team Oowlish, one of Latin America's rapidly expanding software development companies, is seeking experienced technology professionals to enhance our diverse and vibrant team. As a valued member of Oowlish, you will collaborate with premier clients from the United States and Europe, contributing to pioneering digital solutions. Our commitment to creating a nurturing work environment is recognized by our certification as a Great Place to Work, where you will have opportunities for professional development, growth, and a chance to make a significant international impact. We offer the convenience of remote work, allowing you to craft a work-life balance that suits your personal and professional needs. We're looking for candidates who are passionate about technology, proficient in English, and excited to engage in remote collaboration for a worldwide presence. About the Role: We are looking for an experienced Senior Site Reliability Engineer (SRE) to own the reliability, availability, and operational excellence of business-critical production systems. This is a dedicated Site Reliability Engineering role—not a general DevOps or Infrastructure position. You will define how reliability is measured, lead incident response during production outages, drive observability strategy, and continuously improve operational practices across high-availability environments. The ideal candidate has hands-on experience managing SLOs, leading major incidents, improving on-call operations, and building a strong reliability culture through automation, observability, and continuous improvement. Responsibilities: Define, implement, and continuously improve Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets. Develop and maintain observability strategies, including monitoring, logging, tracing, and alerting. Own observability configuration, instrumentation, and alert optimization. Lead Incident Command during production incidents and coordinate cross-functional response efforts. Drive blameless postmortems and ensure corrective actions are completed. Own and continuously improve the on-call program, including rotations, escalation policies, runbooks, and alert tuning. Establish production readiness standards for new services. Partner with engineering teams on capacity planning, scalability, and disaster recovery initiatives. Automate operational processes and reliability improvements using software engineering best practices. Continuously improve system reliability, availability, and operational efficiency. Requirements: 5+ years of experience in Site Reliability Engineering, Production Engineering, Reliability Engineering, or similar roles. Proven experience operating production systems in high-availability environments. Hands-on experience defining and managing SLOs, SLIs, and Error Budgets. Experience leading production incident response and Incident Command. Strong observability and monitoring experience. Strong software engineering skills using Python, Go, or TypeScript. Experience working with cloud platforms. Strong written and verbal English communication skills. Must have: Proven Site Reliability Engineering experience. Experience defining and managing: Service Level Indicators (SLIs) Service Level Objectives (SLOs) Error Budgets Experience leading Incident Command during major production incidents. Experience conducting blameless postmortems and driving follow-up actions. Experience designing, maintaining, and improving on-call programs. Experience developing runbooks and escalation policies. Strong observability experience, including: Monitoring Logging Alerting Distributed Tracing Experience tuning alerts to reduce operational noise. Strong automation skills using Python, Go, or TypeScript. Experience supporting mission-critical production systems. Experience working in high-availability production environments. Nice to have: Experience with Datadog. Experience with AWS. Experience with Heroku. Experience working in regulated industries (Healthcare, HIPAA, Financial Services, etc.). Experience establishing or maturing an SRE practice. Capacity planning experience. Disaster recovery planning and execution. Experience with Kubernetes. Experience with PostgreSQL or SQL Server. Experience supporting modern TypeScript-based applications. Benefits & Perks: Home office; Competitive compensation based on experience; Career plans to allow for extensive growth in the company; International Projects; Oowlish English Program (Technical and Conversational); Oowlish Fitness with Total Pass; Games and Competitions; You can also apply here: Website: https://www.oowlish.com/work-with-us/ LinkedIn: https://www.linkedin.com/company/oowlish/jobs/ Instagram: https://www.instagram.com/oowlishtechnology/

Similar jobs

Similar jobs

Mastercard logo

Site Reliability Engineer II

Mastercard

🌍India, Mexico, United States18 hours ago
Cognativ logo

Senior Site Reliability Engineer (Hiring Globally)

Cognativ

🌍Argentina, Brazil, India, Mexico, Serbia, Spain, United KingdomYesterday
CU

Site Reliability Engineer (SRE) | AWS | Kubernetes | Databricks | Sênior (Remote)

Compass UOL

🇧🇷BrazilYesterday
BairesDev logo

Robot Reliability Engineer - Remote Work

BairesDev

🇨🇴Colombia2 days ago
Encora logo

Site Reliability Engineer (SRE) - DevOps

Encora

🇧🇷Brazil4 days ago
Lumenalta logo

Senior DevSecOps / Site Reliability Engineer (GCP)

Lumenalta

🌍Canada, Colombia, Dominican Republic, Mexico5 days ago