Staff Site Reliability Engineer, Payments Infrastructure
- Hiring from
- Brazil
- Work type
- Remote
- Posted
Show job descriptionHide job description
About hireworks
hireworks is building a community of top talent in key international markets by unlocking unparalleled access to positions at leading U.S. based companies. As your employer, hireworks will ensure you have a seamless interview, onboarding, and employee experience - providing ongoing support and resources along the way. Established in 2023, hireworks is forging corp-to-corp relationships with leading U.S. based organizations looking to grow their teams with best-in-class talent around the world. Working with hireworks means unlocking access to a network of local peers and mentors and career opportunities through our client network.
About Our client
A leading provider of innovative software and services to K-12 schools and educational institutions worldwide. Their mission is to empower schools to transform education through technology. They believe in the power of education to shape the future, and are committed to helping schools prepare students for success and make a positive impact on the world.
Position Overview
Our client's payments infrastructure spans two clouds and two application generations: a VM-based Azure estate that needs security hardening and modernization, and a modern AWS-based events platform. The SRE function is newly combined, bringing infrastructure ownership that used to be spread across teams under one roof, and is growing and maturing quickly. This is one of the two highest-priority hires in Payments: a Staff-level SRE who brings real technical leadership to a team still establishing its structure, owns reliability and security across both estates, and builds the paved road the rest of the organization needs to move fast safely.
Key Responsibilities
- Establish and operate monitoring, alerting, and reliability metrics across the Azure and AWS estates, working toward full critical-flow coverage
- Secure and remediate the Azure environment: network segmentation, fine-grained RBAC, privileged-access reviews, and end-of-life upgrades
- Operate and harden the AWS-based events platform (ECS, Lambda, Step Functions, EventBridge, Aurora, DynamoDB, CloudFront, WAF)
- Own or partner on CI/CD and release controls across both clouds
- Lead security posture work: cloud security posture management, vulnerability scanning, penetration testing, PCI-aware operations
- Help design and support a shared, limited-hours on-call model for the events platform
- Use AI-assisted development tools in a practical, controlled way, with human review as the final gate before anything ships
Required Qualifications
- 12+ years of overall engineering experience, including proven experience as a Staff SRE or equivalent senior IC operating production infrastructure at scale
- Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience
- Azure expertise required: genuine, hands-on operational depth, not just familiarity
- Experience securing a legacy, VM-based environment: RBAC, privileged access, network exposure remediation, and OS/database end-of-life upgrades
Experience with Azure DevOps (ADO) or similar CI/CD tooling
- Expert-level familiarity with Infrastructure-as-Code, including hands-on expertise with at least one IaC language (Terraform, Crossplane, OpenTofu, CDK, CloudFormation, or Bicep), applied as an enforced standard, not just personal practice
- Hands-on database administration and performance-tuning experience (SQL Server and/or a managed relational database such as Aurora MySQL), including indexing, query performance, and locking behavior at production scale
- Expert-level observability discipline: distributed tracing, actionable logging, dashboards, and alerting, including in-depth, hands-on experience with at least one event-based observability platform (Datadog, Honeycomb, New Relic, or comparable)
- Security/DevSecOps experience: posture management, vulnerability scanning, and PCI-aware production operations
- Deep proficiency as a software-oriented engineer, not just an infrastructure operator: comfortable reading and debugging application code (.NET, and ideally React) to separate infrastructure symptoms from application-level defects
- Experience modernizing legacy .NET systems and the infrastructure underneath them
- Proficiency with Claude Code, Codex, or comparable AI-assisted development tools
Preferred Qualifications
- Experience with modern, containerized AWS infrastructure (ECS, Lambda, Step Functions, EventBridge, or comparable)
- Experience with automated testing frameworks
- Experience helping a newly combined or fast-growing SRE or infrastructure team mature: standardizing titles, access, and tooling
- Experience migrating or consolidating CI/CD tooling between providers, with an eye toward reducing operational risk
- Background in payments or fintech infrastructure, particularly highly available environments
Benefits: hireworks is cultivating a growing community of top talent across LATAM. In addition to unlocking access to positions at top tier U.S. based companies, we offer a variety of benefits to enhance your experience:
- Competitive Pay - compensation that reflects your experience and accomplishments Remote Flexibility - work from anywhere within your home country (Brazil, Colombia or Argentina)
- Paid Time Off - ample vacation days to rest and recharge
- Public Holidays - local federal holidays are fully paid days off