Senior DevOps/Cloud Infrastructure Engineer
- Hiring from
- Argentina
- Work type
- Remote
- Posted
- Sep 24, 2026
Is this job info correct?
Svitla Systems Inc. is looking for a Senior DevOps/Cloud Infrastructure Engineer for a full-time position (40 hours per week) on a short-term project in Argentina. Our client is a software company focused on an AI-driven sales automation platform for lead qualification and engagement.
You will help improve, automate, and secure our cloud infrastructure and deployment workflows across AWS and GCP. This is a hands-on role for someone who can quickly understand the current environment, identify gaps, and drive improvements with minimal direction. The primary focus will be to streamline deployments, migrate manual and script-based infrastructure workflows to Terraform, improve CI/CD reliability, tighten production access controls, and support SOC 2 compliance readiness.
The ideal candidate has strong production experience with Terraform, GitHub Actions, AWS, GCP, serverless infrastructure, networking, IAM, observability, and compliance-oriented DevOps practices.
Current Cloud And Deployment Setup Includes
You will help improve, automate, and secure our cloud infrastructure and deployment workflows across AWS and GCP. This is a hands-on role for someone who can quickly understand the current environment, identify gaps, and drive improvements with minimal direction. The primary focus will be to streamline deployments, migrate manual and script-based infrastructure workflows to Terraform, improve CI/CD reliability, tighten production access controls, and support SOC 2 compliance readiness.
The ideal candidate has strong production experience with Terraform, GitHub Actions, AWS, GCP, serverless infrastructure, networking, IAM, observability, and compliance-oriented DevOps practices.
Current Cloud And Deployment Setup Includes
- AWS and GCP are the primary cloud platforms.
- AWS Amplify deployments are currently managed and triggered directly through AWS Amplify.
- AWS Lambda functions are currently managed using the Serverless Framework.
- Vertex AI Pipelines are deployed across multiple tenants, and their deployment times need to be optimized.
- Tenant provisioning is currently handled partially through scripts.
- GitHub is the primary source control and automation platform.
- SOC 2 controls require ongoing maintenance and improvement.
- 7+ years of experience in DevOps, cloud infrastructure, platform engineering, or systems engineering.
- Strong experience with Terraform and infrastructure as code (IaC) in production environments.
- Strong knowledge of and experience with SecOps.
- Expertise in managing infrastructure across AWS and GCP.
- In-depth understanding of building and maintaining GitHub Actions workflows.
- Strong knowledge of GitHub Actions and CI/CD automation.
- Understanding of multi-tenant provisioning automation.
- Knowledge of least-privilege access and just-in-time (JIT) access.
- Experience with SOC 2 technical controls.
- Experience with production incident response and troubleshooting.
- Experience with AWS services, including Lambda, IAM, VPC, Amplify, CloudWatch, Secrets Manager, and Parameter Store.
- Experience with GCP services and capabilities, including Vertex AI, IAM, VPC networking, cloud security, and Cloud Logging/Monitoring/observability.
- Experience migrating manual or framework-based infrastructure to Infrastructure as Code.
- Strong understanding of Linux, networking, DNS, IAM, and cloud security fundamentals.
- Experience implementing least-privilege access and production access controls.
- Experience with monitoring, logging, and observability tools.
- Solid scripting experience with Bash, Python, Go, or similar languages.
- Ability to work independently, clarify ambiguity, and drive implementation with minimal guidance.
- Strong documentation and communication skills.
- Self-directed professional who can identify what needs to be done and take action without waiting for detailed instructions.
- Ownership-oriented professional who treats infrastructure reliability, security, and deployment quality as their responsibility.
- Ability to balance speed, quality, and maintainability without over-engineering.
- Ability to design access, automation, and deployment workflows with least-privilege principles and auditability in mind.
- Ability to work well with engineering, product, and leadership teams.
- Documentation-focused professional who can leave behind clear runbooks, diagrams, and processes that the team can maintain.
- Ability to focus on measurable improvements in deployment speed, reliability, security, and compliance.
- Move AWS Amplify deployments from direct Amplify workflows to GitHub-triggered CI/CD.
- Migrate AWS Lambda/serverless infrastructure to Terraform.
- Migrate tenant provisioning scripts to Terraform modules and GitHub Actions workflows.
- Optimize Vertex AI Pipeline deployment times across multiple tenants.
- Clean up and standardize AWS and GCP VPC configurations.
- Implement GitHub repository rules for branch naming, Jira references, PR approvals, and protected branches.
- Automate production permissions and implement just-in-time access for on-call engineers.
- Improve SOC 2 controls around access, deployments, infrastructure changes, and evidence collection.
- Improve monitoring, logging, and alerting, and create or enhance operational runbooks.
- Document infrastructure architecture and deployment workflows.
- Experience with multi-tenant SaaS infrastructure.
- Familiarity with Vertex AI Pipelines or ML/AI deployment workflows.
- Understanding of optimizing long-running cloud deployment pipelines.
- Familiarity with GitOps or declarative infrastructure patterns.
- Experience with SOC 2, ISO 27001, or similar compliance frameworks.
- Experience with just-in-time access tools or capabilities such as Okta, Teleport, AWS IAM Identity Center, Google IAM Conditions, or similar.
- Knowledge of policy-as-code tools such as OPA, Checkov, Conftest, Sentinel, or Terraform Cloud policies.
- Knowledge of Kubernetes, Docker, or containerized workloads.
- Manage and improve cloud infrastructure across AWS and GCP.
- Convert existing manual provisioning scripts into Terraform.
- Migrate AWS Lambda/serverless infrastructure from Serverless Framework to Terraform.
- Build reusable Terraform modules for tenant provisioning, networking, IAM, deployment resources, and environment setup.
- Ensure infrastructure is repeatable, version-controlled, documented, and auditable.
- Improve infrastructure consistency across development, staging, and production environments.
- Streamline deployment workflows across services and environments.
- Move AWS Amplify deployments from direct Amplify workflows to GitHub-driven workflows.
- Build and maintain GitHub Actions pipelines for application deployment, infrastructure deployment, and tenant provisioning.
- Improve deployment speed, reliability, rollback capabilities, and visibility.
- Optimize deployment workflows for multi-tenant environments.
- Optimize long-running Vertex AI Pipeline deployments across multiple tenants.
- Establish clear promotion workflows between environments.
- Review existing tenant provisioning scripts and workflows.
- Convert tenant provisioning into Terraform-backed infrastructure workflows.
- Automate tenant provisioning using GitHub Actions.
- Improve repeatability, traceability, and rollback capabilities for tenant setup.
- Reduce manual operational work and deployment risk.
- Review, clean up, and improve existing VPC structures.
- Define clear networking patterns across AWS and GCP.
- Improve segmentation between environments and tenants where appropriate.
- Review DNS, routing, security groups, firewall rules, and cloud networking configurations.
- Document cloud network architecture and create or improve operational runbooks.
- Implement least-privilege access across AWS, GCP, GitHub, and deployment systems.
- Automate permission management for engineering and production environments.
- Restrict production access based on role and operational need.
- Implement or improve just-in-time access for on-call engineers.
- Improve auditability of privileged access and production changes.
- Review secrets management and recommend improvements where needed.
- Implement GitHub repository rules and engineering workflow standards, including: Branch naming conventions; Pull request requirements; Required Jira ticket references; Protected branches; Required reviews; Required CI checks; Environment-based approvals.
- Improve consistency of engineering workflows across repositories.
- Ensure GitHub workflows support both developer velocity and compliance needs.
- Review and improve monitoring, logging, metrics, and alerting.
- Help identify deployment bottlenecks, infrastructure risks, and recurring operational issues.
- Improve incident response readiness through runbooks and documentation.
- Support production incident troubleshooting when needed.
- Recommend improvements to reduce operational toil and improve system reliability.
- Help maintain and improve technical controls required for SOC 2.
- Support controls related to: Access management; Change management; Deployment approvals; Infrastructure security; Production access; Audit logging; Evidence collection.
- Ensure infrastructure and deployment processes are auditable and documented.
- Help create or improve runbooks, diagrams, and process documentation needed for compliance
- US and EU projects based on advanced technologies.
- Competitive compensation based on skills and experience.
- Comprehensive private medical insurance.
- Regular performance appraisals to support your growth.
- Flexibility in workspace, either remote, our welcoming office or local coworking.
- Bonuses for recommendations of new employees.
- Bonuses for article writing, public talks, other activities.
- 15 vacation days, 10 national holidays, 10 sick leaves.
- Personalized learning program tailored to your interests and skill development.
- Free tech webinars and meetups organized by Svitla.
- Fun corporate online\offline celebrations and activities.
- Well-established remote culture.
- Awesome team, friendly and supportive community!