Svitla Systems, Inc. logo

Senior DevOps/Cloud Infrastructure Engineer

Hiring from
Argentina
Work type
Remote
Posted
Sep 24, 2026
Is this job info correct?
Svitla Systems Inc. is looking for a Senior DevOps/Cloud Infrastructure Engineer for a full-time position (40 hours per week) on a short-term project in Argentina. Our client is a software company focused on an AI-driven sales automation platform for lead qualification and engagement.

You will help improve, automate, and secure our cloud infrastructure and deployment workflows across AWS and GCP. This is a hands-on role for someone who can quickly understand the current environment, identify gaps, and drive improvements with minimal direction. The primary focus will be to streamline deployments, migrate manual and script-based infrastructure workflows to Terraform, improve CI/CD reliability, tighten production access controls, and support SOC 2 compliance readiness.

The ideal candidate has strong production experience with Terraform, GitHub Actions, AWS, GCP, serverless infrastructure, networking, IAM, observability, and compliance-oriented DevOps practices.

Current Cloud And Deployment Setup Includes

  • AWS and GCP are the primary cloud platforms.
  • AWS Amplify deployments are currently managed and triggered directly through AWS Amplify.
  • AWS Lambda functions are currently managed using the Serverless Framework.
  • Vertex AI Pipelines are deployed across multiple tenants, and their deployment times need to be optimized.
  • Tenant provisioning is currently handled partially through scripts.
  • GitHub is the primary source control and automation platform.
  • SOC 2 controls require ongoing maintenance and improvement.

Requirements

  • 7+ years of experience in DevOps, cloud infrastructure, platform engineering, or systems engineering.
  • Strong experience with Terraform and infrastructure as code (IaC) in production environments.
  • Strong knowledge of and experience with SecOps.
  • Expertise in managing infrastructure across AWS and GCP.
  • In-depth understanding of building and maintaining GitHub Actions workflows.
  • Strong knowledge of GitHub Actions and CI/CD automation.
  • Understanding of multi-tenant provisioning automation.
  • Knowledge of least-privilege access and just-in-time (JIT) access.
  • Experience with SOC 2 technical controls.
  • Experience with production incident response and troubleshooting.
  • Experience with AWS services, including Lambda, IAM, VPC, Amplify, CloudWatch, Secrets Manager, and Parameter Store.
  • Experience with GCP services and capabilities, including Vertex AI, IAM, VPC networking, cloud security, and Cloud Logging/Monitoring/observability.
  • Experience migrating manual or framework-based infrastructure to Infrastructure as Code.
  • Strong understanding of Linux, networking, DNS, IAM, and cloud security fundamentals.
  • Experience implementing least-privilege access and production access controls.
  • Experience with monitoring, logging, and observability tools.
  • Solid scripting experience with Bash, Python, Go, or similar languages.
  • Ability to work independently, clarify ambiguity, and drive implementation with minimal guidance.
  • Strong documentation and communication skills.
  • Self-directed professional who can identify what needs to be done and take action without waiting for detailed instructions.
  • Ownership-oriented professional who treats infrastructure reliability, security, and deployment quality as their responsibility.
  • Ability to balance speed, quality, and maintainability without over-engineering.
  • Ability to design access, automation, and deployment workflows with least-privilege principles and auditability in mind.
  • Ability to work well with engineering, product, and leadership teams.
  • Documentation-focused professional who can leave behind clear runbooks, diagrams, and processes that the team can maintain.
  • Ability to focus on measurable improvements in deployment speed, reliability, security, and compliance.

The Candidate Can Work On Projects Such As

  • Move AWS Amplify deployments from direct Amplify workflows to GitHub-triggered CI/CD.
  • Migrate AWS Lambda/serverless infrastructure to Terraform.
  • Migrate tenant provisioning scripts to Terraform modules and GitHub Actions workflows.
  • Optimize Vertex AI Pipeline deployment times across multiple tenants.
  • Clean up and standardize AWS and GCP VPC configurations.
  • Implement GitHub repository rules for branch naming, Jira references, PR approvals, and protected branches.
  • Automate production permissions and implement just-in-time access for on-call engineers.
  • Improve SOC 2 controls around access, deployments, infrastructure changes, and evidence collection.
  • Improve monitoring, logging, and alerting, and create or enhance operational runbooks.
  • Document infrastructure architecture and deployment workflows.

Will be a plus

  • Experience with multi-tenant SaaS infrastructure.
  • Familiarity with Vertex AI Pipelines or ML/AI deployment workflows.
  • Understanding of optimizing long-running cloud deployment pipelines.
  • Familiarity with GitOps or declarative infrastructure patterns.
  • Experience with SOC 2, ISO 27001, or similar compliance frameworks.
  • Experience with just-in-time access tools or capabilities such as Okta, Teleport, AWS IAM Identity Center, Google IAM Conditions, or similar.
  • Knowledge of policy-as-code tools such as OPA, Checkov, Conftest, Sentinel, or Terraform Cloud policies.
  • Knowledge of Kubernetes, Docker, or containerized workloads.

Responsibilities

  • Manage and improve cloud infrastructure across AWS and GCP.
  • Convert existing manual provisioning scripts into Terraform.
  • Migrate AWS Lambda/serverless infrastructure from Serverless Framework to Terraform.
  • Build reusable Terraform modules for tenant provisioning, networking, IAM, deployment resources, and environment setup.
  • Ensure infrastructure is repeatable, version-controlled, documented, and auditable.
  • Improve infrastructure consistency across development, staging, and production environments.
  • Streamline deployment workflows across services and environments.
  • Move AWS Amplify deployments from direct Amplify workflows to GitHub-driven workflows.
  • Build and maintain GitHub Actions pipelines for application deployment, infrastructure deployment, and tenant provisioning.
  • Improve deployment speed, reliability, rollback capabilities, and visibility.
  • Optimize deployment workflows for multi-tenant environments.
  • Optimize long-running Vertex AI Pipeline deployments across multiple tenants.
  • Establish clear promotion workflows between environments.
  • Review existing tenant provisioning scripts and workflows.
  • Convert tenant provisioning into Terraform-backed infrastructure workflows.
  • Automate tenant provisioning using GitHub Actions.
  • Improve repeatability, traceability, and rollback capabilities for tenant setup.
  • Reduce manual operational work and deployment risk.
  • Review, clean up, and improve existing VPC structures.
  • Define clear networking patterns across AWS and GCP.
  • Improve segmentation between environments and tenants where appropriate.
  • Review DNS, routing, security groups, firewall rules, and cloud networking configurations.
  • Document cloud network architecture and create or improve operational runbooks.
  • Implement least-privilege access across AWS, GCP, GitHub, and deployment systems.
  • Automate permission management for engineering and production environments.
  • Restrict production access based on role and operational need.
  • Implement or improve just-in-time access for on-call engineers.
  • Improve auditability of privileged access and production changes.
  • Review secrets management and recommend improvements where needed.
  • Implement GitHub repository rules and engineering workflow standards, including: Branch naming conventions; Pull request requirements; Required Jira ticket references; Protected branches; Required reviews; Required CI checks; Environment-based approvals.
  • Improve consistency of engineering workflows across repositories.
  • Ensure GitHub workflows support both developer velocity and compliance needs.
  • Review and improve monitoring, logging, metrics, and alerting.
  • Help identify deployment bottlenecks, infrastructure risks, and recurring operational issues.
  • Improve incident response readiness through runbooks and documentation.
  • Support production incident troubleshooting when needed.
  • Recommend improvements to reduce operational toil and improve system reliability.
  • Help maintain and improve technical controls required for SOC 2.
  • Support controls related to: Access management; Change management; Deployment approvals; Infrastructure security; Production access; Audit logging; Evidence collection.
  • Ensure infrastructure and deployment processes are auditable and documented.
  • Help create or improve runbooks, diagrams, and process documentation needed for compliance

WE OFFER

  • US and EU projects based on advanced technologies.
  • Competitive compensation based on skills and experience.
  • Comprehensive private medical insurance.
  • Regular performance appraisals to support your growth.
  • Flexibility in workspace, either remote, our welcoming office or local coworking.
  • Bonuses for recommendations of new employees.
  • Bonuses for article writing, public talks, other activities.
  • 15 vacation days, 10 national holidays, 10 sick leaves.
  • Personalized learning program tailored to your interests and skill development.
  • Free tech webinars and meetups organized by Svitla.
  • Fun corporate online\offline celebrations and activities.
  • Well-established remote culture.
  • Awesome team, friendly and supportive community!

Similar jobs

Apply on LinkedIn