Hiring for Cloud Operations and Engineering role for a USA Based company
Location : Remote
Experience
6+ years in infrastructure, DevOps, platform engineering, cloud operations, or SRE;
~4 years hands-on on GCP.
Mandate skills : Strong GCP (Multiple Services) / IaC / Terraform / GKE
CTC : Open
Job Requirements
Hands-on experience with cloud infrastructure design, provisioning, deployment, and troubleshooting.
Strong understanding of GCP core services: Compute Engine, GKE, Cloud Run, Cloud Storage, Cloud SQL, IAM, VPC, Load Balancing, Cloud Logging, and Cloud Monitoring.
Strong Infrastructure as Code experience using Terraform: HCL, reusable modules, variables, remote state, state locking, workspaces/environments, plan/apply lifecycle, drift handling, and code review.
CI/CD experience using Cloud Build, Jenkins, GitHub Actions, GitLab CI, or equivalent.
Good understanding of networking: VPC, subnetting, firewall rules, private connectivity, load balancers, DNS, NAT, VPN, Interconnect, Cloud Router, Shared VPC, and hub-spoke architecture.
Production troubleshooting exposure: logs, metrics, alerts, incident response, performance, reliability, and cost optimisation.
Scripting knowledge in Bash, Python, or Go.
Good security basics: IAM least privilege, service accounts, secrets, KMS, audit logging, vulnerability awareness.
Ability to explain real-world implementation decisions, trade-offs, and failure scenarios.
Detailled JD
Generic Skills (Must Have)
Terraform HCL, modules, remote state, state locking, reusable templates, environment separation.
Containerisation and Kubernetes basics: deployments, services, ingress, config maps, secrets, autoscaling, troubleshooting.
Linux administration and basic scripting.
GCP Skills (Must Have)
GCP Compute Engine, GKE, Cloud Run, IAM, VPC, Load Balancers, Cloud NAT, Cloud Router, Cloud DNS.
CI/CD using Cloud Build, Jenkins, GitHub Actions, or GitLab CI.
Monitoring and logging using Cloud Monitoring, Cloud Logging, Prometheus, Grafana, or equivalent.
Nice to have (Trainable)
GKE Autopilot vs Standard decisioning.
Anthos / GKE Enterprise exposure.
Cloud Deploy, Artifact Registry.
FinOps and cost optimisation.
Policy as Code using OPA, Sentinel, or Terraform policy controls.Artifact Registry
Skills: jenkins,gke,cloud sql,devops,firewalls,cloud build,finops,autopilot,cloud,linux,google cloud,cloud deploy,python,artifact registry,github actions,compute engine,cloud storage,terraform,cloud logging,terraform policy controls,infrastructure,state locking,resusable templates,hcl,load balancing,gitlab ci,bash,subnetting,anthos,vpc,google cloud platform,cloud run,iam,gcp,sentinel,remote state