Manager, Cloud Services
- Hiring from
- United States
- Work type
- Remote
- Posted
Show job descriptionHide job description
Led by an experienced management team and supported by a strong investor group, including large and experienced institutions and strategic partners, EdgeConneX offers a dynamic, fast-paced work environment where we are bringing flexibility, proximity, power, and connectivity to some of the world’s key businesses. With major offices in Herndon, Denver, Amsterdam, Singapore and Malaysia, we have a global footprint and a unified team of employees committed to providing a premier customer experience and delivering the full spectrum of data center solutions, from core to edge, like no other data center provider can do.
Focused on driving innovation and helping our customers define and deliver their own unique vision for the Edge, at any scale, in any market worldwide, for any requirement, we are building tomorrow’s data center infrastructure, today for some of the world’s most demanding Network, Content, and Cloud customers.
Job Description
We're looking for a Manager, Cloud Services who truly leads as a player-coach within a very lean team serving as an expert individual contributor who also guides and grows a team of DevOps, MLOps, and DataOps engineers. Reporting to the Senior Director of Cloud Services & Hybrid IT, you'll help oversee the engineering teams that build and operate our multi-cloud platform, translate the Senior Director's vision into execution, and stay deeply hands-on across infrastructure, ML operations, and data operations. You'll set standards and practices, unblock your engineers, and still write code and review PRs yourself.
This role spans three disciplines. On the DevOps side you'll drive IaC, CI/CD, and cloud platform operations across Azure (primary), AWS, and GCP. On the MLOps side, you'll help build the platform and pipelines that operationalize AI/ML workloads and the organization's AI tooling. On the DataOps side you'll support the data pipelines and platforms the business depends on. You'll also support the broader enterprise technology organization as part of our Hybrid IT charter. This position reports to our Senior Director, Cloud Services & Hybrid IT and can be based in our US Headquarters in Herndon, VA or can be based remotely in the U.S. The expectation would be standard US EST work hours, so easiest if someone is based on the East Coast. If remote, the expectation is to travel 1 week quarterly to our Herndon, VA HQ.
What you'll do
Leadership & Team Oversight
• Help oversee and mentor a small team of DevOps, MLOps, and DataOps engineers (consultants and 1-2 internal direct reports in the near-term), setting technical direction and day-to-day priorities under the Senior Director's guidance.
• Translate the Senior Director's vision into execution by defining the procedures, standards, and practices that shape how the teams work.
• Grow the team's capability through coaching, code review, hiring input, and career development.
• Coordinate across the DevOps, MLOps, and DataOps functions to keep priorities, tooling, and standards aligned.
• Own delivery for your teams' commitments: planning, tracking in Jira, and reporting progress to leadership.
• Partner with enterprise technology, product engineering, security, and data teams as a trusted technical leader.
Hands-on engineering as an expert Individual Contributor
• Evolve and enforce our infrastructure-as-code standard (set with the Senior Director): Terraform with Azure Verified Modules (pinned), one root and state per account/subscription, and PR-based workflows with automated review.
• Lead CI/CD in GitHub Enterprise (GitHub Actions), including completing the migration from Azure DevOps.
• Manage Kubernetes clusters (EKS and AKS) with GitOps-based delivery (ArgoCD or Flux) across Windows Server and Linux workloads.
• Drive cloud governance automation across Azure, AWS, and GCP: custom RBAC, Azure Policy enforcement, secrets management (Key Vault), and Terraform state management.
• Create and maintain Azure Enterprise Applications and App Registrations (SSO, API permissions, service principal lifecycle), and integrate Entra ID SSO/SCIM across platforms.
• Build and operate the MLOps platform: model deployment and serving pipelines, model registry, and monitoring for AI/ML workloads, using MLOps tooling such as MLflow, Kubeflow, or Azure ML.
• Design and build custom MCP (Model Context Protocol) servers and AI tooling to streamline how the organization uses AI, and enable rapid AI-assisted (“vibe coding”) development.
• Support the DataOps platform and pipelines: ETL/ELT orchestration, data platform reliability, CI/CD for data (Synapse, Power BI reporting, data warehousing), and database schema migration and versioning (Liquibase, Atlas).
• Provide hands-on support to the enterprise technology team on infrastructure and identity needs.
Required qualifications
• 10+ years in cloud, DevOps, platform, or data/ML engineering, with at least 2 years leading or managing engineers (formal management or true player-coach).
• Expert, hands-on Terraform experience at scale: module design, state management, and multi-environment patterns.
• Deep, hands-on expertise in both Azure and AWS (networking, RBAC, Key Vault, compute); familiarity with GCP.
• Proven experience with cloud migrations and multi-cloud setups across Azure, AWS, and GCP.
• Production experience with CI/CD systems (GitHub Actions strongly preferred) and Git-based PR workflows.
• Hands-on Kubernetes experience (EKS and AKS) and GitOps delivery with ArgoCD or Flux.
• Working knowledge across at least two of the three disciplines (DevOps, MLOps, DataOps) and the ability to lead engineers in all three.
• Experience with enterprise identity and SSO (Entra ID, SAML, SCIM, OIDC), including Azure Enterprise Applications and App Registrations.
• Strong scripting and automation skills (PowerShell, Python, or Bash).
• Comfortable administering both Windows Server and Linux.
• Solid written and verbal communication skills, with the ability to explain trade-offs to leadership and mentor engineers.
• Demonstrated use of AI-assisted development and a track record of building automation and internal tooling.
Preferred qualifications
• Hands-on with MLOps tooling (MLflow, Kubeflow, SageMaker, or Azure ML) and DataOps database change management (Liquibase or Atlas).
• Experience building custom MCP servers or comparable AI integrations.
• Experience leading an Azure DevOps-to-GitHub or comparable toolchain migration.
• Familiarity with Jira/Atlassian and Coda, or comparable work-tracking and documentation tooling.
• Background modernizing legacy applications (WebForms, .NET) toward containers and modern deployment patterns.
• Azure SQL / SQL Managed Instance operational experience.
• Relevant certifications (Azure, AWS, HashiCorp Terraform).
What sets you apart
• An innovative mind with high technical aptitude; you spot opportunities to automate and improve, and you act on them.
• Comfort operating in ambiguity in a newly formed, fast-growing team building foundational capabilities.
• A collaborative mindset that defaults to partnership over gatekeeping across organizational boundaries.
• Strong written and verbal communication; you explain trade-offs to leadership and mentor engineers with equal clarity.
• Judgment about when to build versus buy, standardize versus accommodate, and move fast versus set guardrails first.