Senior Site Reliability Engineer (f/m/d) - Eastern Europe (remote)
- Hiring from
- Europe
- Work type
- Remote
- Posted
Is this job info correct?
Show job descriptionHide job description
Hiring in Bulgaria, the Czech Republic, Romania or Slovakia via EOR or staffing partner
cplace is the platform for project and portfolio management that leading companies use to steer their most complex initiatives – grown in the DACH region and, with our launch in the US, on its way to becoming an international provider. As a Senior SRE in our Cloud Operations team, you will run our existing cplace Cloud 1.0 reliably and securely for customers in automotive, life sciences, manufacturing and retail – and actively shape the architecture and operating model of our Kubernetes-based cplace Cloud 2.0. AI is in cplace’s DNA: we use it intensively across all our work and expect you to apply it productively and critically and to help us take it further.
For EOR and staffing providers: Do you employ or can you place suitable candidates in Bulgaria, the Czech Republic, Romania or Slovakia via EOR or staff leasing (ANÜ)? We welcome your proposals.
cplace is the platform for project and portfolio management that leading companies use to steer their most complex initiatives – grown in the DACH region and, with our launch in the US, on its way to becoming an international provider. As a Senior SRE in our Cloud Operations team, you will run our existing cplace Cloud 1.0 reliably and securely for customers in automotive, life sciences, manufacturing and retail – and actively shape the architecture and operating model of our Kubernetes-based cplace Cloud 2.0. AI is in cplace’s DNA: we use it intensively across all our work and expect you to apply it productively and critically and to help us take it further.
- cplace Cloud 2.0: Co-building the Kubernetes platform – from cluster design, networking and storage to tenant isolation and scaling – plus planning and driving the migration of customer environments from Cloud 1.0
- Everything as code: Development of reusable Terraform modules, GitOps repositories and our own platform software (self-service portal, APIs, automation), including code reviews, automated tests and policy as code
- cplace Cloud 1.0: Operation and improvement of our environment of Linux servers, containers, SQL databases and Elasticsearch; lasting resolution of bugs, findings and capacity issues (incident and problem management, post-mortems); conversion of manual procedures into versioned, tested code (Ansible, Terraform, n8n)
- Improvement of monitoring, logging and alerting; ownership of backup & recovery, disaster recovery and business continuity, including regular testing
- Technical implementation of security and compliance requirements (e.g. SOC 2, ISO 27001, GDPR) and optimisation of cost and capacity
- Driving AI in operations, e.g. for incident and log analysis and agents for runbooks and routine tasks
- Close collaboration with product development for smooth releases, technical representation of the team in customer meetings, tenders and customer projects, and knowledge sharing through internal and external documentation and mentoring
- On-call duty in a fair rotation
- Degree in computer science or a related STEM field, or comparable vocational training, plus several years (ideally 5+) of experience as an SRE, DevOps or platform engineer in business-critical production environments
- Solid hands-on experience with Kubernetes in production (operations, upgrades, troubleshooting, storage, networking) and with at least one cloud provider – AWS, GCP, Azure and/or Hetzner Cloud a strong plus
- Deep experience with infrastructure as code (Terraform, Ansible), CI/CD, GitOps and Git-based collaboration via pull requests and code reviews – with modular, tested code that stays maintainable for the team
- Strong Linux skills and experience with SQL databases (e.g. MariaDB), Elasticsearch/OpenSearch and observability stacks (e.g. Prometheus, Grafana, Loki/ELK)
- Solid software engineering skills, ideally in Go or Python – tools and automation with tests and clean structure rather than one-off scripts; confident Bash scripting a given
- Good understanding of cloud security (e.g. network segmentation, secrets management, WAF)
- Hands-on experience with AI tools in everyday engineering and a good sense of their strengths and limits
- Customer-focused, structured way of working, ability to explain technical topics clearly, and fluent German and English
- The chance to shape a new cloud platform from the start, with real influence on architecture and technology.
- An experienced, collaborative team spanning SRE, Support, Escalation Engineering and Technical Account Management.
- Exciting enterprise customers and varied technical challenges.
- A working environment where modern AI tools are readily available and actively developed further.
For EOR and staffing providers: Do you employ or can you place suitable candidates in Bulgaria, the Czech Republic, Romania or Slovakia via EOR or staff leasing (ANÜ)? We welcome your proposals.