Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
UN

Sr. Engineer II, DevOps, NG-SIEM (Hybrid, Bucharest)

Undelucram.ro
Posted 6 hours ago
🇷🇴Romania🏢Hybrid📁Engineering & Development
Is this job info correct?
Undelucram.ro on behalf of:

Crowdstrike SRL

About the Role:

Our mission is to make all of our customers' security-relevant data continuously available for automated detection and response, threat hunting, and other Falcon platform use cases. To enable this, the systems behind NG-SIEM (next-generation security information and event management) are growing to accommodate >100 PB of event and action data ingested every day, up to 10 years of retention, and dozens of millions of queries per hour across large sections of the data stored, for tens of thousands of customers.

As a Senior Engineer II on the newly established NG-SIEM EPICS (End-to-End Performance, Incident-response, Cost, and Scaling) team, you will own the reliability and scalability of the security industry's largest SIEM platform — treating these as software engineering problems rather than purely operational ones.

The NG-SIEM platform comprises many decoupled components interacting across complex pipelines. As we scale, ensuring end-to-end health across ingest, search, and workflow execution requires deep cross-service expertise and coordinated action. You will be the engineer who builds the observability, automation, and scaling systems that keep the entire platform performing — not just individual components. You will join a distributed team of high-ownership technical leaders who share a strong passion for our mission: to stop breaches.

This is a hybrid role based in one of our offices in Bucharest (Romania), 2-3x a week.

What You'll Do:

  • End-to-end observability: Design, build, and maintain monitoring and synthetic test suites that provide deep visibility into the health of the entire NG-SIEM pipeline — from ingest through search and workflow execution — enabling rapid root cause analysis across component boundaries.
  • Coordinated scaling: Engineer orchestrated scaling solutions that treat the NG-SIEM pipeline as a unified system, proportionally increasing resources across all dependent components (Kafka, ingest pipelines, downstream services) to eliminate cascading bottleneck patterns.
  • Incident response engineering: Serve as a subject matter expert during platform-wide incidents (P2 and above), applying cross-service knowledge to diagnose and resolve multi-component failures. Partake in follow-the-sun on-call rotations, providing incident commander coordination for critical platform-wide events.
  • Capacity planning and cost management: Build and refine models for end-to-end capacity forecasting that account for all pipeline dimensions, including partner team dependencies (data services, GPS). Develop tooling to continuously track and surface cost drivers across the platform.
  • Automation and runbooks: Transform manual standard operating procedures into automated remediation workflows — including pipeline-wide scaling responses, CID rebalancing, and infrastructure healing — with the goal of resolving issues before customers are impacted.
  • Cross-team collaboration: Partner with cell-level teams, product engineering, GDI/3PI, and external stakeholders (e.g., CSM) to triage SLO breaches, drive problem management for large reliability efforts, and ensure consistent communication during incidents.
  • Platform improvements: Use your broad NG-SIEM knowledge to identify and drive systemic improvements across teams, contributing to the platform's long-term resilience and efficiency.

What You'll Need:

  • A passion for reliability engineering and curiosity about how large-scale running systems behave under pressure;
  • 10+ years of experience in software engineering, site reliability engineering, or platform engineering, with significant time spent on large-scale distributed systems, and the ability to make pragmatic tradeoffs between short-term delivery needs and long-term platform goals;
  • Proven experience utilizing AI technologies to enhance decision-making, streamline workflows and processes, improve efficiency and drive business outcomes.
  • Strong proficiency in at least one systems programming language (Go, Java, Rust, or C++) and one scripting language (Python, Bash);
  • Deep experience with end-to-end observability — building monitoring pipelines, defining SLIs/SLOs, and creating dashboards that drive actionable insights across multi-service architectures;
  • Demonstrated ability to diagnose and resolve complex incidents spanning multiple distributed components operating 24/7;
  • Experience with coordinated capacity planning and scaling for systems with significant infrastructure footprints;
  • Hands-on experience with streaming platforms (Kafka or similar) and understanding of backpressure, partition management, and consumer group dynamics at scale;
  • Familiarity with infrastructure-as-code, CI/CD pipelines, and automated deployment practices;
  • A can-do attitude — you thrive collaborating in a team and are not afraid of taking on responsibilities;
  • Strong written and verbal communication skills — you will lead incident communications and produce post-incident analyses that drive lasting improvements;
  • Comfort working across time zones with globally distributed teams.

Bonus Points:

  • Experience in a similar reliability or platform engineering role at a hyperscaler (AWS, Azure, GCP) or large-scale SaaS provider;
  • Track record of building automated remediation and self-healing infrastructure;
  • Experience with cost modeling and unit economics for large compute and storage footprints;
  • Familiarity with cloud-native architectures and serverless computing paradigms;
  • Hands-on experience operating platforms processing over 1 trillion events per day or more than 10 PB of data per day;
  • Exposure to or experience with Log Management, cybersecurity products, or security operations workflows;
  • Experience with disaster recovery planning and execution for multi-region systems.

Similar jobs

Similar jobs

CI

JavaScript Engineer

Ciklum

🌍Poland, Romania4 hours ago
Vodafone logo

Security Engineer

Vodafone

🇷🇴Romania4 hours ago
Lifted, an Upwork Company™ logo

PHP & Azure Developer

Lifted, an Upwork Company™

🌍Philippines, Romania4 hours ago
Proxify logo

Senior MS Power BI Developer

Proxify

🌍Argentina, Italy, Poland, Romania4 hours ago
Global logo

Requirements Engineer

Global

🇷🇴Romania4 hours ago
Global logo

Senior Fullstack Developer (.Net & Angular)

Global

🇷🇴Romania4 hours ago