Cloudera logo

Senior Incident Response & Resiliency Engineer

Cloudera
Posted 1 hour ago
United StatesHybrid$114K–$142KEngineering & Development
Is this job info correct?

Business Area:

IT

Seniority Level:

Mid-Senior level

Job Description:

At Cloudera, we empower people to transform complex data into clear and actionable insights. With as much data under management as the hyperscalers, we're the preferred data partner for the top companies in almost every industry. Powered by the relentless innovation of the open source community, Cloudera advances digital transformation for the world’s largest enterprises.

We are seeking an experienced, strategic, and proactive Senior Incident Response & Resiliency Engineer to join our InfoSec team. In this role, you will lead and mature the operational readiness of Cloudera’s Incident Response (IR), Business Continuity Planning (BCP), Disaster Recovery (DR), and Backup/Recovery programs. Reporting to the Director of InfoSecOps, you will take ownership of maintaining emergency playbooks, leading tabletop exercises, facilitating crisis communications, and driving enterprise-wide resilience across systems and business operations.

The successful candidate is a structured, results-driven professional who excels at both strategic planning and operational execution. You bridge the gap between technical teams, infrastructure operations, and executive leadership, ensuring that our response plans are continuously validated, regulatory readiness remains sharp, and technical recovery capabilities align with organizational RTO/RPO targets.

As a Senior Incident Response & Resiliency Engineer you will:

  • Proactively map and maintain clear application and service ownership chains across matrixed teams to establish zero-day readiness and eliminate attribution gaps before threats emerge.

  • Serve as the primary operational connector during critical advisories and zero-day events, rapidly resolving ownership gaps, tracking down unassigned assets, and executing swift escalation workflows between IR, SecOps, and engineering leads.

  • Maintain, test, and optimize dynamic call trees, out-of-band crisis communication channels, and external outreach protocols (including Legal, PR, regulatory agencies, and retained forensics firms) to ensure operational readiness during major outages.

  • Support Response Engineers during high-severity (Sev-1/Sev-0) security incidents or disaster activations by facilitating crisis workflows, cross-functional stakeholder updates, and operational timelines.

  • SupportResponse Engineers during high-severity (Sev-1/Sev-0) security incidents or disaster activations by facilitating crisis workflows, cross-functional stakeholder updates, and operational timelines.

  • Lead the continuous review, expansion, and alignment of Cloudera’s Incident Response plans, BCP/DR frameworks, and technical playbooks against evolving threat vectors, regulatory requirements, and cloud infrastructure architectures.

  • Oversee enterprise Business Impact Analyses (BIAs) to establish baseline Recovery Time Objectives (RTOs) and Recovery Point Objectives (RPOs) across critical applications, mapping vendor dependencies and ensuring SLA alignment.

  • Partner with IT Infrastructure, Cloud, and Engineering teams to audit, validate, and document backup retention policies, immutable snapshot architectures, and air-gapped controls necessary to guarantee operational survival during ransomware scenarios.

  • Organize and lead hands-on technical DR restoration tests and failover simulations to validate data integrity, recovery workflows, and cross-functional business continuity during catastrophic infrastructure failures.

  • Proactively audit SIEM ingestion coverage, security tooling, and backup monitoring platforms to identify operational blind spots across cloud and on-premises environments.

  • Partner directly with application owners, IT infrastructure, and business unit leaders to define required log sources, audit event logging levels, and field-level telemetry needed for effective detection and incident forensics.

  • Document telemetry gaps and collaborate with the IR team to onboard new log feeds and validate data ingestion pipelines.

  • Design, execute, and evaluate comprehensive tabletop exercises (TTXs) spanning IR, BCP, and DR scenarios for both technical engineering teams and executive leadership.

  • Lead Post-Incident Reviews (PIRs) and Post-Exercise After-Action Reports (AARs) to capture critical takeaways, establish remediation roadmaps, and track cross-functional action items to resolution.

  • Develop key performance indicators (KPIs) and maturity metrics for IR readiness, backup success rates, BCP coverage, and DR testing results.

  • Centralize and maintain all resilience documentation, SOPs, and contact matrixes across primary corporate knowledge repositories.

We are excited about you if you have:

  • Education: Bachelor’s degree in Cybersecurity, Information Systems, Computer Science, Business Continuity, or a related field (or equivalent practical experience).

  • Experience: 4+ years of experience in cybersecurity operations, Incident Response, Business Continuity/Disaster Recovery (BCP/DR), IT Risk Governance or GRC.

  • Resiliency Expertise: Proven experience designing or maintaining enterprise BCP/DR plans, conducting BIAs, and overseeing technical backup restoration testing within cloud environments (AWS, GCP, Azure).

  • Technical & Framework Mastery: Strong working knowledge of NIST SP 800-61, ISO 22301, NIST SP 800-34, security automation frameworks (e.g., SOAR playbooks, workflow orchestration), and common cybersecurity and resiliency standards.

  • Operational Leadership: Demonstrated ability to lead cross-functional initiatives independently, facilitate complex tabletop scenarios, and drive technical post-incident remediations across matrixed teams.

  • Communication: Exceptional written and verbal communication skills, with a track record of translating complex technical risks into actionable operational requirements for non-technical stakeholders (Legal, PR, Executive Leadership).

  • Tools: Proficiency with modern security platforms, ticket management systems (Jira, Confluence), emergency messaging systems, and enterprise BCP/DR tracking tools.

  • Certifications: Preferred industry certifications include CISSP, CBCP (Certified Business Continuity Professional), CISM, GIAC (GCIH, GCFA), CySA+, or PMP.

The anticipated annual base salary range for this position is:

  • California: $114,000-$142,000

Individual compensation within the published range is determined by the candidate's skills, experience, qualifications, and primary work location. In addition to base pay, sales roles are eligible for Cloudera's commission plan, while non-sales roles are eligible for the corporate incentive plan. All employees receive a comprehensive benefits package.

This role is not eligible for immigration support or relocation.

What you can expect from us:

  • Generous PTO Policy

  • Support work life balance with Unplugged Days

  • Flexible WFH Policy

  • Mental & Physical Wellness programs

  • Phone and Internet Reimbursement program

  • Access to Continued Career Development

  • Comprehensive Benefits and Competitive Packages

  • Paid Volunteer Time

  • Employee Resource Groups

EEO/VEVRAA

#LI-SZ1
#LI-REMOTE

Similar jobs