Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
SigNoz logo

Sr Site Reliability Engineer

SigNoz
Posted Jun 23, 2026, 7:39 PM UTC
🇮🇳India🏠Remote📁Engineering & Development
Is this job info correct?

About SigNoz SigNoz is an open-source observability platform that helps modern engineering teams monitor, debug, and optimize their applications with deep visibility into metrics, traces, and logs — all in one place. We're built natively on OpenTelemetry and offer both self-hosted and cloud options, so teams can run observability the way they want, without vendor lock-in. We are growing fast and building core developer infra products. And we are not fooling around: 27,000+ GitHub stars 800+ customers 7,000+ members in our Slack community Role: Sr Site Reliability Engineer (SRE) We're looking for an SRE to own the reliability, scalability, and operability of the SigNoz cloud platform. You'll keep a petabyte-scale observability system fast and dependable — making sure the people who trust us to watch their systems can always trust ours. The platform team handles infra, scalability of SaaS, ingest pipelines, staging environments, automation, and the operational backbone of the product. This is a deeply hands-on role for someone who understands what actually breaks in production at scale — and enjoys fixing it for good. What we're looking for Kubernetes at scale — not just "I've deployed to k8s," but real fluency with the nuances and gotchas: resource tuning, autoscaling behavior, networking, stateful workloads, upgrades, and the failure modes that only show up under load Working knowledge of ClickHouse — operating it, tuning queries, and understanding its behavior at scale — is a strong plus Knowledge of Golang is a plus (most of our stack and tooling is in Go) Familiarity with OpenTelemetry and running large-scale data ingest pipelines is a plus What you'll work on You'll work with a high-caliber team across areas like: Reliability of the SigNoz cloud platform: SLOs/SLIs, error budgets, incident response, and on-call practices that don't burn people out Scaling the ingest path — making it robust to bursts while maintaining data freshness SaaS auto-scalability and capacity planning across a petabyte-scale system Operating and tuning ClickHouse and the data layer for performance and cost Kubernetes infrastructure: cluster operations, upgrades, multi-tenancy, and the automation that keeps it boring Observability of SigNoz itself — we dogfood our own product, so you'll help make it world-class Infrastructure-as-code, CI/CD, and the tooling that lets a small team operate big systems What will make you successful 5–8 years in SRE, infrastructure, or platform/backend roles operating production systems at scale Deep, practical Kubernetes experience — you know where the bodies are buried Strong grasp of distributed systems failure modes, performance debugging, and capacity planning Comfortable in code (Go preferred) — you automate and fix things, not just configure them Loves open source — ideally with prior contributions to OSS projects (any size) Comfortable in a high-ownership, fast-moving, remote-first environment Strong communication — can write clear runbooks and tech docs and explain trade-offs Nice-to-haves Past experience on platform/infra/SRE teams of Series B+ startups Hands-on experience operating ClickHouse, Kafka, or similar high-throughput data systems Experience in observability (monitoring / logging / tracing) and with OpenTelemetry Why you'll love working at SigNoz Work on a globally used open-source project that engineers actually love Huge scope and ownership — your work directly shapes how teams adopt SigNoz Collaborate with a high-caliber team who just can't stop shipping Remote-first, async-friendly culture Opportunity to help define the future of open-source observability

Similar jobs

Similar jobs

TE

AI Reliability & Red Teaming Engineer

Technopals

🇮🇳India5 hours ago
Mastercard logo

Site Reliability Engineer II

Mastercard

🌍India, Mexico, United States18 hours ago
Cognativ logo

Senior Site Reliability Engineer (Hiring Globally)

Cognativ

🌍Argentina, Brazil, India, Mexico, Serbia, Spain, United KingdomYesterday
Nutanix logo

Systems Reliability Engineer 2 - Enterprise Tech Support (Virtualization & Networking)

Nutanix

🇮🇳India2 days ago
Sabre logo

Site Reliability Engineer

Sabre

🇮🇳India2 days ago
Eurofins logo

Azure Site Reliability Engineer

Eurofins

🇮🇳India2 days ago