Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact mahmoud@relomote.com · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Tribus logo

Site Reliability Engineer

Tribus
Posted 8 hours ago
🛂Visa sponsorship
🇦🇺Australia
📁Engineering & Development
Is this job info correct?

Site Reliability Engineer (Trading Infrastructure) Multiple organisations

Sponsorship Available

Sydney or Hong Kong | Onsite | Global Trading Environment


Join a high-performance engineering team responsible for the reliability, scalability and operational excellence of the infrastructure powering a global electronic trading platform.

This is a hands-on Site Reliability Engineering role sitting close to the trading stack, where you'll help build resilient production systems, improve automation, and work alongside software engineers to support latency-sensitive applications operating across global financial markets.


Unlike traditional SRE environments focused primarily on SLI/SLO metrics, this team takes an event-driven approach to reliability engineering. You'll design intelligent monitoring and automated operational workflows that identify abnormal system behaviour, infrastructure anomalies and production events before they impact trading. The focus is on actionable signals, rapid diagnosis and engineering-led remediation rather than simply measuring service health.


What You'll Be Doing


  • Design, build and maintain highly reliable Linux-based production infrastructure.
  • Develop Infrastructure as Code using Terraform and Ansible.
  • Build observability platforms using Prometheus, Grafana and Splunk.
  • Create event-driven monitoring, intelligent alerting and automated remediation workflows.
  • Improve operational tooling and incident response across business-critical trading systems.
  • Work with PostgreSQL and InfluxDB supporting production data platforms.
  • Support workflow orchestration using Prefect.
  • Administer virtualisation and storage platforms including CEPH, VMware, KVM and Proxmox.
  • Partner closely with software engineers developing C++, Java, C# and Python applications.
  • Improve CI/CD, deployment tooling and overall platform reliability through automation.


What We're Looking For


  • Strong Linux systems administration and troubleshooting experience.
  • Python software development skills, specifically building tools for internal teams
  • Experience with Infrastructure as Code using Terraform, Ansible or similar tools.
  • Hands-on experience with observability platforms such as Prometheus, Grafana or Splunk.
  • Experience designing monitoring strategies that focus on operational events, anomaly detection and actionable alerts rather than purely SLI/SLO-driven metrics.
  • Familiarity with PostgreSQL, InfluxDB or other production databases.
  • Strong networking fundamentals including TCP/IP, routing and firewalls.
  • Experience with Docker and modern development tooling.
  • Exposure to virtualisation platforms such as VMware, KVM, Proxmox or CEPH.
  • Experience supporting low-latency, distributed or mission-critical production environments is highly regarded.


Why This Role?


  • Work on infrastructure that directly supports real-time trading.
  • Solve complex reliability challenges where milliseconds matter.
  • Influence how monitoring, automation and operational engineering are built from the ground up.
  • Collaborate with experienced infrastructure and software engineers in a highly technical environment.
  • Work in a culture that values engineering ownership, continuous improvement and pragmatic problem solving.


If you're passionate about Linux, automation, observability and building resilient production platforms, we'd love to hear from you.

Similar jobs

Similar jobs

Optiver Private Jobs logo

Data - Site Reliability Engineer

Optiver Private Jobs

🇦🇺AustraliaJul 23, 2026, 2:52 PM UTC
coreflow logo

Software Engineer (Site Reliability)

coreflow

🇦🇺AustraliaJul 29, 2026, 9:56 PM UTC
UQ

Postdoctoral Research Fellow in Power System Stability

Uq

🇦🇺AustraliaYesterday
UE

Registered Nurse

UPA External

🇦🇺Australia2 hours ago
Anthropic logo

Manager, APAC Recruiting

Anthropic

🇦🇺Australia2 hours ago
LH

Front Office Manager Lead the Front Office team at Alamanda Palm Cove

Lancemore Hotel Group

🇦🇺Australia8 hours ago