Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Arcadion logo

HPC Specialist

Arcadion
Posted Jul 2, 2026, 7:17 AM UTC
🇨🇦Canada🏢Hybrid📁Other
Is this job info correct?

🔬 HPC Specialist – Role Overview Arca dion is seeking an HPC (High-Performance Computing) Specialist responsible for the design, deployment, optimization, and management of high-performance computing systems, often used in scientific research, engineering simulations, AI/ML workloads, and large-scale data analytics. 🧠 Core Responsibilities Architecture & System Design Design scalable HPC clusters (on-prem, cloud, or hybrid) Choose appropriate CPUs, GPUs, interconnects (e.g., InfiniBand), and storage Configure Slurm, PBS, or OpenHPC job schedulers Cluster Deployment & Maintenance Install and manage Linux-based compute nodes Maintain job schedulers and resource managers Integrate monitoring tools (Prometheus, Grafana, Nagios) Performance Tuning & Optimization Benchmark workloads and tune for performance (e.g., MPI, CUDA, OpenMP) Optimize I/O and inter-node communication Ensure efficient job execution and queue handling Cloud & Hybrid HPC Integration Deploy and manage cloud-based HPC environments (Azure CycleCloud, AWS ParallelCluster, Google Cloud HPC Toolkit) Optimize workload portability and orchestration (e.g., Singularity, Kubernetes with KubeFlow or Volcano) AI/ML & GPU Workload Support Manage AI pipelines that require HPC acceleration (e.g., LLM training) Optimize GPU usage (NVIDIA A100/H100, AMD MI300) Interface with TensorFlow, PyTorch, and HPCML tools Security & Compliance Implement security best practices for multi-user environments Support data governance for sensitive or regulated workloads Maintain audit trails and role-based access control User & Application Support Assist researchers and data scientists with job submissions and optimization Develop documentation, training materials, and run code validation sessions 🛠️ Technical Skills & Tools Category Technologies Schedulers Slurm, PBS, Torque, LSF HPC OS & Config RHEL/CentOS, Rocky, Ubuntu Server HPC File Systems Lustre, BeeGFS, GPFS Parallel Computing MPI, OpenMP, CUDA Monitoring/Telemetry Prometheus, Grafana, Ganglia Cloud HPC AWS HPC, Azure CycleCloud, GCP HPC Toolkit Containers Singularity, Apptainer, Docker, Kubernetes AI/ML Support PyTorch, TensorFlow, Horovod, MLFlow DevOps Tools Ansible, Terraform, Git, Jenkins 🎓 Qualifications Bachelor’s or Master’s in Computer Science, Engineering, Physics, or a related technical field 3–7 years of experience in HPC environments Experience supporting AI/ML teams is a major asset Certifications: NVIDIA DLI, AWS Certified HPC Specialist, Linux+ or RHCE

Similar jobs

Similar jobs

Distribution stox logo

Conseiller(ère) rémunération et avantages sociaux

Distribution stox

🇨🇦Canada5 hours ago
Distribution stox logo

Directeur(trice) de comptes

Distribution stox

🇨🇦Canada5 hours ago
AO Globe Life logo

Entry Level Management-WFH

AO Globe Life

🇨🇦Canada5 hours ago
GV

Major Gifts Manager

Greater Vancouver Food Bank

🇨🇦Canada5 hours ago
Mercor logo

Feedback Synthesis Specialist - Remote | Upto $120/hr

Mercor

🌍Australia, Canada, India, United Kingdom, United States1 hour ago
Mercor logo

AI Red Team Specialist - Remote | Upto $22/hr

Mercor

🌍Australia, Canada, India, United Arab Emirates, United Kingdom, United States1 hour ago