Browse DevOps & SRE Jobs
Search 37 curated tech job listings scraped in real-time from LinkedIn, Glassdoor, RemoteOK, and more. Filter by role, location, seniority, and source to find your next opportunity.
37 jobs found for "HPC"
Page 1 of 2
Staff Engineer, CI/CD & Cloud Infrastructure
…through builds, tests, artifacts, and deployments. You will work cross-functionally with firmware, application, and HPC engineers to keep the entire delivery pipeline fast, reliable, and observable. Key Responsibilities CI/CD & Build Engineering … strategies as needed Familiarity with Azure and GCP for secondary or hybrid requirements On-Premises HPC & Hybrid Infrastructure Provision, configure, and manage on-premises Linux HPC nodes used for secondary and tertiary…
DevOps Engineer
…intersection of systems engineering, software engineering, cloud-native infrastructure, and high-performance computing (HPC), contributing to platforms and services used by researchers and engineers tackling complex scientific and AI challenges. We value … operate scalable infrastructure and cloud-native platform services Contribute to Kubernetes-based AI/ML and HPC platforms, including CI/CD, GitOps, observability, security, and operational tooling Collaborate with researchers and engineers to support complex…
DevOps Engineer
…intersection of systems engineering, software engineering, cloud-native infrastructure, and high-performance computing (HPC), contributing to platforms and services used by researchers and engineers tackling complex scientific and AI challenges. We value … operate scalable infrastructure and cloud-native platform services Contribute to Kubernetes-based AI/ML and HPC platforms, including CI/CD, GitOps, observability, security, and operational tooling Collaborate with researchers and engineers to support complex…
Site Reliability Engineer
…manage AWS-based IT environments supporting enterprise applications, STAP platform rollouts, and high-performance computing (HPC) workloads End-to-end application rollout on STAP platforms, ensuring scalability, security, compliance, and performance readiness … Design, implement, and manage HPC platforms on AWS to support high end genomics, bioinformatics, and data intensive workloads Drive AWS migration initiatives, including assessment, planning, execution, and post migration optimization…
Developer Experience Engineer
…teams. You will work at the intersection of DevOps, software engineering, and high-performance computing (HPC), building systems that accelerate chip design, simulation, and AI model deployment in a cloud … Python skills for automation, scripting, and infrastructure development. Experience with Slurm job scheduling in an HPC or hybrid environment. Hands-on experience with observability and monitoring tools like Prometheus, Grafana, and OpenTelemetry…
Software Infrastructure Kubernetes Engineer
…productisation processes of our Machine Learning Software components. Work with our High-Performance Computing (HPC) AI platforms and gain invaluable experience in distributed system The Team The Software Infrastructure team provides critical … with Infrastructure as Code (IaC) tools (e.g. Terraform/OpenTofu) Experience with GitHub Actions Experience with distributed HPC systems Experience with modern observability tooling (e.g. Prometheus) Knowledge of Python/C++ (or similar language) Benefits…
Research Software Engineer (Automation and DevOps) - IT Services - 102683 - Grade 7
…post holder will contribute to initiatives in all areas of service innovation and delivery (including HPC, HTC, data analytics, image rendering and visualisation) enabling researchers to fully exploit the compute power … review and maintaining high engineering standards. Knowledge of Higher Education, Research and High-Performance-Computing (HPC). Informal enquiries to James Carpenter, email: [email protected] View our staff values and behaviours here…
Software Infrastructure Kubernetes Engineer
…productisation processes of our Machine Learning Software components. Work with our High-Performance Computing (HPC) AI platforms and gain invaluable experience in distributed system The Team The Software Infrastructure team provides critical … with Infrastructure as Code (IaC) tools (e.g. Terraform/OpenTofu) Experience with GitHub Actions Experience with distributed HPC systems Experience with modern observability tooling (e.g. Prometheus) Knowledge of Python/C++ (or similar language) Benefits…
DevOps Engineer New
…shell scripts and config files to streamline setup, reproducibility, and remote development (e.g., SSH-based HPC access). Define and document standardized ML development practices across teams. Collaborate with ML researchers, software engineers…
DevOps Engineer
…open-source technology to create and maintain the full stack of a High-Performance Computing (HPC) system that hosts a diverse range of applications served out to thousands of users. In This…
Staff ML Systems Engineer
…team is composed of experts in electrical engineering, electromagnetic simulation, ML/AI, and high-performance computing (HPC). We are inventing and leveraging novel techniques to solve the decades-old problem of automating circuit…
Senior DevOps Engineer
…Bonus Points Observability stack experience (Azure Monitor, Grafana, Prometheus, or similar) Automation for GPU, HPC, or ML infrastructure Database and data pipeline operations in cloud environments Relevant certifications preferred: Azure DevOps Engineer…
Platform Engineer / DevOps Engineer - Trading - $130,000-$250,000 CAD + Bonus
…Desirable skills: IaC – Terraform, Helm, Ansible, Pulumi, Bicep AWS or GCP DR Block / Blob Storage HPC Please apply ASAP for more information…
Senior Agent Platform Engineer
…Experience with model gateways and routing Exposure to scientific or high-performance computing: batch schedulers, HPC clusters, or simulation-heavy environments Open-source contributions to AI/agent tooling A track record of technical…
DevOps Engineer (TS/SCI)
…open-source technology to create and maintain the full stack of a High-Performance Computing (HPC) system that hosts a diverse range of applications served out to thousands of users. About…
Software Engineer, Devops Platform Engineering
…CI/CD pipelines, including Jenkins, GitHub Actions, GitLab CI, or equivalent. Experience working or supporting HPC or large-scale accelerated computing platforms. Hands-on experience with containerization (Docker, Kubernetes) and microservices. Deep expertise…
DevOps Engineer (USA)
…fast-paced environment Preferred Prior experience supporting quantitative research, trading platforms, or high-performance computing (HPC) environments Benefits Competitive salary plus bonus based on individual and company performance Collaborative, casual, and friendly…
DevOps Engineer
…infrastructure that the Infrastructure team manages – e.g. data catalogs and repositories, at-scale HPC compute services, user interfaces and APIs, and intelligent operations (AI/MLOps). Your primary job responsibilities will be to support…
Head of Physical Infrastructure
…operate What Sets You Apart Experience designing, deploying, or operating high-density compute infrastructure, HPC clusters, AI infrastructure, or similarly complex physical systems Strong understanding of how power, cooling, rack design, networking…
DevOps Engineer - AWS
…document infrastructure and collaborate across engineering teams Preferred Qualifications Experience with AI/ML, GPU, or HPC workloads Kubernetes on AWS (EKS or self-managed) Observability platforms: Prometheus, Grafana, Loki, OpenTelemetry, Datadog AWS cost…