Browse DevOps & SRE Jobs

Search 1005 curated tech job listings scraped in real-time from LinkedIn, Glassdoor, RemoteOK, and more. Filter by role, location, seniority, and source to find your next opportunity.

Filters:
× Clear filters

1,005 jobs found for "Distributed Systems"

Page 1 of 51

Distributed Systems Engineer

Cadence Burnaby, British Columbia, Canada

…load balancing across compute clusters Build monitoring and observability for distributed workflows Optimize task granularity and dependency management Required Expertise Distributed Systems 3+ years building distributed systems with Python Experience with distributed … work alongside experienced engineers building greenfield distributed infrastructure with modern tools. This is an opportunity to grow your expertise in production-scale distributed systems while solving challenging problems in chip design. What…

LinkedIn

Java Distributed Systems Engineer - Remote Work | REF#300859

BairesDev Guayas, Ecuador $301k+

…with our vacancies, setting you on a path to exceptional career development and success. Java Distributed Systems Engineer at BairesDev As a Java Distributed Systems Engineer, you will design and build complex … performance through rigorous testing of distributed failure scenarios. What We Are Looking For 8+ years of experience in Software Engineering with a focus on distributed systems. Proven expertise in designing and building…

LinkedIn

Java Distributed Systems Engineer - Remote Work | REF#299353

BairesDev Chile $299k+

…with our vacancies, setting you on a path to exceptional career development and success. Java Distributed Systems Engineer at BairesDev As a Java Distributed Systems Engineer, you will design and build complex … performance through rigorous testing of distributed failure scenarios. What We Are Looking For 8+ years of experience in Software Engineering with a focus on distributed systems. Proven expertise in designing and building…

LinkedIn

SRE/Dev Ops Engineer New

CrowdStrike Sunnyvale, CA $401k+

…security with the world’s most advanced AI-native platform. We work on large scale distributed systems, processing almost 3 trillion events per day and this traffic is growing daily. Our customers … resilience validation for production systems Proven ability to build and maintain systems with zero tolerance for data loss or downtime Advanced observability experience including Prometheus, Grafana, distributed tracing (Jaeger/OpenTelemetry), and large-scale…

LinkedIn

Staff Software Development Engineer SDM

Delinea Redwood City $150k – $185k

…clusters, and web applications. As a Staff Software Engineer, you will apply deep expertise across distributed systems, security, and cloud architecture to help shape the technical direction of StrongDM’s proxy … span multiple teams, and serve as a recognized subject matter expert on secure, high-scale distributed systems written in Go. You will report to the team’s engineering manager. What Youâ…

RemoteOK

Java Engineer Lead - Capital Markets

SMBC Group New York, NY $265k+

…Markets functions. This role will lead the development of highly scalable, event-driven microservices and distributed systems that enable Trading, Risk Management, Settlements, Collateral Management, Finance, Regulatory Reporting, and other core Capital … engineering ecosystems PostgreSQL and relational database design Redis and distributed caching technologies Event-Driven Architecture Event Sourcing and CQRS patterns Distributed systems design and resiliency engineering Real-time data processing and streaming…

LinkedIn

Cloud Infrastructure Engineer

Rune Technologies Seattle, WA $190k – $215k

systems (metrics, logs, tracing) to monitor system health and performance. Collaborate with backend and infrastructure teams to ensure services are designed for scalability, resilience, and operability. Support integration with edge-deployed systems … Kubernetes at the edge (e.g., k3s, RKE2) or deploying to constrained/on-prem systems. Experience building or operating systems that interact with distributed or intermittently connected nodes. Exposure to GovCloud or regulated environments. Experience…

LinkedIn

Platform Engineer

Citadel Securities Miami, FL

…York. Our Research Platform teams are at the forefront of designing and developing the distributed systems, data platforms, and research infrastructure that power large-scale simulations, analytics, and model development across … distributed systems, data-intensive applications, and system design fundamentals Experience with cloud platforms (AWS, GCP, Azure) and modern infrastructure technologies Experience with technologies such as Kafka, Kubernetes, Spark, Airflow, distributed databases…

LinkedIn

DevOps Engineer

MTN Global Fort Lauderdale, FL

…communications and real-time services for customers around the world. Our engineering team develops scalable distributed systems that process millions of events, orchestrate cloud services, integrate with multiple providers, and deliver highly … environments Microservices architecture Event-driven systems Nice to Have Kubernetes Administrator (CKA) Terraform Associate Certification Google Cloud Certification Experience supporting 24/7 production systems Experience working with distributed event-driven platforms What…

LinkedIn

Senior Platform Engineer, Event Streaming Platform (Kafka)

wolt Berlin

…thousands of services across 3 companies. This is deeply technical work: you'll own distributed systems at scale, contribute to long-term architecture decisions, and be working on the reliability of critical … reliability, scalability, observability and automation of Kafka, Event Bus and related streaming components. Debug complex distributed systems issues — latency issues, performance bottlenecks, capacity constraints, partition imbalances — and drive them to resolution. Build…

Arbeitnow

Senior Platform Engineer, Event Streaming Platform (Kafka)

Wolt - English Berlin, Berlin, Germany

…thousands of services across 3 companies. This is deeply technical work: you'll own distributed systems at scale, contribute to long-term architecture decisions, and be working on the reliability of critical … reliability, scalability, observability and automation of Kafka, Event Bus and related streaming components. Debug complex distributed systems issues — latency issues, performance bottlenecks, capacity constraints, partition imbalances — and drive them to resolution. Build…

Arbeitnow

Software Development Engineer III

Expedia Group London, England, United Kingdom

…algorithms in Java or Kotlin, with familiarity across the JVM stack, system design, and distributed systems—and can understand highly complex systems, design moderately complex services, and guide integrations across teams within … platform tooling and/or infrastructure as code. Hands-on experience designing, building, and operating large-scale, distributed systems and services. Strong proficiency in system design, API design, and data modelling. Experience using modern…

LinkedIn

Engineer

Tata Consultancy Services Pittsburgh, PA $401k+

…Java/Kafka Developer with strong expertise in Java, J2EE, Spring Boot, Microservices, Kafka, REST APIs, and Distributed Systems. The ideal candidate will be responsible for designing, developing, and implementing scalable, high-performance, event … topics, partitions, and consumer groups. Monitor message delivery, performance, and reliability. Troubleshoot distributed messaging and data consistency issues. Distributed Systems Engineering Develop highly available, fault-tolerant distributed applications. Implement asynchronous communication…

LinkedIn

DevOps Software Engineer

Garmin Cochrane, Alberta, Canada

…Essential Functions Work on modern tech stacks and frameworks to deliver high-performant, highly reliable distributed systems for data processing and analysis for Global Garmin teams Design database systems to manage … MongoDB or Elastic/Opensearch Experience with maintaining small to medium sized network infrastructure Experience with distributed systems System administrator / IT experience with Windows (10, 11 and server) System administrator / IT experience with Linux…

LinkedIn

Software Engineer, DevOps New

FieldAI Irvine, CA

…maintaining CI/CD systems in complex engineering environments Strong scripting or programming skills in Python, Bash, Go, or similar languages Solid understanding of Linux systems, networking, and distributed systems fundamentals Strong problem-solving … ArgoCD, or related infrastructure tooling Experience managing large monorepos, build systems, and developer platforms Experience building observability and reliability systems for distributed infrastructure Kubernetes certifications such as CKA, CKAD, or CKS Bias…

LinkedIn

Software Engineer, DevOps

FieldAI Irvine, CA

…maintaining CI/CD systems in complex engineering environments Strong scripting or programming skills in Python, Bash, Go, or similar languages Solid understanding of Linux systems, networking, and distributed systems fundamentals Strong problem-solving … ArgoCD, or related infrastructure tooling Experience managing large monorepos, build systems, and developer platforms Experience building observability and reliability systems for distributed infrastructure Kubernetes certifications such as CKA, CKAD, or CKS Bias…

LinkedIn

Site Reliability Engineer

Helsing London, England, United Kingdom

…manifest templating. Experience with policy-as-code tools like OPA/Gatekeeper System Administration: deep understanding of Linux/Unix system administration and highly available, distributed systems Comfortable building out data and telemetry pipelines for debugging … right up to the state of the art in technical innovation, be it reinforcement learning, distributed systems, generative AI, or deployment infrastructure. The defence industry is entering the most exciting phase…

LinkedIn

Systems Engineer

Langdock Berlin €90k – €140k

…company priorities change. The constant is sustained investigation of unfamiliar systems and responsibility for turning that understanding into a production system. Systems Engineers own the difficult mechanism and the measurable improvement … looking for You have owned technically difficult production systems in areas such as distributed systems, databases, runtimes, inference, networking, or storage. You operate what you build and remain responsible for it after…

Arbeitnow

Site Reliability Engineer

Vynca Nevada, United States $140k – $150k

…production Kubernetes environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay … tracing, alerting, and incident response. Strong understanding of Linux systems administration, networking, cloud architecture, and distributed systems fundamentals. Experience designing, implementing, and maintaining CI/CD pipelines and deployment automation. Strong problem-solving skills…

LinkedIn

Site Reliability Engineer

Vynca Florida, United States $140k – $150k

…production Kubernetes environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay … tracing, alerting, and incident response. Strong understanding of Linux systems administration, networking, cloud architecture, and distributed systems fundamentals. Experience designing, implementing, and maintaining CI/CD pipelines and deployment automation. Strong problem-solving skills…

LinkedIn
Filters:
× Clear filters