Browse DevOps & SRE Jobs
Search 958 curated tech job listings scraped in real-time from LinkedIn, Glassdoor, RemoteOK, and more. Filter by role, location, seniority, and source to find your next opportunity.
958 jobs found for "Event"
Page 1 of 48
Senior Platform Engineer, Event Streaming Platform (Kafka)
…your life. About the Role We're looking for an experienced Platform Engineer to join Event Streaming team — the distributed team responsible for the Kafka infrastructure and event streaming systems that power … within the broader Storage organisation. The team builds and operates highly reliable, scalable and efficient event streaming infrastructure, including Kafka, in-house Event Streaming abstraction (internally called Event Bus) and any future…
Senior Platform Engineer, Event Streaming Platform (Kafka)
…your life. About the Role We're looking for an experienced Platform Engineer to join Event Streaming team — the distributed team responsible for the Kafka infrastructure and event streaming systems that power … within the broader Storage organisation. The team builds and operates highly reliable, scalable and efficient event streaming infrastructure, including Kafka, in-house Event Streaming abstraction (internally called Event Bus) and any future…
Java Engineer Lead - Capital Markets
…Office and enterprise Capital Markets functions. This role will lead the development of highly scalable, event-driven microservices and distributed systems that enable Trading, Risk Management, Settlements, Collateral Management, Finance, Regulatory Reporting … highly scalable Java-based microservices, APIs, Kafka producers and consumers, streaming applications, and real-time event processing solutions that support trading, risk, settlements, collateral, finance, and regulatory platforms. Establish engineering roadmaps that…
Senior Technical Architect – Hybrid (Product / Platform) New
…Cloud-Native Product Architecture (Primary Focus) Design of microservices architectures on Node.js API-first and event-driven architecture concepts Definition of service boundaries, data flows, and integration patterns Ensuring multi-tenant capability … integration layer between cloud-native products and SAP S/4HANA Private Cloud API and event contracts across the hybrid boundary (OData, REST, SAP Business Events, Event Mesh) Data ownership, consistency, and synchronisation patterns…
DevOps Engineer, Studios
…sales, multi-channel content production and distribution, brand partnerships, strategic consulting, digital services, and events management. It powers growth of revenues, fanbases and IP for more than 200 federations, associations, events … reach 1 billion households across 210 countries and territories and organize more than 500 live events year-round, attracting more than three million fans. TKO also services and partners with major sports…
Python with AKS
…hands-on experience with REST API development, API lifecycle management via APIM, and exposure to event-driven or service-oriented architectures. Familiarity with CI/CD pipelines, container orchestration, monitoring, and troubleshooting in cloud … hands-on experience with REST API development, API lifecycle management via APIM, and exposure to event-driven or service-oriented architectures. Familiarity with CI/CD pipelines, container orchestration, monitoring, and troubleshooting in cloud…
Sr Software Engineer - Python / DevOps / Message Brokers
…scalable Python services for cloud and IoT-scale systems on Google Cloud Platform (GCP). APIs & Event-Driven Systems: Build REST/gRPC APIs and pub/sub architectures using message brokers such as MQTT, Kafka, RabbitMQ … Messaging: Enable secure device-to-cloud and service-to-service communication using MQTT and event-driven patterns. Automate & Deploy: Build and operate automated CI/CD pipelines and cloud-native deployments using Docker…
Site Reliability Engineer
…drive measurable results in brick-and-mortar. Under the hood, that means high-volume event ingestion, real-time decisioning, and infrastructure that has to be both reliable and low-latency to influence … heavy concurrent load from both consumer traffic and agent-driven workflows Build and maintain our event-driven GCP infrastructure, including Cloud Functions, Pub/Sub, Workflows, and GKE/Kubernetes deployments Investigate performance inconsistencies (e.g., identical…
Site Reliability Engineer
…manage applications using Helm and infrastructure automation tools. Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. Define, monitor, and maintain … environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault…
Site Reliability Engineer
…manage applications using Helm and infrastructure automation tools. Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. Define, monitor, and maintain … environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault…
Site Reliability Engineer
…manage applications using Helm and infrastructure automation tools. Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. Define, monitor, and maintain … environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault…
Site Reliability Engineer
…manage applications using Helm and infrastructure automation tools. Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. Define, monitor, and maintain … environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault…
Site Reliability Engineer
…manage applications using Helm and infrastructure automation tools. Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. Define, monitor, and maintain … environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault…
Site Reliability Engineer
…manage applications using Helm and infrastructure automation tools. Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. Define, monitor, and maintain … environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault…
Site Reliability Engineer
…manage applications using Helm and infrastructure automation tools. Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. Define, monitor, and maintain … environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault…
Site Reliability Engineer
…manage applications using Helm and infrastructure automation tools. Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. Define, monitor, and maintain … environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault…
Site Reliability Engineer
…manage applications using Helm and infrastructure automation tools. Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. Define, monitor, and maintain … environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault…
Site Reliability Engineer
…manage applications using Helm and infrastructure automation tools. Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. Define, monitor, and maintain … environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault…
Site Reliability Engineer
…manage applications using Helm and infrastructure automation tools. Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. Define, monitor, and maintain … environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault…
Site Reliability Engineer
…Cloudwatch Application Environment: Core Platform: Java Spring Boot microservices for betting, trading, and customer management Event Processing: High-throughput event queues with latency monitoring (Kakfa, JMS) Data Pipeline: Avro-based S3 buffer … tracing are actionable for detecting and fixing production issues Implement chaos engineering experiments targeting our event queue systems and high-availability betting services Design and execute pre-deployment readiness checks and post…