Browse DevOps & SRE Jobs

Search 1219 curated tech job listings scraped in real-time from LinkedIn, Glassdoor, RemoteOK, and more. Filter by role, location, seniority, and source to find your next opportunity.

Filters:
× Clear filters

1,219 jobs found for "Prometheus"

Page 1 of 61

DevOps Specialist

SCIENTIFIC GAMES Montreal, Quebec, Canada

…configuration sécurisée. Améliorer la visibilité sur les systèmes à l'aide d'outils tels que Prometheus, Grafana, Alertmanager, ELK/OpenSearch, Splunk et CloudWatch ; gérer les tableaux de bord, les alertes, les guides … stockage, certificats, application de correctifs, renforcement de la sécurité et analyse des performances. Observabilité / Opérations Prometheus, Grafana, Alertmanager, ELK/OpenSearch, Splunk, CloudWatch, logs, métriques, traces, tableaux de bord, guides d'intervention, concepts SLO/SLI…

LinkedIn

404

intellicryst Mobile County, AL

…Swift)✦ PostgreSQL✦ MySQL✦ MongoDB✦ Redis✦ Firebase✦ AWS✦ Google Cloud✦ Docker✦ Kubernetes✦ Linux✦ Terraform✦ Ansible✦ Prometheus✦ Grafana✦ Nginx✦ Apache✦ Load Balancers✦ CDN & Edge✦ Zero Trust Architecture✦ WAF✦ OAuth2 / JWT✦ Penetration Testing✦ Elasticsearch✦ Kibana … Swift)✦ PostgreSQL✦ MySQL✦ MongoDB✦ Redis✦ Firebase✦ AWS✦ Google Cloud✦ Docker✦ Kubernetes✦ Linux✦ Terraform✦ Ansible✦ Prometheus✦ Grafana✦ Nginx✦ Apache✦ Load Balancers✦ CDN & Edge✦ Zero Trust Architecture✦ WAF✦ OAuth2 / JWT✦ Penetration Testing✦ Elasticsearch✦ Kibana…

LinkedIn

DevOps Engineer New

Holidu Munich, Bavaria, Germany $40k+

…Container Orchestration: Kubernetes with Helm Infrastructure as Code: Terraform + Terragrunt, [Pulumi/ CDK](Optional) Monitoring & Observability: Prometheus, Grafana, Elastic Stack, OpenTelemetry CI/CD: Jenkins, GitHub Actions, ArgoCD, ArgoRollouts Scripting: Python, Go, Bash Version Control … across Jenkins, GitHub Actions, ArgoRollouts and ArgoCD. Monitoring & Alerting : Maintain and extend our monitoring stack (Prometheus, Grafana). Build dashboards, configure alerts, and improve observability to ensure comprehensive visibility into system health…

LinkedIn

Senior DevOps / Site Reliability Engineer (SRE) New

Stellar Technologies Abu Dhabi, Abu Dhabi Emirate, United Arab Emirates

…Functions, App Gateway, VNets, Key Vault). Observability & Monitoring – Implement monitoring with Azure Monitor, Grafana, Prometheus, Application Insights and set up custom alerts/dashboards. Secrets Management – Manage and secure secrets via Azure Key Vault … services (Functions, App Gateway, VNets, Key Vault). Proficiency in observability tools – Azure Monitor, Application Insights, Prometheus, Grafana. Solid understanding of Linux, Docker, Kubernetes, and CI/CD workflows. DevOps Tech Stack Category Tools / Technologies…

LinkedIn

Senior DevOps Engineer

thinkbridge Pune City, Maharashtra, India

…with Docker, and orchestrating deployments with Kubernetes (AKS). Skilled in infrastructure as code (ARM/Bicep/Terraform), implementing Prometheus-based monitoring stacks, and enforcing cloud security controls across environments. Strong independent contributor with experience … with Docker, and orchestrating deployments with Kubernetes (AKS). Skilled in infrastructure as code (ARM/Bicep/Terraform), implementing Prometheus-based monitoring stacks, and enforcing cloud security controls across environments. Strong independent contributor with experience…

LinkedIn

AWS DevOps / Observability Consultant

Anagh Technologies Inc Houston, TX

…Staff AWS Observability Consultant with 5+ years of hands-on experience in AWS, Kubernetes, Prometheus, Grafana, OpenTelemetry, CI/CD, and Infrastructure-as-Code . The consultant will act as an embedded AWS SME, delivering … regulated Financial Services environments. Main Responsibilities: Design and implement AWS observability and monitoring using CloudWatch, Prometheus, Grafana, AWS X-Ray, and OpenTelemetry. Manage EKS, ECS, and Fargate monitoring, including Kubernetes observability…

LinkedIn

DevOps Engineer

Quarterhill Inc. Frisco, TX

…deployment. Monitoring & Incident Response Set up and manage monitoring, logging, tracing, and alerting tools (e.g., Prometheus, Grafana, OpenTelemetry, ELK Stack). Proactively monitor systems for performance issues and outages. Participate in a 24/7 … scripting languages (Python, Bash, etc.) for automation. Familiarity with monitoring and logging solutions (e.g., Prometheus, Grafana, ELK Stack). Required Skills/Abilities Cloud Infrastructure Management: Strong experience with cloud platforms such as AWS, Azure…

LinkedIn

Senior DevOps Engineer

thinkbridge Pune Division, Maharashtra, India

…with Docker, and orchestrating deployments with Kubernetes (AKS). Skilled in infrastructure as code (ARM/Bicep/Terraform), implementing Prometheus-based monitoring stacks, and enforcing cloud security controls across environments. Strong independent contributor with experience … with Docker, and orchestrating deployments with Kubernetes (AKS). Skilled in infrastructure as code (ARM/Bicep/Terraform), implementing Prometheus-based monitoring stacks, and enforcing cloud security controls across environments. Strong independent contributor with experience…

LinkedIn

DevOps Engineer

Tata Consultancy Services London Area, United Kingdom

…Kibana) for centralized logging, monitoring, and troubleshooting. Implement and maintain monitoring and alerting solutions using Prometheus and related tools. Administer Linux servers, ensuring platform stability, security, and high availability. Deploy and manage … expertise in Ansible automation and configuration management. Experience managing ELK Stack environments. Strong knowledge of Prometheus monitoring and alerting. Linux administration experience (RHEL/CentOS/OEL). Hands-on experience with Docker and Podman containers. Scripting…

LinkedIn

Senior DevOps Engineer

thinkbridge Pune District, Maharashtra, India

…with Docker, and orchestrating deployments with Kubernetes (AKS). Skilled in infrastructure as code (ARM/Bicep/Terraform), implementing Prometheus-based monitoring stacks, and enforcing cloud security controls across environments. Strong independent contributor with experience … with Docker, and orchestrating deployments with Kubernetes (AKS). Skilled in infrastructure as code (ARM/Bicep/Terraform), implementing Prometheus-based monitoring stacks, and enforcing cloud security controls across environments. Strong independent contributor with experience…

LinkedIn

Senior Java Engineer

ING Hubs Philippines Metro Manila

…production operations, ensuring stable and observable services. Enhance reliability using monitoring and logging tools including Prometheus, Grafana, OpenTracing, ELKaaS. Participate in incident analysis and drive improvements in resilience and operational maturity. Ensure … Capabilities/Experience Good knowledge of observability and monitoring tools like Grafana, Kibana, Loki, Tempo and Prometheus Familiarity with containerization and orchestration tools (e.g., Docker, Kubernetes). Excellent problem-solving skills and ability to work…

LinkedIn

DevOps Engineer

Holidu Munich, Bavaria, Germany $40k+

…Container Orchestration: Kubernetes with Helm Infrastructure as Code: Terraform + Terragrunt, [Pulumi/ CDK](Optional) Monitoring & Observability: Prometheus, Grafana, Elastic Stack, OpenTelemetry CI/CD: Jenkins, GitHub Actions, ArgoCD, ArgoRollouts Scripting: Python, Go, Bash Version Control … across Jenkins, GitHub Actions, ArgoRollouts and ArgoCD. Monitoring & Alerting : Maintain and extend our monitoring stack (Prometheus, Grafana). Build dashboards, configure alerts, and improve observability to ensure comprehensive visibility into system health…

LinkedIn

DevOps Engineer (Customer Experience and Data)

Luminor Group Tallinn, Harjumaa, Estonia €42k – €64k

…leveraging Kubernetes, Kafka, Flink, Spark, Iceberg, AWS services, and observability tooling such as Grafana and Prometheus. Working closely with cross-functional teams, you will champion DevOps best practices, enhance platform performance … practices. Implement and maintain monitoring, alerting, and observability capabilities using tools such as Grafana and Prometheus to ensure platform health, reliability, and performance. Create and maintain dashboards, metrics, and alerts that provide…

LinkedIn

DevOps Engineer (Customer Experience and Data)

Luminor Group Riga, Riga, Latvia €38k – €57k

…leveraging Kubernetes, Kafka, Flink, Spark, Iceberg, AWS services, and observability tooling such as Grafana and Prometheus. Working closely with cross-functional teams, you will champion DevOps best practices, enhance platform performance … practices. Implement and maintain monitoring, alerting, and observability capabilities using tools such as Grafana and Prometheus to ensure platform health, reliability, and performance. Create and maintain dashboards, metrics, and alerts that provide…

LinkedIn

Site Reliability Engineer

NOV Houston, TX

…Observability & Insights Set up and maintain full observability stacks (logging, metrics, tracing) using tools like Prometheus, Grafana, Datadog, OpenTelemetry, or ELK. Analyze telemetry and logs to identify trends, anomalies, and opportunities … frameworks. Proficiency with scripting and automation (Bash, PowerShell, Python). Experience with observability tools (Phobos,Datadog, Prometheus, Grafana, OpenTelemetry, ELK). Hands-on experience with cloud platforms (AWS, Azure, or GCP). Strong PostgreSQL knowledge…

LinkedIn

Senior DevOps Engineer für MARE

Deutsche Telekom Košice, Kosice, Slovakia

…schnell und sicher ausliefern, Docker & Kubernetes für moderne Runtime-Operationen und starkes Monitoring mit Grafana, Prometheus und ELK , damit wir Probleme erkennen, bevor es die User tun. Aufgaben DevOps & Platform Lifecycle Ownership … Anwendungskomponenten. Observability & Reliability Engineering: Du etablierst umfassendes Monitoring, Logging und Alerting mit Tools wie Prometheus, Grafana und zentralisierten Logging-Stacks und stellst hohe Plattformzuverlässigkeit, schnelle Incident-Erkennung und effizientes Troubleshooting sicher. Incident…

LinkedIn

DevOps Engineer für MARE

Deutsche Telekom Košice, Kosice, Slovakia

…GitLab CI -Pipelines, unterstützt containerisierte Workloads mit Docker & Kubernetes und stärkst die Observability mit Grafana, Prometheus und ELK , damit wir Probleme frühzeitig erkennen und beheben können. Du arbeitest eng mit erfahrenen Engineers … Plattformkomponenten zu ermöglichen. Monitoring & Observability: Mitwirkung an Monitoring-, Logging- und Alerting-Setups mit Tools wie Prometheus, Grafana oder ähnlichen Technologien, um Plattform-Transparenz zu gewährleisten und effizientes Troubleshooting zu unterstützen. Incident Support…

LinkedIn

Infrastructure Engineer

The Phoenix Group New York City Metropolitan Area

…network configurations for secure and efficient system operation. Implement monitoring, observability, and alerting solutions (Prometheus, Grafana, ELK Stack, DataDog, or similar) for comprehensive system health tracking and incident response. Collaborate with engineering … tools (Jenkins, GitLab CI, CircleCI, or similar). Proficiency in monitoring, observability, and incident management tools (Prometheus, Grafana, DataDog, ELK, Splunk). Excellent communication skills with the ability to articulate technical options and system…

LinkedIn

Infrastructure Engineer

The Phoenix Group® New York City Metropolitan Area

…network configurations for secure and efficient system operation. Implement monitoring, observability, and alerting solutions (Prometheus, Grafana, ELK Stack, DataDog, or similar) for comprehensive system health tracking and incident response. Collaborate with engineering … tools (Jenkins, GitLab CI, CircleCI, or similar). Proficiency in monitoring, observability, and incident management tools (Prometheus, Grafana, DataDog, ELK, Splunk). Excellent communication skills with the ability to articulate technical options and system…

LinkedIn

DevOps Engineer New

TeamViewer Porto, Porto, Portugal

…operational workflow Assist in the implementation of robust monitoring, logging, and observability frameworks (e.g., Prometheus, Grafana, Datadog) to ensure platform visibility and high service health Proactively identify and resolve performance, security … using tools like Terraform Familiarity with monitoring, logging, and observability platforms such as Grafana, Prometheus, Datadog, or similar technologies Understanding of GitOps and source control management tools like Git backed by scripting…

LinkedIn
Filters:
× Clear filters