1 to 25 of 165 OpenTelemetry Jobs in England

Observability Engineer - Assistant Vice President

Location
Greater London, England, United Kingdom
drive the migration of applications from existing monitoring tools (Geneos ITRS, Prometheus, ELK, Splunk, AppDynamics, etc.) to Google Cloud Observability (GCO) and Grafana using OpenTelemetry (OTel) as the instrumentation standard. You will act as a hands‐on technical authority, authoring reusable deployment solutions, configuring telemetry collectors, and providing direct technical … transparency, innovation, and technical excellence that encourages continuous improvement and automation. Collaborative Enablement: Partner with development and SRE teams to drive the adoption of OpenTelemetry (OTel) and Google Cloud Observability (GCO) and Grafana standards. Regulatory Compliance: Operate effectively within a highly regulated environment, ensuring all observability and deployment solutions comply ...

Observability Engineer - Assistant Vice President

Location
Greater London, England, United Kingdom
drive the migration of applications from existing monitoring tools (Geneos ITRS, Prometheus, ELK, Splunk, AppDynamics, etc.) to Google Cloud Observability (GCO) and Grafana using OpenTelemetry (OTel) as the instrumentation standard. You will act as a hands-on technical authority, authoring reusable deployment solutions, configuring telemetry collectors, and providing direct technical … transparency, innovation, and technical excellence that encourages continuous improvement and automation. Collaborative Enablement: Partner with development and SRE teams to drive the adoption of OpenTelemetry (OTel) and Google Cloud Observability (GCO) and Grafana standards. Regulatory Compliance: Operate effectively within a highly regulated environment, ensuring all observability and deployment solutions comply ...

ML Ops Engineer

Location
Greater London, England, United Kingdom
telemetry to track unit economics and throughput for training and serving AI models. Implement end-to-end observability using tools like Prometheus, Grafana, OpenTelemetry, and Weights & Biases or MLflow. Your Skills Hands-on production experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering, with some experience dedicated ...

ML Ops Engineer

Hiring Organisation
Anaplan
Location
London, United Kingdom
Salary
£ 80 K
benchmarking and telemetry to track unit economics and throughput for training and serving AI models.Implement end-to-end observability using tools like Prometheus, Grafana, OpenTelemetry, and Weights & Biases or MLflow.Your SkillsHands-on production experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering, with some experience dedicated to AI/ ...

Senior DevOps Engineer (Azure)

Hiring Organisation
Darktrace
Location
London, United Kingdom
Salary
£ 80 K
supporting technologies like ArgoCD and Helm. You should also be familiar and well-versed in monitoring, logging, and observability tools including: Prometheus; Grafana; Loki; OpenTelemetry; and the ELK stack. Amongst this, you should be able to demonstrate:Proven experience as a DevOps Engineer, with a solid background in building ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Location
Birmingham, England, United Kingdom
specifically building and operating highly resilient cloud-native architectures. Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch) Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures and software ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Location
Birmingham, England, United Kingdom
specifically building and operating highly resilient cloud-native architectures. Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch) Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures and software ...

Strategic DevSecOps Consultant

Hiring Organisation
CloudBees
Location
London, United Kingdom
Salary
£ 80 K
IDPs).Familiarity with AI-enabled software development, agentic workflows, large language models (LLMs), or AI governance practices.Experience with observability and telemetry platforms such as OpenTelemetry, Splunk, Dynatrace, Datadog, AppDynamics, Grafana, or similar technologies.Experience working with large-scale enterprise architecture, governance, compliance, and regulated environments.Thought leadership experience through technical publications, conference ...

Senior Cloud Engineer, AI Platform SRE

Location
Manchester, England, United Kingdom
Built and maintained CI/CD with GitHub Actions, GitLab CI, Argo CD, Jenkins or similar. Observability: Hands‐on with Datadog, Prometheus, Grafana or OpenTelemetry, and opinionated about what's worth alerting on. Incident management: Calm, methodical instincts under pressure, and a habit of fixing the class of problem rather ...

Senior Cloud Engineer, AI Platform SRE

Location
Leeds, England, United Kingdom
Built and maintained CI/CD with GitHub Actions, GitLab CI, Argo CD, Jenkins or similar. Observability: Hands‐on with Datadog, Prometheus, Grafana or OpenTelemetry, and opinionated about what's worth alerting on. Incident management: Calm, methodical instincts under pressure, and a habit of fixing the class of problem rather ...

Senior Cloud Engineer, AI Platform SRE

Location
Greater London, England, United Kingdom
Built and maintained CI/CD with GitHub Actions, GitLab CI, Argo CD, Jenkins or similar. Observability: Hands‐on with Datadog, Prometheus, Grafana or OpenTelemetry, and opinionated about what's worth alerting on. Incident management: Calm, methodical instincts under pressure, and a habit of fixing the class of problem rather ...

Cloud Native Specialist

Location
Greater London, England, United Kingdom
OpenShift) and container orchestration.* Demonstrate how Dynatrace provides automated, code-level visibility into microservices without manual instrumentation or sidecar overhead.* Advocate for OpenTelemetry (OTel) integration and explain how Dynatrace extends the value of open-source telemetry in a production-grade environment.2. Domain Execution:* Lead technical discovery and high-stakes Proof ...

Platform Compute Specialist

Hiring Organisation
SQUAREPOINT CAPITAL
Location
London, United Kingdom
Salary
£ 100 K
storage, networking, IAM)Infrastructure-as-Code (e.g., Terraform, Pulumi)Monitoring and alerting (e.g., Prometheus, Grafana, Datadog, Zabbix)Logging and tracing (e.g., ELK stack, Fluentd, OpenTelemetry, JaegerIdentity and access management (e.g., LDAP, Kerberos, OAuth2)Cloud-native services (e.g., S3, EBS, GKE, EKS, Cloud Functions)Git-based workflows and version controlAgile methodologies ...

Head of Cloud

Location
Norwich, England, United Kingdom
expert), Kubernetes (EKS) IaC & Orchestration: Terraform, Helm, Terragrunt Languages: Go, Node.js, Python (automation/tooling) CI/CD: GitHub Actions, ArgoCD Observability: Prometheus, Grafana, OpenTelemetry What We’re Looking For Proven Leadership: Experience managing and scaling high-performing engineering teams. Cloud Expertise: Deep hands-on experience architecting and operating cloud ...

Site Reliability Engineer - Service Assurance Systems

Location
Greater London, England, United Kingdom
systems at scale. Familiarity with infrastructure-as-code tools such as Terraform or Ansible. Experience with log aggregation and analysis platforms such as the OTEL Stack or AWS CloudWatch Logs Insights. Exposure to Kubernetes or other container orchestration platforms. Experience working in an Agile or DevOps team environment. EEO Statement ...

Site Reliability Engineering (SRE) / Observability Technical Lead

Hiring Organisation
NTT DATA
Location
London, UK
Employment Type
Full-time
roles, with leadership responsibilities. Proven expertise with Application Performance Monitoring (APM) tools such as New Relic, Datadog, AppDynamics, or Dynatrace. Hands-on experience with OpenTelemetry (OTel) for distributed tracing and observability instrumentation. Strong proficiency in Infrastructure as Code (IaC) using Terraform. Solid understanding of cloud platforms including ...

Site Reliability Engineering (SRE) / Observability Technical Lead

Hiring Organisation
NTT
Location
London, United Kingdom
Salary
£ 80 K
DevOps roles, with leadership responsibilities.Proven expertise with Application Performance Monitoring (APM) tools such as New Relic, Datadog, AppDynamics, or Dynatrace.Hands-on experience with OpenTelemetry (OTel) for distributed tracing and observability instrumentation.Strong proficiency in Infrastructure as Code (IaC) using Terraform.Solid understanding of cloud platforms including AWS, GCP, or Azure.Experience with automation ...

Database Platform Engineer

Location
Greater London, England, United Kingdom
services relevant to data platforms such as RDS, Aurora, S3, EC2 or EKS Familiarity with modern observability stacks such as Prometheus, Grafana, Elk or OTel Desirable: experience with cloud‐native and distributed SQL databases such as Aurora, YugabyteDB or TiDB Desirable: knowledge of data streaming and integration tools such ...

Cloud Engineering & Architecture - Senior Platform Engineer AI - Vice President

Location
Greater London, England, United Kingdom
custom agentic loops) to coordinate multi-step diagnostic and remediation tasks. AIOps & Intelligent Observability: Ability to integrate traditional observability stacks (e.g., Datadog, Prometheus, OpenTelemetry) with AI/ML models to automate root-cause analysis, anomaly detection, and semantic log clustering. Self-healing Infrastructure Engineering: Experience designing closed-loop, self-healing ...

Cloud Engineering & Architecture - Senior Platform Engineer AI - Vice President

Location
Greater London, England, United Kingdom
custom agentic loops) to coordinate multi‐step diagnostic and remediation tasks. AIOps & Intelligent Observability: Ability to integrate traditional observability stacks (e.g., Datadog, Prometheus, OpenTelemetry) with AI/ML models to automate root‐cause analysis, anomaly detection, and semantic log clustering. Self‐Healing Infrastructure Engineering: Experience designing closed‐loop, self‐healing ...

Senior Backend Developer (Python) - Remote

Hiring Organisation
IO
Location
Bristol, Gloucestershire, United Kingdom
Employment Type
Contract
Contract Rate
GBP 45 - 50 Hourly
Databricks Microsoft Azure Event-driven architectures (Kafka, MQTT, Azure Event Hubs, etc.) Geospatial or time-series data Observability tools such as Prometheus, Grafana or OpenTelemetry Eligibility Due to project requirements, applicants must be citizens of a NATO member country and currently reside within a NATO member country. Contract Details Fully ...

Lead Software Engineer

Location
City Of London, England, United Kingdom
customer support requirements. Experience with Kafka and event-driven architectures. Experience with OAuth2, OpenID Connect, SAML, and Keycloak. Experience with Datadog, Grafana, Elastic, OpenTelemetry, or similar observability tooling. Experience implementing AI-assisted software development practices. WSD is an employer that values diversity.We highly encourage applications from appropriately qualified and eligible ...

Senior Software Engineer (Infrastructure)

Location
Greater London, England, United Kingdom
Experience developing production‐ready infrastructure management tooling with either Python or Golang Familiarity with at least one of the following: Observability Tools (e.g. Prometheus, OpenTelemetry, Grafana) Databases (e.g. Postgres, DuckDB) Event Streaming platforms (e.g. Kafka) Container Orchestration (e.g. Docker, Kubernetes) Familiarity with cloud platforms such as AWS, Azure ...

Platform Engineer

Hiring Organisation
9fin
Location
London, United Kingdom
Salary
£ 80 K
Good working knowledge of AWS services including ECS, EC2, Lambda, VPC, IAM, Route53, CloudFront, S3, RDSGood understanding of monitoring and logging solutions. We use OpenTelemetry, AWS Cloudwatch and SigNoz so experience with them is a bonus.Basic SRE knowledge, and experience in alerting and incident management platforms (eg. incident.io, Pagerduty)Proven ...

Senior Platform Engineer

Location
Greater London, England, United Kingdom
working knowledge of AWS services including ECS, EC2, Lambda, VPC, IAM, Route53, CloudFront, S3, RDS Good understanding of monitoring and logging solutions. We use OpenTelemetry, AWS Cloudwatch and SigNoz so experience with them is a bonus. Basic SRE knowledge, and experience in alerting and incident management platforms (eg. incident.io, Pagerduty ...