1 to 25 of 94 OpenTelemetry Jobs in England

Observability Engineer - Assistant Vice President

Location
Greater London, England, United Kingdom
drive the migration of applications from existing monitoring tools (Geneos ITRS, Prometheus, ELK, Splunk, AppDynamics, etc.) to Google Cloud Observability (GCO) and Grafana using OpenTelemetry (OTel) as the instrumentation standard. You will act as a hands‐on technical authority, authoring reusable deployment solutions, configuring telemetry collectors, and providing direct technical … transparency, innovation, and technical excellence that encourages continuous improvement and automation. Collaborative Enablement: Partner with development and SRE teams to drive the adoption of OpenTelemetry (OTel) and Google Cloud Observability (GCO) and Grafana standards. Regulatory Compliance: Operate effectively within a highly regulated environment, ensuring all observability and deployment solutions comply ...

Observability Engineer - Assistant Vice President

Location
Greater London, England, United Kingdom
drive the migration of applications from existing monitoring tools (Geneos ITRS, Prometheus, ELK, Splunk, AppDynamics, etc.) to Google Cloud Observability (GCO) and Grafana using OpenTelemetry (OTel) as the instrumentation standard. You will act as a hands-on technical authority, authoring reusable deployment solutions, configuring telemetry collectors, and providing direct technical … transparency, innovation, and technical excellence that encourages continuous improvement and automation. Collaborative Enablement: Partner with development and SRE teams to drive the adoption of OpenTelemetry (OTel) and Google Cloud Observability (GCO) and Grafana standards. Regulatory Compliance: Operate effectively within a highly regulated environment, ensuring all observability and deployment solutions comply ...

ML Ops Engineer

Location
Greater London, England, United Kingdom
telemetry to track unit economics and throughput for training and serving AI models. Implement end-to-end observability using tools like Prometheus, Grafana, OpenTelemetry, and Weights & Biases or MLflow. Your Skills Hands-on production experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering, with some experience dedicated ...

Software Engineering Tech Lead (SRE + AI)

Location
Greater London, England, United Kingdom
agents, MCP tool integrations, and deterministic evaluation pipelines for automated operational decision support. Telemetry & Insights: Architect ingestion and correlation pipelines across distributed logs, metrics, OpenTelemetry traces, change events, and runbooks to accelerate Mean Time to Detection (MTTD) and Resolution (MTTR). Safe Production Automation: Develop proactive anomaly detection and Human … Hands‐on experience building LLM pipelines, AI Agents, Model Context Protocol (MCP) servers/clients, RAG architectures, and evaluation frameworks. Observability & Telemetry: Experience with OpenTelemetry (OTel), Prometheus, Grafana, Splunk, ThousandEyes, or distributed tracing systems. Cloud & Infrastructure: Expertise in public cloud providers (AWS, GCP, Azure), Terraform/IaC, and GitOps/ ...

Cloud Native Specialist

Location
Greater London, England, United Kingdom
OpenShift) and container orchestration.* Demonstrate how Dynatrace provides automated, code-level visibility into microservices without manual instrumentation or sidecar overhead.* Advocate for OpenTelemetry (OTel) integration and explain how Dynatrace extends the value of open-source telemetry in a production-grade environment.2. Domain Execution:* Lead technical discovery and high-stakes Proof ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Location
Birmingham, England, United Kingdom
specifically building and operating highly resilient cloud-native architectures. Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch) Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures and software ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President

Location
Birmingham, England, United Kingdom
specifically building and operating highly resilient cloud‐native architectures. Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch) Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures and software ...

Senior Cloud Engineer, AI Platform SRE

Location
Leeds, England, United Kingdom
Built and maintained CI/CD with GitHub Actions, GitLab CI, Argo CD, Jenkins or similar. Observability: Hands‐on with Datadog, Prometheus, Grafana or OpenTelemetry, and opinionated about what's worth alerting on. Incident management: Calm, methodical instincts under pressure, and a habit of fixing the class of problem rather ...

Senior Cloud Engineer, AI Platform SRE

Location
Manchester, England, United Kingdom
Built and maintained CI/CD with GitHub Actions, GitLab CI, Argo CD, Jenkins or similar. Observability: Hands‐on with Datadog, Prometheus, Grafana or OpenTelemetry, and opinionated about what's worth alerting on. Incident management: Calm, methodical instincts under pressure, and a habit of fixing the class of problem rather ...

Senior Cloud Engineer, AI Platform SRE

Location
Greater London, England, United Kingdom
Built and maintained CI/CD with GitHub Actions, GitLab CI, Argo CD, Jenkins or similar. Observability: Hands‐on with Datadog, Prometheus, Grafana or OpenTelemetry, and opinionated about what's worth alerting on. Incident management: Calm, methodical instincts under pressure, and a habit of fixing the class of problem rather ...

Head of Cloud

Location
Norwich, England, United Kingdom
expert), Kubernetes (EKS) IaC & Orchestration: Terraform, Helm, Terragrunt Languages: Go, Node.js, Python (automation/tooling) CI/CD: GitHub Actions, ArgoCD Observability: Prometheus, Grafana, OpenTelemetry What We’re Looking For Proven Leadership: Experience managing and scaling high-performing engineering teams. Cloud Expertise: Deep hands-on experience architecting and operating cloud ...

Site Reliability Engineer - Service Assurance Systems

Location
Greater London, England, United Kingdom
systems at scale. Familiarity with infrastructure-as-code tools such as Terraform or Ansible. Experience with log aggregation and analysis platforms such as the OTEL Stack or AWS CloudWatch Logs Insights. Exposure to Kubernetes or other container orchestration platforms. Experience working in an Agile or DevOps team environment. EEO Statement ...

Staff Engineer - Money, Risk & Payment Ancillaries Yuno Totalmente remoto · Mundial ayer

Location
Greater London, England, United Kingdom
gRPC, REST Frameworks — Spring Boot, Spring WebFlux; Go standard library Messaging — Apache Kafka, SQS Databases — PostgreSQL, Redis Infrastructure — AWS, Kubernetes, Docker, Terraform Observability — Datadog, OpenTelemetry CI/CD — GitHub Actions, ArgoCD Version Control — Git/GitHub What We Offer at Yuno Competitive Compensation Remote Work — you can work from everywhere ...

Cloud Engineering & Architecture - Senior Platform Engineer AI - Vice President

Location
Greater London, England, United Kingdom
custom agentic loops) to coordinate multi-step diagnostic and remediation tasks. AIOps & Intelligent Observability: Ability to integrate traditional observability stacks (e.g., Datadog, Prometheus, OpenTelemetry) with AI/ML models to automate root-cause analysis, anomaly detection, and semantic log clustering. Self-healing Infrastructure Engineering: Experience designing closed-loop, self-healing ...

Senior Forward Deployment Engineer

Location
Greater London, England, United Kingdom
failover, and production-recovery exercises, highlighting skills in system reliability and continuity planning. - Experience with enterprise observability tools such as Splunk, ELK, Grafana, Prometheus, OpenTelemetry, AppDynamics, or Dynatrace, reflecting proficiency in monitoring and diagnostics. - Experience modernizing monolithic or legacy enterprise applications into maintain #J-18808-Ljbffr ...

Engineer C# (Full Stack)

Location
City Of London, England, United Kingdom
with microservices and event‐driven architectures. Experience with AI‐assisted design and coding. Knowledge of GraphQL and WebSockets. Familiarity with observability tools such as OpenTelemetry and Grafana. Experience with Infrastructure as Code (Terraform). Understanding of financial markets or trading systems. Contribution to open‐source projects. Awareness of security principles ...

QA Engineer

Hiring Organisation
Moorepay
Location
Manchester, Lancashire, United Kingdom
Employment Type
Full-Time
Salary
Competitive salary
with performance and load testing tools (e.g. JMeter, k6, Locust). Understanding of contract testing (e.g., Pact). Familiarity with observability tools (e.g., Grafana, OpenTelemetry). ...

Senior DevOps Engineer

Hiring Organisation
Hackajob Ltd
Location
Leicester, Leicestershire, East Midlands, United Kingdom
Employment Type
Permanent
Salary
£70,000
scripting Networking skills/fundamentals? SQL Server/NoSQL databases Windows & Linux servers Source Control Management (Git) Docker containers Kubernetes Microservices Monitoring tool - Dynatrace, OTel, Grafana Docker Compose/Helm charts DevSecOps Tooling - SonarCloud/PrismaCloud/CrowdStrike ...

Engineer C# (Full Stack)

Location
Greater London, England, United Kingdom
collaboration skills.Desired* Experience with microservices and event-driven architectures.* Experience with AI assisted design and coding.* GraphQL, and WebSockets.* Knowledge of observability tools (e.g., OpenTelemetry, Grafana).* Familiarity with Infrastructure as Code (Terraform).* Understanding of financial markets or trading systems.* Contribution to open-source projects.* Awareness of security principles ...

AI DevOps Engineer

Location
Greater London, England, United Kingdom
direction). Infrastructure as Code — Terraform. Cloud platform — AWS (ECS/EKS, Lambda, S3, IAM, networking); container orchestration with Docker + Kubernetes. Observability & monitoring — OpenTelemetry, CloudWatch, logging/tracing for AI workloads. Scripting/automation — Python and/or Bash. Security & compliance operations — secrets management, IAM/least privilege ...

Lead Observability Engineer / Senior Software Engineer

Location
Nottingham, England, United Kingdom
based culture, so you’ll have plenty of opportunity to talk, coach, and learn with many great and diverse individuals. Observability & Telemetry tools , including OTel, Grafana, and Prometheus Good understanding of programming languages used for high-performance engineering such as Golang and Java AWS , leveraging cloud-based services such ...

Principal Site Reliability Engineer

Location
Greater London, England, United Kingdom
GitLab Pipelines, ArgoCD, Octopus Deploy Data - ElasticSearch hosted with Kubernetes Operator, PostgreSQL, SQL Server, BigQuery Monitoring and Security - Splunk, Grafana/Grafana Tempo, OpenTelemetry, Cloud Armor Enterprise, OpsGenie, Renovate, Sentry AI Tools - Claude, Amazon Bedrock, Gemini, Vertex AI Key Responsibilities Design, implement, and operate scalable, reliable, and secure infrastructure across ...

Senior Platform Engineer IRC296090

Location
Greater London, England, United Kingdom
systems (Helm, Terraform modules) Background in developer experience research — understanding how engineers consume platform tooling and designing for adoption Experience with observability and monitoring (OpenTelemetry, Grafana, Datadog) — particularly instrumenting developer workflows Experience in financial services or similarly regulated environments Job responsibilities Design and build reusable CI/CD templates, pipeline ...

Backend Python Engineer

Hiring Organisation
Lorien
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£450.00 per day
Strong Python development experience (FastAPI, Pydantic, SQLAlchemy, Pytest). Experience with PostgreSQL, Docker, Git/GitHub, GitHub Actions, and microservices. Knowledge of Pandas, NumPy, OpenTelemetry, and cloud-native development. Strong communication, problem-solving, and stakeholder management skills. Ability to take ownership and thrive in a fast-moving environment. Desirable ...

MLOps Engineer

Hiring Organisation
DGH Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
automation, and DevOps best practices - Experience with Kubernetes/containerisation technologies - Strong programming skills (e.g. Python, Go, Node.js) - Experience with observability tools (e.g. OpenTelemetry, Datadog) - Understanding of security, performance optimisation, and scalability Desirable Skills - Experience working on AI/ML platforms or deployments - Exposure to large-scale distributed systems ...