About the Role
Title: Principal Observability DevOps Engineer
Location: Bulgaria (Remote)
Job Description:
LivePerson (NASDAQ: LPSN) is the global leader in enterprise conversations. Hundreds of the world’s leading brands – including HSBC, Chipotle, and Virgin Media – use our award-winning Conversational Cloud platform to connect with millions of consumers. We power nearly a billion conversational interactions every month, providing a uniquely rich data set and safety tools to unlock the power of Conversational AI for better customer experiences.
At LivePerson, we foster an inclusive workplace culture that encourages meaningful connection, collaboration, and innovation. Everyone is invited to ask questions, actively seek new ways to achieve success and reach their full potential. We are continually looking for ways to improve our products and make things better. This means spotting opportunities, solving ambiguities, and seeking effective solutions to the problems our customers care about.
Overview:
The Observability Platform team is building a state of the art system for logging, motoring, and tracing across cloud and on-prem data centers. We’re looking for an experienced Principal DevOps Lead to head our Logging and Monitoring, ensuring robust, scalable solutions within our Google Cloud Platform. In this role, you will be helping to bring systems to life that give superpowers to an entire organization of software developers.
You will:
- Lead the design, implementation, and maintenance of our observability infrastructure.which manages trillions of observability events (logs, traces, metrics) per day.
- Develop and manage monitoring, logging, and alerting systems using GrafanaLab, CaptainHook, Zabbix, fluentd, filebeat ,ELK, Kafka, Prometheus, OpenTelemetry and related technologies.
- Provide strategic direction and leadership to the Logging and Monitoring team
- Design and develop parts of a highly scalable software observability platform which manages trillions of observability events (logs, traces, metrics) per day.
- Develop and maintain Kubernetes Helm charts that deploy hundreds of pods across nodes every day.
- Collaborate closely with DevOps teams in delivering cloud solutions aligned with our observability platform.
- Ensure high availability and performance of observability platforms and tools.
- Partner with vendors, and stay up to date with emerging technologies around observability, improving your own skills and those of the team around you.
- Design and develop end-to-end Synthetic Tests Monitoring solutions on GCP. with self-service capabilities for engineering teams.
- Located in Bulgaria
You have:
- Bachelor’s degree in Computer Science, Engineering, or related work experience.
- 6+ years as DevOps Engineer (or equal role) with a passion for technology and strong motivation and responsibility for high reliability and service level
- Proficient in Kubernetes and containerization technologies (Docker, etc.)
- Extensive experience with observability tools such as GrafanaLab, CaptainHook, Zabbix, Fluentd, ELK, Kafka, and Prometheus.
- Familiarity with infrastructure as code (IaC) tools like Terraform, Ansible, or CloudFormation.
- Experience with cloud platforms (AWS, Azure, GCP) and their services related to computing, storage, and networking – preferred GCP.
- Strong programming skills in one or more languages (Bash, Python, Go, etc.).
- The ideal candidate will have experience with OpenTelemetry Collector and Grafana Agent.