PandaDoc

PandaDoc

PandaDoc offers an all-in-one document automation platform that enables sales professionals to create, approve, track, and electronically sign proposals and quotes more efficiently, significantly reducing the time spent on business document workflows.

Internet Software & Services
251-1K
Founded 2013
$51M raised

Description

  • Own and influence the incident management process end-to-end.
  • Maintain and evolve the on-prem observability stack and alerting systems.
  • Participate in the on-call rotation to keep production applications running smoothly.
  • Develop automations and tools that support platform reliability.
  • Contribute to production services with performance and resiliency in mind.
  • Collaborate with product engineers to embed SRE principles across the R&D organization.
  • Mentor SRE team members or product engineers.
  • Support proactive prevention of incidents and resolution of performance bottlenecks.

Requirements

  • Solid programming experience with Python (Django and AsyncIO) and/or Java (Spring Boot).
  • Experience maintaining an observability tools suite, specifically LGTM (Loki, Grafana, Tempo, Mimir).
  • Experience developing and maintaining Python services in production.
  • Strong experience with AWS and Kubernetes.
  • Solid proficiency with relational databases, especially PostgreSQL.
  • Experience with messaging systems such as RabbitMQ, NATS, or Kafka.
  • Experience as an on-call SRE engineer.
  • Hands-on troubleshooting experience with distributed systems in production environments.
  • Proficiency in English, both written and spoken.
  • Experience with incident management, capacity planning, and reliability operations is highly relevant.

Benefits

  • Remote-first work with optional hybrid work from offices in Kyiv, Warsaw, and Lisbon.
  • Work schedule aligned with EU time zones.
  • Long-term collaboration options, including employment, employer of record, or B2B arrangements.
  • Honest, open culture that values constructive feedback.
  • Professional and personal development in a collaborative, supportive team.
  • Stable yet growing SaaS product with ownership, startup energy, and strong technical challenges.
  • Work-life balance and kind, globally distributed teammates.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Site Reliability Engineer

PENN Entertainment 10K-50K Hotels, Restaurants & Leisure

PENN Entertainment is seeking a Senior Site Reliability Engineer to build and operate the cloud infrastructure supporting large-scale sports betting and media platforms across regulated jurisdictions.

Argo CD AWS Bash CI/CD Datadog GCP GitHub Actions GitOps Go Helm Kubernetes Linux PostgreSQL Python Shell Scripting Terraform
17 hours, 20 minutes ago

Staff Site Reliability Engineer, Ads

Reddit 1K-5K Internet Software & Services

Reddit is hiring a Staff Site Reliability Engineer to provide technical leadership for reliability, scalability, and operational excellence across its advertising infrastructure and revenue-critical systems.

Apache Spark ClickHouse GCP Go Kafka Kubernetes
17 hours, 51 minutes ago

Sr Lead Network Reliability Engineer

Coupa Software 1K-5K Internet Software & Services

Coupa is hiring a Sr. Lead Network Development Engineer to scale and operate its global SaaS platform’s cloud networking infrastructure through automation, reliability engineering, and technical leadership.

Ansible AWS Azure Chef DNS Fortinet Go Java Kubernetes Linux Python Ruby TCP/IP Terraform TLS
1 day, 17 hours ago

Senior Monitoring/Observability Architect

Makpar 51-250 Internet Software & Services

Makpar is seeking a Senior Monitoring/Observability Architect to lead enterprise monitoring strategy, architecture, and implementation guidance for a large federal government program.

Datadog Splunk
1 day, 17 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers