PandaDoc

PandaDoc

PandaDoc offers an all-in-one document automation platform that enables sales professionals to create, approve, track, and electronically sign proposals and quotes more efficiently, significantly reducing the time spent on business document workflows.

Internet Software & Services
251-1K
Founded 2013
$51M raised

Description

  • Own and influence the incident management process end-to-end.
  • Maintain and evolve the on-prem observability stack and alerting systems.
  • Participate in the on-call rotation to keep production applications running smoothly.
  • Develop automations and tools that support platform reliability.
  • Contribute to production services with performance and resiliency in mind.
  • Collaborate with product engineers to embed SRE principles across the R&D organization.
  • Mentor SRE team members or product engineers.
  • Support proactive prevention of incidents and resolution of performance bottlenecks.

Requirements

  • Solid programming experience with Python (Django and AsyncIO) and/or Java (Spring Boot).
  • Experience maintaining an observability tools suite, specifically LGTM (Loki, Grafana, Tempo, Mimir).
  • Experience developing and maintaining Python services in production.
  • Strong experience with AWS and Kubernetes.
  • Solid proficiency with relational databases, especially PostgreSQL.
  • Experience with messaging systems such as RabbitMQ, NATS, or Kafka.
  • Experience as an on-call SRE engineer.
  • Hands-on troubleshooting experience with distributed systems in production environments.
  • Proficiency in English, both written and spoken.
  • Experience with incident management, capacity planning, and reliability operations is highly relevant.

Benefits

  • Remote-first work with optional hybrid work from offices in Kyiv, Warsaw, and Lisbon.
  • Work schedule aligned with EU time zones.
  • Long-term collaboration options, including employment, employer of record, or B2B arrangements.
  • Honest, open culture that values constructive feedback.
  • Professional and personal development in a collaborative, supportive team.
  • Stable yet growing SaaS product with ownership, startup energy, and strong technical challenges.
  • Work-life balance and kind, globally distributed teammates.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Site Reliability engineer

Filevine 251-1K Specialized Consumer Services

Filevine is seeking a Senior Site Reliability Engineer to help build and operate the reliability, observability, and deployment systems behind its Legal AI platform.

AWS Bash CI/CD Go Kubernetes Machine Learning Python
13 hours, 58 minutes ago

Director, Engineering - Infrastructure

Federato 11-50 Insurance

Federato is hiring a Director of Infrastructure to lead the platform, SRE, and Security functions that keep its AI-native insurance workflow system scalable, reliable, secure, and cost-efficient.

2 days, 12 hours ago

Technical Support Engineer (GPU Clusters) - US Weekends

Together 1-10 IT Services

Together AI is hiring a Technical Support Engineer to support customers building and operating AI training, fine-tuning, and inference systems on Kubernetes GPU infrastructure.

Ansible Kubernetes Machine Learning
3 days, 12 hours ago

Staff Site Reliability Engineer

BeyondTrust 1K-5K Professional Services

BeyondTrust is hiring a Staff Site Reliability Engineer to lead the evolution of its Password Safe platform, infrastructure, and deployment ecosystem across cloud and on-premises environments.

Ansible AWS Azure C# CI/CD Datadog DevSecOps Docker GitOps Go Java Kubernetes Linux Microservices OpenTelemetry Secrets Management Terraform
3 days, 12 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers