PandaDoc

PandaDoc

PandaDoc offers an all-in-one document automation platform that enables sales professionals to create, approve, track, and electronically sign proposals and quotes more efficiently, significantly reducing the time spent on business document workflows.

Internet Software & Services
251-1K
Founded 2013
$51M raised

Description

  • Own and influence the incident management process end-to-end.
  • Maintain and evolve the on-prem observability stack and alerting systems.
  • Participate in the on-call rotation to keep production applications running smoothly.
  • Develop automations and tools that improve platform reliability.
  • Contribute to production services with a focus on performance and resiliency.
  • Collaborate with product engineers to embed SRE principles across the R&D organization.
  • Mentor SRE team members or product engineers.
  • Support incident investigation and mitigation through monitoring and observability tooling.
  • Help with capacity planning and operational reliability efforts.

Requirements

  • Solid programming experience with Python (Django and AsyncIO) and/or Java (Spring Boot).
  • Experience maintaining an observability tools suite, specifically LGTM: Loki, Grafana, Tempo, and Mimir.
  • Experience developing and maintaining Python services in production.
  • Strong experience with AWS and Kubernetes.
  • Solid proficiency with PostgreSQL and messaging systems such as RabbitMQ, NATS, or Kafka.
  • Experience as an on-call SRE engineer.
  • Hands-on troubleshooting experience with distributed systems in production environments.
  • Proficiency in English, both written and spoken.
  • Ownership mindset and a strong drive to deliver high-quality work.
  • Experience sharing knowledge and communicating about reliability is preferred.

Benefits

  • Remote-first work with the option for hybrid work in Kyiv, Warsaw, and Lisbon.
  • Work schedule aligned with EU time zones.
  • Flexible collaboration options including employment contract, employment of record, or B2B arrangements.
  • Honest, open culture that values constructive feedback.
  • Professional and personal development in a collaborative, supportive team.
  • Stable yet growing SaaS product with strong technical challenges and ownership.
  • Work-life balance and creative virtual team-bonding activities.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Site Reliability engineer

Filevine 251-1K Specialized Consumer Services

Filevine is seeking a Senior Site Reliability Engineer to help build and operate the reliability, observability, and deployment systems behind its Legal AI platform.

AWS Bash CI/CD Go Kubernetes Machine Learning Python
13 hours, 51 minutes ago

Director, Engineering - Infrastructure

Federato 11-50 Insurance

Federato is hiring a Director of Infrastructure to lead the platform, SRE, and Security functions that keep its AI-native insurance workflow system scalable, reliable, secure, and cost-efficient.

2 days, 12 hours ago

Technical Support Engineer (GPU Clusters) - US Weekends

Together 1-10 IT Services

Together AI is hiring a Technical Support Engineer to support customers building and operating AI training, fine-tuning, and inference systems on Kubernetes GPU infrastructure.

Ansible Kubernetes Machine Learning
3 days, 12 hours ago

Staff Site Reliability Engineer

BeyondTrust 1K-5K Professional Services

BeyondTrust is hiring a Staff Site Reliability Engineer to lead the evolution of its Password Safe platform, infrastructure, and deployment ecosystem across cloud and on-premises environments.

Ansible AWS Azure C# CI/CD Datadog DevSecOps Docker GitOps Go Java Kubernetes Linux Microservices OpenTelemetry Secrets Management Terraform
3 days, 12 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers