Senior Site Reliability Engineer

1 month ago
Full-time
Senior
DevOps and Infrastructure
Oxylabs

Oxylabs

Oxylabs is a leading provider of premium proxy services with a vast network of residential and datacenter IP proxies. They offer industry-leading web scraping solutions and a trusted ethical proxy network for global clients.

IT Services
251-1K
Founded 2015

Description

  • Own and evolve Webshare's production infrastructure, including the migration from Docker Swarm to Kubernetes or a hybrid Kubernetes + Ansible setup.
  • Maintain high availability across hundreds of servers and approximately 50 services.
  • Drive observability in cooperation with the development team.
  • Establish and enforce Infrastructure as Code practices, CI/CD pipeline reliability, and change management processes.
  • Participate in the on-call rotation alongside backend developers.
  • Respond to incidents, lead resolution efforts, run post-mortems, and drive systematic remediation.
  • Build platform tooling that improves developer experience and reduces infrastructure toil.
  • Keep backend engineers informed and capable through shared infrastructure ownership.

Requirements

  • Experience building and operating highly available infrastructure at comparable scale, including hundreds of servers and dozens of services in production.
  • Hands-on experience with Kubernetes in self-hosted or bare-metal environments.
  • Strong Infrastructure as Code experience.
  • Experience owning CI/CD pipelines end-to-end, such as GitLab CI or an equivalent system.
  • Experience being on call in a production environment.
  • Proactive communication and problem-solving mindset.
  • Scripting and development skills.
  • Experience leading at least one major infrastructure migration, from planning through stabilization, is preferred.
  • Familiarity with Python and/or Go is preferred.
  • Exposure to proxy or networking-heavy infrastructure is preferred.
  • Experience in a small team where developers shared infrastructure responsibility is preferred.
  • Familiarity with edge clusters or split compute/edge architectures is preferred.

Benefits

  • Gross salary of 26,000 PLN to 34,000 PLN per month, with flexibility to discuss a different salary based on skills and experience.
  • 40+ internal learning options, external conferences, mentorship, and year-round knowledge sharing.
  • Private health insurance, a gym allowance, and a wellness app.
  • Team events, an overseas workation, and opportunities to celebrate milestones together.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Site Reliability Engineer

Counterpart Health 51-200 hospital & health care

Counterpart Health is hiring a Senior Site Reliability and Infrastructure Engineer to support and evolve the technology platform behind its primary care tool and maintain reliable infrastructure for domestic and international workloads.

AWS Azure CI/CD Containerd DNS Docker GCP Go gRPC Helm Kubernetes Linux Load Balancing Prometheus Python Shell Scripting TCP/IP
1 day, 19 hours ago

Senior Test Platform & Reliability Engineer - Star Trek Fleet Command

Scopely 1K-5K Internet Software & Services

Scopely is hiring a Senior Test Platform & Reliability Engineer in Ireland to build validation, reliability, and developer enablement platforms for Star Trek Fleet Command’s large-scale live-service backend systems.

AWS Bash CI/CD Docker GitLab Go Python Terraform
1 day, 20 hours ago

Senior Software Engineer - Databases, SRE | Canada | Remote

Grafana 1K-5K IT Services

Grafana Labs is hiring a Senior Software Engineer for its remote SRE team to improve reliability and operability of Grafana Cloud database services for high-SLA customers across AWS, GCP, and Azure.

AWS Azure GCP Go Helm Java Kubernetes Linux Microservices Python Terraform
2 days, 19 hours ago

Senior Site Reliability Engineer

Semios 51-250 Food Products

Semios Group is hiring a Senior Site Reliability Engineer to help scale, secure, and improve the reliability of its global agricultural technology platform.

AWS Azure Bash Buildkite CI/CD Datadog Docker Envoy GCP Git GitHub GitHub Actions GitLab Go Jenkins Kubernetes Linux NATS New Relic Prometheus Python Ruby Splunk Terraform
2 days, 20 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers