Senior Site Reliability Engineer

2 months, 4 weeks ago
Full-time
Senior
DevOps and Infrastructure
Oxylabs

Oxylabs

Oxylabs is a leading provider of premium proxy services with a vast network of residential and datacenter IP proxies. They offer industry-leading web scraping solutions and a trusted ethical proxy network for global clients.

IT Services
251-1K
Founded 2015

Description

  • Own and evolve Webshare's production infrastructure, including the migration from Docker Swarm to Kubernetes or a hybrid Kubernetes + Ansible setup.
  • Maintain high availability across hundreds of servers and approximately 50 services.
  • Drive observability in cooperation with the development team.
  • Establish and enforce Infrastructure as Code practices, CI/CD pipeline reliability, and change management processes.
  • Participate in the on-call rotation alongside backend developers.
  • Respond to incidents, lead resolution efforts, run post-mortems, and drive systematic remediation.
  • Build platform tooling that improves developer experience and reduces infrastructure toil.
  • Keep backend engineers informed and capable through shared infrastructure ownership.

Requirements

  • Experience building and operating highly available infrastructure at comparable scale, including hundreds of servers and dozens of services in production.
  • Hands-on experience with Kubernetes in self-hosted or bare-metal environments.
  • Strong Infrastructure as Code experience.
  • Experience owning CI/CD pipelines end-to-end, such as GitLab CI or an equivalent system.
  • Experience being on call in a production environment.
  • Proactive communication and problem-solving mindset.
  • Scripting and development skills.
  • Experience leading at least one major infrastructure migration, from planning through stabilization, is preferred.
  • Familiarity with Python and/or Go is preferred.
  • Exposure to proxy or networking-heavy infrastructure is preferred.
  • Experience in a small team where developers shared infrastructure responsibility is preferred.
  • Familiarity with edge clusters or split compute/edge architectures is preferred.

Benefits

  • Gross salary of 26,000 PLN to 34,000 PLN per month, with flexibility to discuss a different salary based on skills and experience.
  • 40+ internal learning options, external conferences, mentorship, and year-round knowledge sharing.
  • Private health insurance, a gym allowance, and a wellness app.
  • Team events, an overseas workation, and opportunities to celebrate milestones together.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Cloud Site Reliability Engineer

Cadwell 51-250 Health Care Providers & Services

Cadwell is seeking a Cloud Site Reliability Engineer to operate and improve AWS infrastructure supporting healthcare customers and ensure reliable, secure, and compliant hosted neurodiagnostic software environments.

AWS Bash CI/CD Encryption HIPAA JavaScript JSON Python Terraform TypeScript YAML
1 hour, 48 minutes ago

Site Reliability Engineer

GiveCampus 51-250 Internet Software & Services

GiveCampus is seeking a hands-on Site Reliability Engineer to strengthen the reliability, performance, observability, and operational maturity of its AWS-based fundraising platform in a remote-first U.S. role.

AWS CI/CD CircleCI Datadog GitHub Actions Kubernetes Linux New Relic OpenSearch PostgreSQL Redis Ruby Ruby on Rails Terraform
2 hours, 18 minutes ago

DevOps / SRE / DevSecOps Engineer (AWS) - Latin America, Remote

Bluelight Consulting 11-50 Internet Software & Services

Bluelight is hiring a DevOps/SRE/DevSecOps professional to operate secure, reliable AWS-based systems and delivery infrastructure for client software projects in a remote consultancy environment.

AWS CI/CD DevSecOps Docker K6 OpenTelemetry PostgreSQL Secrets Management Terraform
1 day, 1 hour ago

DevOps / SRE / DevSecOps Engineer (AWS) - Latin America, Remote

Bluelight Consulting 11-50 Internet Software & Services

Bluelight is hiring a DevOps/SRE/DevSecOps professional to support complex client systems by building secure, reliable, and observable AWS infrastructure and delivery operations.

AWS CI/CD DevSecOps Docker K6 OpenTelemetry PostgreSQL Secrets Management Terraform
1 day, 1 hour ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers