Senior Site Reliability Engineer (DevTools)

2 months, 2 weeks ago
Full-time
Senior
DevOps and Infrastructure
Nebius

Nebius

Nebius enables B2B companies to build local hyperscaling cloud platforms with cost-effective GPUs, InfiniBand network, and 50% less compute cost. They offer managed Kubernetes and a launch-ready business model for innovative cloud solutions.

Internet Software & Services
51-250

Description

  • Improve services based on user feedback and user problems.
  • Build fault-tolerant, self-healing architecture.
  • Identify ways to speed up systems and reduce user friction.
  • Modify and extend closed-source and open-source solutions, including GitLab and TeamCity plugins.
  • Support users and help resolve their requests and issues.
  • Define metrics that measure user problems and verify that fixes actually resolve them.
  • Work with large-scale build, artifact, and monorepo systems in a production environment.

Requirements

  • Experience combining SRE and software engineering work in roughly a 50/50 split.
  • Experience with Java, Kotlin, Go, Python, and/or Ruby.
  • Understanding of Unix-like systems and the JVM under the hood.
  • Strong focus on improving user experience.
  • Ability to adapt quickly in a fast-changing environment.
  • Experience in Platform Engineering is a plus.
  • Experience operating GitLab or another version control system is a plus.
  • Experience operating TeamCity or another CI system is a plus.
  • Experience with Spring and operating Java monoliths is a plus.
  • Coding interview participation is part of the hiring process.
  • Must be authorized to work in the country of application and provide proof of employment eligibility.

Benefits

  • Competitive compensation.
  • Career growth and learning opportunities.
  • Flexibility and work-life balance.
  • Collaborative and innovative culture.
  • Opportunity to work on impactful AI projects.
  • International environment with talented teams.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Site Reliability Engineer

PENN Entertainment 10K-50K Hotels, Restaurants & Leisure

PENN Entertainment is seeking a Senior Site Reliability Engineer to build and operate the cloud infrastructure supporting large-scale sports betting and media platforms across regulated jurisdictions.

Argo CD AWS Bash CI/CD Datadog GCP GitHub Actions GitOps Go Helm Kubernetes Linux PostgreSQL Python Shell Scripting Terraform
17 hours, 7 minutes ago

Staff Site Reliability Engineer, Ads

Reddit 1K-5K Internet Software & Services

Reddit is hiring a Staff Site Reliability Engineer to provide technical leadership for reliability, scalability, and operational excellence across its advertising infrastructure and revenue-critical systems.

Apache Spark ClickHouse GCP Go Kafka Kubernetes
17 hours, 37 minutes ago

Sr Lead Network Reliability Engineer

Coupa Software 1K-5K Internet Software & Services

Coupa is hiring a Sr. Lead Network Development Engineer to scale and operate its global SaaS platform’s cloud networking infrastructure through automation, reliability engineering, and technical leadership.

Ansible AWS Azure Chef DNS Fortinet Go Java Kubernetes Linux Python Ruby TCP/IP Terraform TLS
1 day, 17 hours ago

Senior Monitoring/Observability Architect

Makpar 51-250 Internet Software & Services

Makpar is seeking a Senior Monitoring/Observability Architect to lead enterprise monitoring strategy, architecture, and implementation guidance for a large federal government program.

Datadog Splunk
1 day, 17 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers