Principal Site Reliability Engineer, Platform

2 days, 10 hours ago
Full-time
Lead
DevOps and Infrastructure
Blue River Technology

Blue River Technology

Blue River Technology serves the agricultural industry with revolutionary computer vision-based robotics, optimizing chemical usage, reimagining routine processes, and improving farming yields while minimizing environmental impact. Through collaboratio...

Industrial Conglomerates
251-1K
Founded 2011

Description

  • Architect, scale, and own essential infrastructure and high-availability services.
  • Build and maintain a Kubernetes-based platform supporting multiple teams and services.
  • Develop Golang backend services and internal tooling for autonomous systems and platform teams.
  • Partner with product teams to launch new products on the platform.
  • Analyze end-to-end performance, identify improvements, and implement robust solutions.
  • Design and maintain observability infrastructure, including dashboards, alerts, and log aggregation.
  • Participate in on-call rotations, resolve production incidents, and document root causes and postmortems.
  • Collaborate with cloud vendors, external technical support, and security teams on upgrades, risk assessments, and threat detection.
  • Maintain the security risk register and implement mitigation plans.
  • Establish and maintain payment processes for SaaS services with IT and purchasing teams.

Requirements

  • 8+ years of experience building and maintaining infrastructure for data-intensive, highly available applications.
  • 6+ years of experience building and maintaining public cloud solutions.
  • Deep knowledge of Kubernetes, Terraform, and cloud orchestration.
  • Strong understanding of software design methodologies, information systems architecture, object-oriented design, and design patterns.
  • Deep experience securing cloud infrastructure, preferably AWS and Kubernetes.
  • Professional experience with Golang preferred; Python, JavaScript, or Rust also applicable.
  • Deep experience with CI/CD tools such as GitHub Actions, ArgoCD, ArgoCD Image Updater, and Artifactory.
  • Interest in robotics, autonomous systems, and the future of automation.
  • Self-directed approach with strong problem-solving ability and enthusiasm for collaborative engineering.

Benefits

  • Base salary range of $174,000–$305,000 per year.
  • Annual performance bonus and competitive benefits package.
  • Remote work within the United States.
  • Visa sponsorship may be available.
  • Career development, mentorship, and learning and development programs.
  • Inclusive workplace with reasonable disability accommodations.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Site Reliability Engineer

Lodgify 251-1K Internet Software & Services

Lodgify, a Barcelona-based vacation-rental technology company, is hiring a Senior Site Reliability Engineer to improve the reliability, scalability, observability, and operational ownership of its cloud platform and critical product services.

Datadog Grafana Kubernetes Microservices Prometheus Python
10 hours, 7 minutes ago

Senior Site Reliability Engineer

PointClickCare 1K-5K Health Care Providers & Services

PointClickCare is seeking a Senior Site Reliability Engineer to provide technical leadership and improve the reliability, automation, observability, and operational efficiency of cloud-based healthcare applications.

Agile Ansible AWS Azure C C++ Chef Docker Go Java Kubernetes Linux Perl Puppet Python Ruby TCP/IP Terraform Windows Server
10 hours, 22 minutes ago

AWS - Incident Handler

Caseware 251-1K Internet Software & Services

Caseware is hiring a fully remote Incident Commander in Colombia to lead incident response for its 24/7 SaaS operations, coordinating resolution, communication, root-cause analysis, and post-incident improvements.

AWS JIRA New Relic PagerDuty
2 days, 9 hours ago

Senior Site Reliability Engineer - Cloud Platform

GoDaddy 5001-10000 Technology, Information and Internet

GoDaddy’s Global Compute team is seeking a remote infrastructure engineer to operate and scale AWS production infrastructure that supports the company’s engineering teams.

AWS AWS CDK CI/CD CloudFormation GitOps Python
2 days, 10 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers