Sr. Site Reliability Engineer - Top Secret Clearance (Starlink)

1 hour, 59 minutes ago
Full-time
Senior
DevOps and Infrastructure
SpaceX

SpaceX

SpaceX designs, manufactures, and launches advanced rockets and spacecraft with the aim of revolutionizing space technology and enabling human life on other planets.

Aerospace & Defense
10K-50K
Founded 2002

Description

  • Upgrade distributed systems to be sharded and geo-redundant across multiple data centers.
  • Advance deployment, monitoring, and alerting infrastructure for a multi-region environment.
  • Manage petabyte-scale bare metal compute clusters.
  • Collaborate with engineers across programs to build highly operable, scalable, and maintainable products.
  • Participate in the full software development lifecycle from design through deployment, operation, and refinement.
  • Identify and address performance bottlenecks and improve system performance.

Requirements

  • Bachelor's degree in computer science, engineering, math, or a scientific discipline with 5 years of software development experience, or 7+ years of professional software experience in site reliability or DevOps in lieu of a degree.
  • Experience with Linux operating systems.
  • Active Top Secret or Top Secret SCI clearance.
  • 5+ years of rigorous experience with site reliability or DevOps (preferred).
  • Experience with Kubernetes and Istio for on-premise deployment (preferred).
  • Experience with stream processing and analytics platforms such as Apache Kafka, Spark, HBase, HDFS, and Flink (preferred).
  • Experience troubleshooting hardware and network-layer issues (preferred).
  • Programming experience in Python, C#, Java, Scala, Go, or similar languages (preferred).
  • Good understanding of version control, testing, continuous integration, build, deployment, and monitoring (preferred).
  • Willingness to work extended hours and weekends when needed.
  • Must meet ITAR eligibility requirements as a U.S. citizen/national, lawful permanent resident, refugee, or asylee, or be eligible for required U.S. Department of State authorizations.

Benefits

  • Base salary range of $165,000 to $265,000 for Level 3.
  • Eligibility for long-term incentives, including company stock or long-term cash awards.
  • Potential discretionary bonuses and access to discounted stock purchases through an Employee Stock Purchase Plan.
  • Comprehensive medical, vision, and dental coverage.
  • 401(k) retirement plan.
  • Short- and long-term disability insurance plus life insurance.
  • Paid parental leave.
  • Approximately 3 weeks of paid vacation and 10 or more paid holidays per year.
  • Paid sick leave in accordance with company policy.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

SITE RELIABILITY ENGINEER III

Harford County Public Library 51-250 Diversified Consumer Services

Site Reliability Engineer na Stone, atuando no time de Foundation Platform para fortalecer a plataforma interna de tecnologia com foco em observabilidade, automação e estabilidade dos sistemas.

Ansible Argo CD AWS Azure Datadog Docker GCP GitHub Actions Go Grafana Kubernetes Linux Node.js OpenTelemetry Prometheus Python Splunk Terraform
44 minutes ago

Head of Platform Engineering

dLocal 251-1K Diversified Financial Services

dLocal is seeking a senior leader to own its engineering platform, reliability posture, and AI-assisted development transformation across a global payments business serving emerging markets.

CI/CD Microservices
1 hour, 14 minutes ago

Database Reliability Engineer

Alex Staff Agency 11-50 Professional Services

Senior Database Reliability Engineer for an infrastructure DBA team, responsible for keeping production database services reliable and automating operational work across a multi-database environment.

Ansible ClickHouse DNS Grafana Linux MongoDB OpsGenie PostgreSQL Redis Terraform TLS
1 hour, 44 minutes ago

SRE/DevOps Specialist

iFood 5K-10K Air Freight & Logistics

O iFood busca uma pessoa para atuar com engenharia de confiabilidade e desenvolvimento interno, apoiando a operação de serviços em AWS e Kubernetes e a evolução da maturidade técnica do time.

AWS Datadog Go K6 Kubernetes Python
2 days ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers