SpaceX

SpaceX

SpaceX designs, manufactures, and launches advanced rockets and spacecraft with the aim of revolutionizing space technology and enabling human life on other planets.

Aerospace & Defense
10K-50K
Founded 2002

Description

  • Install, manage, scale, and optimize Kubernetes and RKE clusters in production environments using Ansible, Terraform, and related tools.
  • Collaborate with SpaceX engineers to gather requirements, design, plan, deploy, and support software platforms running in Kubernetes.
  • Build highly resilient, high-performance, scalable, and robust systems that support demanding engineering teams.
  • Make and implement improvements through an accepted change control process.
  • Design and deliver solutions with internal business units while resolving problems proactively and in a timely manner.
  • Define, document, and follow standards and best practices for systems design, testing, and implementation.
  • Promote collaboration and cross-training to build Kubernetes expertise across the team.
  • Develop scripting, self-service, and automation to reduce administrative overhead and TOIL.
  • Participate in on-call rotation for urgent after-hours support when needed.

Requirements

  • Bachelor’s degree in Computer Science or a STEM discipline with 5+ years of systems engineering experience, or 7+ years of systems engineering experience in lieu of a degree.
  • Experience deploying and supporting Linux servers in physical and virtualized environments, including VMware via automation.
  • Experience with the Linux shell and configuring/extending Linux instances, including kernel modules, cgroups, PKI, iptables, and network interfaces.
  • Experience supporting and scaling containerized applications in Linux environments.
  • Experience using automation frameworks such as Ansible and Terraform to manage infrastructure provisioning and Kubernetes lifecycles.
  • Strong understanding of Linux container runtime.
  • Experience with source control tools such as Git and Subversion, including pull requests and Git-based workflows.
  • Experience with Infrastructure as Code, CI/CD, and GitOps tools such as Ansible, AWX/Tower, Vagrant, Puppet, Redfish, Jenkins, cloud-init, and ArgoCD.
  • Experience writing test automation for backwards compatibility of automation processes and Kubernetes deployments.
  • Experience with Python and Golang for automation and RESTful API integration.
  • Experience troubleshooting Kubernetes internals and related components such as CNI, CRI, CSI, Docker, Cri-O, Ceph, Cilium, MetalLB, Istio, and rook-ceph.
  • Experience building custom Kubernetes solutions using patterns such as webhooks, controllers, operators, and sidecars.
  • Experience creating monitoring and alerting dashboards with tools such as Prometheus, Grafana, and InfluxDB.
  • Experience with dynamic configuration templating using Jinja, Jsonnet, YAML, and Helm.
  • Must be willing to work extended hours and weekends as needed.
  • Must meet ITAR eligibility requirements as a U.S. citizen/national, lawful permanent resident, refugee, asylee, or otherwise eligible for required U.S. Department of State authorization.

Benefits

  • Senior Kubernetes Engineer base salary range of $160,000 to $225,000 per year.
  • Eligibility for long-term incentives, including company stock, stock options, or long-term cash awards.
  • Potential discretionary bonuses and access to an Employee Stock Purchase Plan.
  • Comprehensive medical, vision, and dental coverage.
  • 401(k) retirement plan.
  • Short- and long-term disability insurance, plus life insurance.
  • Paid parental leave.
  • Accrual of 3 weeks of paid vacation and 10 or more paid holidays per year.
  • Paid sick leave in accordance with company policy.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

ML Infrastructure Engineer

x.ai 51-250 Internet Software & Services

SpaceXAI is hiring an ML Infrastructure Engineer to build and optimize the machine learning platform that powers recommendations on X.

Ansible C++ Linux Puppet Python PyTorch Rust
16 hours, 16 minutes ago

Staff Network Engineer (AI Fabric, Datacenter and Edge Networking) - Radian Arc (EMEA)

Submer 51-250 IT Services

Radian Arc is hiring a Staff Network Engineer to design and operate the networking infrastructure behind its GPU cloud platform for AI, machine learning, and cloud gaming workloads in telecom carrier networks.

Bash Python PyTorch TensorFlow WAF
17 hours, 1 minute ago

Cloud Platform Architect (AI Infrastructure Environments)

Gramian Consultancy Group Professional Services

Gramian Consultancy is hiring a fully remote Cloud Platform Architect to build realistic cloud engineering environments used to evaluate how advanced AI systems solve production infrastructure problems.

AWS Azure GCP Kubernetes Terraform
17 hours, 1 minute ago

Engineering Manager I, Applications

StackAdapt 1K-5K Media

StackAdapt is hiring an Engineering Manager to lead the EngOps Applications team, owning delivery and operational health for infrastructure work that supports critical production systems.

AWS Kubernetes Linux Terraform
17 hours, 1 minute ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers