Skip to content
DevOps Jobs
SpaceX

Site Reliability Engineer, AI Infrastructure (Starshield)

SpaceX

Location
Onsite (Redmond, Washington)
Compensation
$125k - $200k/yr
Employment
Full-time
Level
Mid Level
Posted 5 days ago

About the Role

SpaceX is seeking a Site Reliability Engineer to design, operate, and scale GPU and CPU infrastructure for its Starshield national security satellite constellation. This role focuses on managing high-scale AI clusters and ensuring high availability for critical government missions.

Skills

Site Reliability Engineering DevOps Linux Kubernetes Python C++ Go Terraform Ansible GPU infrastructure AI clusters Networking TCP/IP Distributed systems Monitoring CI/CD

Benefits

  • Medical coverage
  • Vision coverage
  • Dental coverage
  • 401(k) plan
  • Disability insurance
  • Life insurance
  • Paid parental leave
  • Paid vacation
  • Paid holidays

Perks

  • Stock options
  • Employee stock purchase plan
  • Company shuttles

Full job details