Skip to content
DevOps Jobs
SpaceX

Sr. Site Reliability Engineer, AI Infrastructure (Starshield)

SpaceX

Location
Onsite (Palo Alto, California)
Compensation
$165k - $265k/yr
Employment
Full-time
Level
Senior Level
Posted 4 days ago

About the Role

SpaceX is seeking a Senior Site Reliability Engineer to design, operate, and scale GPU/CPU infrastructure for AI clusters within Top Secret data centers. This role supports national security missions by ensuring high system availability and developing automation for Kubernetes and on-premise compute resources.

Skills

Linux Kubernetes Python Terraform Ansible GPU infrastructure Site reliability engineering Containerization Bash C++ Go Distributed systems Networking TCP/IP Performance tuning CI/CD

Benefits

  • Medical coverage
  • Vision coverage
  • Dental coverage
  • 401(k) retirement plan
  • Disability insurance
  • Life insurance
  • Paid parental leave
  • Paid vacation
  • Paid holidays
  • Paid sick leave

Perks

  • Company stock
  • Long-term cash awards
  • Discretionary bonuses
  • Employee stock purchase plan

Full job details