Entry Level
Sponsorship restricted
Site Reliability Engineer, AI Infrastructure (Starshield)
SpaceX · Hawthorne, CA + 3 more
About this role
Manage GPU/CPU infrastructure deployments to Top Secret data centers Manage and provide support for GPU as a service for external customers on bare metal hardware and virtualized platforms Design, validate, and productize solutions for AI clusters (100k+ GPU scale) Develop…
Sign up free to open the full description on SpaceX's careers page
Skills mentioned
PythonC++GoKubernetesTerraformLinux
Also posted as
- Site Reliability Engineer, AI Infrastructure (Starshield) — Washington, DC
- Site Reliability Engineer, AI Infrastructure (Starshield) — Redmond, WA
- Site Reliability Engineer, AI Infrastructure (Starshield) — Palo Alto, CA