PB✓
PBridge

Full-time jobsthe United States

Site Reliability Engineer (Raptor)

spacex · Hawthorne, CA · Full-time

About this role

SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.

SITE RELIABILITY ENGINEER (RAPTOR)

SpaceX is looking for a Site Reliability Engineer with a strong drive to solve challenging problems in the Raptor engine organization. You will be empowered to solve a wide range of systems engineering problems – including High Performance Computing, System performance, networking, manufacturing infrastructure – with the singular goal of accelerating the pace of rocket engine development and delivering excellent results. You will work with engineering, analysis, IT, and facilities teams to identify and solve fundamental challenges to scaling the department engineering output.

RESPONSIBILITIES:

• Manage server infrastructure, HPC systems, storage systems, networks, and high-speed interconnect (Infiniband).

• Design, procure, and integrate infrastructure systems.

• Work with propulsion engineering staff to solve critical bottlenecks.

• Coordinate and communicate with company -wide infrastructure and IT teams.

• Support application deployment (ANSYS, StarCCM+, manufacturing software) for best performance of software on real-world systems.

BASIC QUALIFICATIONS:

• 1+ years of hands-on experience with client and server hardware/software, management tools, enterprise networking, virtualization, and security technologies.

• Bachelor's degree in computer science, engineering, math, or scientific discipline; OR 2+ years of professional experience building software in lieu of a degree.

• Experience with Linux and Windows server software.

PREFERRED SKILLS AND EXPERIENCE:

• 1+ year of systems engineering experience

• Experience with scripting languages (Bash, Python), automation (Puppet, Ansible), and other common sysadmin tools.

• Experience building, deploying, and troubleshooting large-scale compute systems.

• Familiarity with resource development and management (Kubernetes, Docker).

• Familiarity with engineering and analysis applications, such as CFD and FEA.

• Familiarity with diagnosing bottlenecks and designing systems for performance.

• Able to work effectively in a dynamic environment while assuming high levels of responsibility and demonstrating accountability for rocket engine-level outcomes.

COMPENSATION AND BENEFITS:

Pay range: Site Reliability Engineer/Level 1: $125,000.00 - $150,000.00/per year Site Reliability Engineer/Level 2: $145,000.00 - $175,000.00/per year

Your actual level and base salary will be determined on a case-by-case basis and may vary based on the following considerations: job-related knowledge and skills, education, and experience.

Base salary is just one part of your total rewards package at SpaceX. You may also be eligible for long-term incentives, in the form of company stock, st

Tired of applying one by one?

Our Career Success Team finds roles in the United States that fit you, tailors your CV to each, and submits the applications — tracked end to end. You just show up to interviews.

We apply, you interview →