Live · read from company career sites daily at 06:00 UTC

Sr. Site Reliability Engineer (Starlink) Verified today

SpaceX · Redmond, WA · Production · first seen 2026-07-29

Role

As a Sr. Site Reliability Engineer on the Starlink program, you will solve challenges that improve how we utilise the hardware deployed across the world's largest satellite constellation. Your goal is to provide customers with the best possible satellite internet experience, often enabling under-served communities to access affordable, life-changing broadband.

You will have full ownership of challenging problems, working with a team of engineers to design and produce solutions that move SpaceX toward its goals at pace. The success of Starlink depends on the software you and your team produce.

What you will do

  • Upgrade existing distributed systems to become sharded and geo-redundant in multiple data centers
  • Advance existing deployment, monitoring, and alerting infrastructure to support a multi-region environment
  • Manage petabyte scale bare metal compute clusters
  • Closely collaborate with engineers across all programs to create highly operable, scalable, and maintainable products
  • Engage throughout the whole software development lifecycle of services - from inception to design, deployment, operation, and iterative refinement
  • Focus on performance bottlenecks and performance improvement techniques

Required qualifications

  • Bachelor's degree in computer science, engineering, math, or scientific discipline and 5 years of software development experience; OR 7+ years of professional experience building software with site reliability or DevOps in lieu of a degree
  • Experience with Linux operating systems

Preferred skills and experience

  • 5+ years of rigorous experience with site reliability or DevOps
  • Experience with Kubernetes and Istio for on-premise deployment
  • Experience with in-stream data processing and analytics using open source platforms such as Apache Kafka, Spark, HBase, HDFS, Flink
  • Experience troubleshooting hardware and network-layer issues
  • Programming experience in Python, C#, Java, Scala, Go or similar languages
  • Good understanding of version control, testing, continuous integration, build, deployment and monitoring

Additional requirements

  • Willing to work extended hours and weekends when needed
  • Must be a U.S. citizen or national, U.S. lawful permanent resident, Refugee under 8 U.S.C. § 1157, or Asylee under 8 U.S.C. § 1158, or be eligible to obtain required authorizations from the U.S. Department of State

Compensation

Level 3: $165,000 - $270,000 base salary. Actual level and base salary determined on a case-by-case basis based on job-related knowledge and skills, education, and experience.

Additional benefits include long-term incentives in the form of company stock or long-term cash awards, discretionary bonuses, Employee Stock Purchase Plan, comprehensive medical, vision, and dental coverage, 401(k) retirement plan, short and long-term disability insurance, life insurance, paid parental leave, 3 weeks paid vacation, 10+ paid holidays per year, and paid sick leave.