Back to all jobs
SpaceX logo

Application Software Engineer, Inference

SpaceX
Palo Alto1d ago

About the role

<div class="content-intro"><p>SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of&nbsp;enabling human life on Mars.</p></div><p><strong>APPLICATION SOFTWARE ENGINEER, INFERENCE</strong></p> <p>The application software team is the central nervous system of SpaceX – we create mission critical applications that are used throughout SpaceX to accelerate launch vehicle production and flight as well as systems that allow Starlink to grow into a worldwide fast, reliable Internet service. We are looking for engineers who treat fellow teammates with fairness, respect, and support.</p> <p>Our team maintains a high-performance AI inference platform that serves the best models internally at SpaceX to accelerate our most ambitious engineering goals. As part of this effort in Palo Alto, you will design and optimize large-scale model serving systems end-to-end, owning everything from distributed infrastructure to deep low-level optimizations. You will work on systems that deliver reliable, high-throughput inference to power SpaceX’s mission-critical applications while maintaining the highest standards of performance and availability.</p> <p>Aerospace experience is not required to be successful here - rather we look for smart, motivated, respectful, collaborative engineers who love solving problems and want to make an impact on a super inspiring mission. You will have full ownership of challenging problems, working with a team of enthusiastic engineers with diverse perspectives to design and produce solutions that enable SpaceX to achieve its loftiest engineering goals at a rapid pace. The success of the missions at SpaceX depends on the software that you and your team produce.</p> <p>This role will report through SpaceX Application Software while also working closely with xAI engineering teams.&nbsp;</p> <p><strong>RESPONSIBILITIES:</strong></p> <ul> <li>Develop highly reliable, high-throughput inference systems that serve the best AI models internally across SpaceX</li> <li><span class="TextRun SCXW71236998 BCX0" lang="EN-US" data-contrast="auto"><span class="NormalTextRun SCXW71236998 BCX0" data-ccp-parastyle="List Bullet">Architect and implement scalable distributed infrastructure for model serving, including load balancing, auto-scaling, batch scheduling, global KV cache, and continuous batching</span></span><span class="EOP SCXW71236998 BCX0" data-ccp-props="{}">&nbsp;</span></li> <li>Optimize latency and throughput of model inference under real production workloads, including low-level GPU kernel work, quantization, speculative decoding, and other acceleration techniques<span data-ccp-props="{}">&nbsp;</span></li> <li>Build reliable, high-concurrency serving systems with 100% uptime, low tail latency, and excellent observability<span data-ccp-props="{}">&nbsp;</span></li> <li>Own end-to-end components such as request routing, SDK development, rate limiting, and efficient scaling for internal SpaceX AI inference platforms<span data-ccp-props="{}">&nbsp;</span></li> <li>Benchmark, fine-tune, and accelerate inference engines (e.g., SGLang, vLLM, TensorRT-LLM)<span data-ccp-props="{}">&nbsp;</span></li> <li>Develop custom tools for tracing, replaying, and resolving issues across the full stack — from orchestration down to GPU kernels<span data-ccp-props="{}">&nbsp;</span></li> <li>Create robust CI/CD infrastructure for seamless endpoint deployment, image publishing, and inference engine updates<span data-ccp-props="{}">&nbsp;</span></li> <li>Collaborate across SpaceXAI<span data-ccp-parastyle="List Bullet">&nbsp;</span><span data-ccp-parastyle="List Bullet">teams to integrate inference capabilities into broader systems and workflows</span><span data-ccp-props="{}">&nbsp;</span></li> </ul> <p><strong>BASIC QUALIFICATIONS:</strong></p> <ul> <li>Bachelor's degree in computer science, engineering, math, or scientific discipline; OR 2+ years of professional experience building software in lieu of a degree</li> <li>Experience in designing, implementing, and maintaining reliable and horizontally scalable distributed systems</li> <li>1+ years of experience in full stack development or backend development with production systems</li> <li>1+ years of experience with Rust or C++</li> </ul> <p><strong>PREFERRED SKILLS AND EXPERIENCE:</strong></p> <ul> <li>Experience with LLM inference engines and serving frameworks (e.g., SGLang, vLLM, Triton, TensorRT-LLM)<span data-ccp-props="{}">&nbsp;</span></li> <li>Deep low-level systems programming and optimizations: GPU kernels, code generation, batching, caching, parallelism, quantization, and speculative decoding<span data-ccp-props="{}">&nbsp;</span></li> <li>Experience with large-scale, high-concurrency production serving systems<span data-ccp-props="{}">&nbsp;</span></li> <li>Knowledge of service observability and reliability best practices<span data-ccp-props="{}">&nbsp;</span></li> <li>Experience operating commonly used databases such as PostgreSQL, ClickHouse, or MongoDB<span data-ccp-props="{}">&nbsp;</span></li> <li>Experience designing or building with agent SDKs and agent orchestration frameworks<span data-ccp-props="{}">&nbsp;</span></li> <li>Experience with Docker, Kubernetes, and containerized applications<span data-ccp-props="{}">&nbsp;</span></li> <li>Expert knowledge of gRPC (unary, response streaming, bi-directional streaming, REST mapping)<span data-ccp-props="{}">&nbsp;</span></li> <li>Programming experience in Python, Go, or similar languages<span data-ccp-props="{}">&nbsp;</span></li> <li>Experience with version control, continuous integration, continuous delivery, build systems, and monitoring<span data-ccp-props="{}">&nbsp;</span></li> <li>Expertise in profiling and improving application performance<span data-ccp-props="{}">&nbsp;</span></li> </ul> <p><strong>ADDITIONAL REQUIREMENTS:</strong></p> <ul> <li>You may be asked to work extended hours/weekends dependent on launch cadence and platform demands<span data-ccp-props="{}">&nbsp;</span></li> <li>This role requires you to be onsite in Palo Alto. Remote and/or hybrid work will not be considered<span data-ccp-props="{}">&nbsp;</span></li> </ul> <p><strong>COMPENSATION AND BENEFITS:</strong><br>&nbsp;<br>Pay Range:<br>Software Engineer/Level I: $135,000.00 - $160,000.00/per year<br>Software Engineer/Level II: $155,000.00 - $185,000.00/per year</p> <p>Your actual level and base salary will be determined on a case-by-case basis and may vary based on the following considerations: job-related knowledge and skills, education, and experience.</p> <p>Base salary is just one part of your total rewards package at SpaceX. You may also be eligible for long-term incentives, in the form of company stock, stock options, or long-term cash awards, as well as potential discretionary bonuses and the ability to purchase additional stock at a discount through an Employee Stock Purchase Plan. You will also receive access to comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short and long-term disability insurance, life insurance, paid parental leave, and various other discounts and perks. You may also accrue 3 weeks of paid vacation and will be eligible for 10 or more paid holidays per year. Employees accrue paid sick leave pursuant to Company policy which satisfies or exceeds the accrual, carryover, and use requirements of the law.</p><div class="content-conclusion"><p><strong>ITAR REQUIREMENTS:</strong></p> <ul> <li>To conform to U.S. Government export regulations, applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii) Refugee under 8 U.S.C. § 1157, or (iv) Asylee under 8 U.S.C. § 1158, or be eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR <a href="https://www.pmddtc.state.gov/?id=ddtc_kb_article_page&amp;sys_id=24d528fddbfc930044f9ff621f961987">here</a>. &nbsp;</li> </ul> <p>SpaceX is an Equal Opportunity Employer; employment with SpaceX is governed on the basis of merit, competence and qualifications and will not be influenced in any manner by race, color, religion, gender, national origin/ethnicity, veteran status, disability status, age, sexual orientation, gender identity, marital status, mental or physical disability or any other legally protected status.</p> <p>Applicants wishing to view a copy of SpaceX’s Affirmative Action Plan for veterans and individuals with disabilities, or applicants requiring reasonable accommodation to the application/interview process should reach out to&nbsp;<a href="mailto:EEOCompliance@spacex.com">EEOCompliance@spacex.com</a><em>.&nbsp;</em></p></div>

Perks & benefits

  • 401k
  • Dental Insurance
  • Paid Time Off
  • Equity Compensation

755,000+ hidden jobs like this

SpaceX and thousands of companies post here first — often days before LinkedIn or Indeed. Your first 5 applications are free; go Pro to apply without limits.

Everything Pro unlocks:

  • Unlimited applications — free stops at 5
  • Track every application in one place
  • Apply straight to the source, one click
  • Save & organize roles you love
  • Roles pulled from company boards before the big sites

Weekly

$9.99
$4.99/week

For an active search. Cancel anytime.

Most popular

Monthly

$24.99
$12.99/month

The smart pick. Save 35% vs weekly.

Lifetime

$99
$49.99once

Pay once. Every future feature, forever.