logo

AI Job Summary

Build evaluation infrastructure that runs AI model testing at scale, working with modeling and optimization teams to develop new evaluations and resolve performance bottlenecks across the stack. You'll need experience building and optimizing large-scale distributed systems, with preferred expertise in LLM inference, GPU compute management, and workload scheduling. The role is hands-on across all infrastructure layers. Salary ranges from $180,000 to $440,000.

Written from this posting by Neural Jobs AI. The full description is below.

AI Resume Tailoring Sign in to use this AI Cover Letter Sign in to use this

Job Description

SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.

ABOUT THE ROLE:

The RL infrastructure team is looking for an engineer to help develop our evaluation infrastructure.

RESPONSIBILITIES:

  • Build highly reliable, efficient, and easy to use infrastructure that runs all our evaluations at scale
  • Collaborate closely with modeling & evaluation teams on developing new evals, maintaining evaluation signal quality, and building internal tooling to support training our next-generation of models
  • Identify and resolve performance bottlenecks in all layers of the eval infrastructure stack, including inference, capacity fleet management, asynchronous eval orchestration, and more

BASIC QUALIFICATIONS:

  • Experience in building, debugging, and optimizing efficiency of large-scale distributed systems
  • Willingness to dive deep and solve hard problems at all levels of the stack

PREFERRED SKILLS AND EXPERIENCE:

  • Experience in LLM inference
  • Experience in GPU compute management, workload scheduling, and dynamic resource optimization
  • Experience in developing interfaces for comparing models and evals

COMPENSATION AND BENEFITS:

$180,000 - $440,000 USD

Base salary is just one part of our total rewards package at SpaceXAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short & long-term disability insurance, life insurance, and various other discounts and perks.

SpaceXAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice.

Do you match this job?

Here is what this employer asked for. Sign in and we will fill in your half.

  • Role Member of Technical Staff - Evaluation Infrastructure
  • Experience 8-9 years
  • Work type On-site
  • Location United States
Check my match (free)
xAI
Foundation Models · 500+ Members · Palo Alto, CA, United States

xAI develops Grok, a family of large language models, and builds the large-scale training and inference infrastructure behind them, including the Colossus supercomputing cluster.

Founded in 2023, the company focuses on frontier model research and deployment across the X platform and its own products.

All jobs at xAI
Job Overview

Approx. salary range

209K – 286K

Our estimate — this employer did not publish a salary

Our estimate, not the employer’s. Worked out from the middle half of 38 comparable roles on Neural Jobs that did publish a salary, in the same field, country and experience band. The real figure for this job may be different.

Eligibility
United States Right to work in the United States required.
Workplace
On-site
Job Posted:
1 day ago
Job Type
Full Time
Experience
8-9 years

Share This Job: