Member of Technical Staff - Mid-training
SpaceXAI · Not provided by source
MeritLog read this listing from SpaceXAI's Greenhouse job board and last checked it on September 12, 2026.
Source: the employer's Greenhouse job board. Open the original listing for current details.
Job details
- Work model
- Not provided by source
- Salary
- $180,000 – $440,000
- Location
- Palo Alto, CA
Hiring context
How this role compares at SpaceXAI
SpaceXAI has 255 live roles in MeritLog’s catalog across 9 job families, and 110 of them are in data & analytics. 3 of those listings publish a pay range, a disclosure rate of 1%.
SpaceXAI concentrates this hiring in:
Counted across the job boards MeritLog tracks, at the time this page was served. Pay comparisons use only listings that publish a complete range in the same currency and period.
What the role asks for
What you'd do
- Scale synthetic coding data to trillions of tokens with large-scale docker verification.
- Distill the intelligence of flagship models into flash models through synthetic data generation.
- Optimize mid-training data mixtures to boost the ceiling for RL.
- Engineer long-context data recipes.
- Develop robust and diverse evaluation for mid-training checkpoints.
What they're asking for
- Expertise in ML and large model scaling, with familiarity across all kinds of scaling laws.Skill
- Strong ability to design ML experiments.Skill
- Familiarity with state-of-the-art techniques for curating AI training data for text, image, audio, and video modalities.Skill
- Strong engineering abilities in Spark, Ray, and other frameworks for large-scale data processing.Skill
Parsed by MeritLog from the employer’s own posting. The full description follows below.
Job description
SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates. RESPONSIBILITIES: • Scale synthetic coding data to trillions of tokens with large-scale docker verification. • Distill the intelligence of flagship models into flash models through synthetic data generation. • Optimize mid-training data mixtures to boost the ceiling for RL. • Engineer long-context data recipes. • Develop robust and diverse evaluation for mid-training checkpoints. BASIC QUALIFICATIONS: • Expertise in ML and large model scaling, with familiarity across all kinds of scaling laws. • Strong ability to design ML experiments. • Familiarity with state-of-the-art techniques for curating AI training data for text, image, audio, and video modalities. • Strong engineering abilities in Spark, Ray, and other frameworks for large-scale data processing. COMPENSATION AND BENEFITS: $180,000 - $440,000 USD Base salary is just one part of our total rewards package at SpaceXAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short & long-term disability insurance, life insurance, and various other discounts and perks. SpaceXAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice.
Keep exploring