Back to search
Data & AnalyticsOn-site

Senior Software Engineer, Machine Learning Infrastructure

Handshake · On-site

Apply
Last seen by MeritLog September 10, 2026Source: AshbySource version: ashby-public-job-posting-v1

MeritLog read this listing from Handshake's Ashby job board and last checked it on September 10, 2026.

Source: the employer's Ashby job board. Open the original listing for current details.

Job details

Work model
On-site
Salary
$176K - $220K
Location
San Francisco, CA
Occupation
Software Developers(O*NET 15-1252.00)

Hiring context

How this role compares at Handshake

Handshake has 67 live roles in MeritLog’s catalog across 8 job families, and 22 of them are in data & analytics. 54 of those listings publish a pay range, a disclosure rate of 81%.

This role's posted range of $176K - $220K sits above 54% of the 46 other Handshake roles quoted over the same currency and period.

Handshake concentrates this hiring in:

Counted across the job boards MeritLog tracks, at the time this page was served. Pay comparisons use only listings that publish a complete range in the same currency and period.

What the role asks for

What you'd do

  • Build and operate the shared infrastructure behind production ML and AI, including data pipelines, feature stores, training, and model serving.
  • Develop and scale our LLM platform, including provider integrations, orchestration, observability, and controls for cost, latency, and reliability.
  • Build evaluation infrastructure, including LLM eval harnesses, benchmarks, and quality measurement pipelines.
  • Support post-training workflows, including fine-tuning, reinforcement learning pipelines, and supporting data infrastructure.
  • Optimize inference infrastructure for open and fine-tuned models, including GPU serving, batching, and autoscaling.
  • Partner with AI, Data Science, and Product teams to productionize new models and establish best practices for ML infrastructure across Handshake.
  • Improve the reliability, scalability, and developer experience of our ML platform.
  • 5+ years of production software engineering experience using Python, Go, TypeScript, or similar languages.
  • Experience building and operating cloud infrastructure on AWS, GCP, or similar platforms.
  • Strong experience with Kubernetes, Docker, Terraform, CI/CD, and operating production services.
  • Hands-on experience building ML infrastructure, including model serving, training pipelines, feature stores, embeddings, or ML observability.
  • Experience with modern data platforms such as BigQuery, Airflow, Spark, Beam/Dataflow, or streaming pipelines.
  • Practical experience building production systems with LLMs or generative AI, including orchestration, provider APIs, observability, and performance optimization.
  • Strong systems design skills, sound engineering judgment, and the ability to thrive in ambiguous, fast-moving environments.

What they're asking for

  • Experience with Ray, Anyscale, KubeRay, Ray Serve, vLLM, Triton, PyTorch, or GPU-backed inference and training.SkillPreferred
  • Experience designing LLM evaluation frameworks, benchmarking systems, or quality regression testing.SkillPreferred
  • Experience with Vertex AI, Bigtable, Redis, or feature platform infrastructure.SkillPreferred
  • Experience with post-training techniques such as fine-tuning, RLHF, reinforcement learning, or reward modeling.SkillPreferred
  • Experience building agentic systems, MCP integrations, tool use, memory systems, or voice AI applications.SkillPreferred

Parsed by MeritLog from the employer’s own posting. The full description follows below.

Job description

About Handshake Handshake was founded on a simple belief that everyone deserves a path to a great career, regardless of where they went to school or who they know. Today, we power 25 million job seekers, 1 million+ employers, and 1,600 educational institutions. In 2025, we started Handshake AI and built the fastest-growing AI data business in history. We work directly with frontier AI lab researchers to create evaluations, publish benchmarks, and push the boundary of data. We’ve grown from $0 to ~$1B run rate and pay ~$60M to over 30K individuals every month. Why join Handshake now: - Shape how every career evolves in the AI economy, at global scale, with impact your friends, family and peers can see and feel - Partner hand-in-hand with world-class AI labs, Fortune 500 partners and the world’s top educational institutions - Work together with engineers, scientists, operators, and more from Palantir, Meta, Scale AI, and former YC founders - Build a massive, fast-growing business with billions in revenue About Handshake AI Human data is the core infrastructure to AI advancement. Frontier AI labs currently improve model capabilities with various data-intensive post-training techniques. We believe that data spend for AI training will increase by 3-5x in the next few years and continue for much longer as models take on new domains. Handshake AI supports all of the frontier AI labs, working on their most complex data at the largest scale. ABOUT THE ROLE We’re looking for a Senior Software Engineer to join our ML Infrastructure & Platform team. This team powers both Handshake’s core career marketplace and Handshake AI by building the shared infrastructure behind our production ML and AI systems. This is an infrastructure-heavy role for an engineer who enjoys building scalable platforms at the intersection of software engineering, machine learning, and generative AI. You’ll help teams move quickly from prototype to production while building the reliable, high-performance systems that power training, evaluation, and inference across Handshake. WHAT YOU’LL DO - Build and operate the shared infrastructure behind production ML and AI, including data pipelines, feature stores, training, and model serving. - Develop and scale our LLM platform, including provider integrations, orchestration, observability, and controls for cost, latency, and reliability. - Build evaluation infrastructure, including LLM eval harnesses, benchmarks, and quality measurement pipelines. - Support post-training workflows, including fine-tuning, reinforcement learning pipelines, and supporting data infrastructure. - Optimize inference infrastructure for open and fine-tuned models, including GPU serving, batching, and autoscaling. - Partner with AI, Data Science, and Product teams to productionize new models and establish best practices for ML infrastructure across Handshake. - Improve the reliability, scalability, and developer experience of our ML platform. DESIRED CAPABILITIES - 5+ years of production software engineering experience using Python, Go, TypeScript, or similar languages. - Experience building and operating cloud infrastructure on AWS, GCP, or similar platforms. - Strong experience with Kubernetes, Docker, Terraform, CI/CD, and operating production services. - Hands-on experience building ML infrastructure, including model serving, training pipelines, feature stores, embeddings, or ML observability. - Experience with modern data platforms such as BigQuery, Airflow, Spark, Beam/Dataflow, or streaming pipelines. - Practical experience building production systems with LLMs or generative AI, including orchestration, provider APIs, observability, and performance optimization. - Strong systems design skills, sound engineering judgment, and the ability to thrive in ambiguous, fast-moving environments. EXTRA CREDIT - Experience with Ray, Anyscale, KubeRay, Ray Serve, vLLM, Triton, PyTorch, or GPU-backed inference and training. - Experience designing LLM evaluation frameworks, benchmarking systems, or quality regression testing. - Experience with Vertex AI, Bigtable, Redis, or feature platform infrastructure. - Experience with post-training techniques such as fine-tuning, RLHF, reinforcement learning, or reward modeling. - Experience building agentic systems, MCP integrations, tool use, memory systems, or voice AI applications. Perks Handshake delivers benefits that help you feel supported-and thrive at work and in life. The below benefits are for full-time US employees. 🎯 Ownership: Equity in a fast-growing company 💰 Financial Wellness: 401(k) match, competitive compensation, financial coaching 🍼 Family Support: Paid parental leave, fertility benefits, parental coaching 💝 Wellbeing: Medical, dental, and vision, mental health support, $500 wellness stipend 📚 Growth: $2,000 learning stipend, ongoing development 💻 Remote & Office: Internet, commuting, and free lunch/gym in our SF office 🏝 Time Off: Flexible PTO, 15 holidays + 2 flex days 🤝 Connection: Team outings & referral bonuses Explore our mission, values, and comprehensive US benefits at joinhandshake.com/careers http://joinhandshake.com/careers.

Privacy choices

Analytics and advertising stay off unless you allow them. Private data stays out.

Read the privacy notice