AI Hiring Index

Handshake · Infrastructure · Senior · Posted 2026-07-10

Senior Software Engineer, LLM Platform

Handshake · San Francisco, CA · $176k–220k base

This range's midpoint is above 34% of posted infrastructure ranges at AI companies right now. See the salary index.

Apply on Handshake's site Watch Handshake for new roles

About Handshake

Handshake was founded on a simple belief that everyone deserves a path to a great career, regardless of where they went to school or who they know. Today, we power 25 million job seekers, 1 million+ employers, and 1,600 educational institutions.

In 2025, we started Handshake AI and built the fastest-growing AI data business in history. We work directly with frontier AI lab researchers to create evaluations, publish benchmarks, and push the boundary of data. We’ve grown from $0 to ~$1B run rate and pay ~$60M to over 30K individuals every month.

Why join Handshake now:

- Shape how every career evolves in the AI economy, at global scale, with impact your friends, family and peers can see and feel

- Partner hand-in-hand with world-class AI labs, Fortune 500 partners and the world’s top educational institutions

- Work together with engineers, scientists, operators, and more from Palantir, Meta, Scale AI, and former YC founders

- Build a massive, fast-growing business with billions in revenue

About Handshake AI

Human data is the core infrastructure to AI advancement. Frontier AI labs currently improve model capabilities with various data-intensive post-training techniques. We believe that data spend for AI training will increase by 3-5x in the next few years and continue for much longer as models take on new domains. Handshake AI supports all of the frontier AI labs, working on their most complex data at the largest scale.

ABOUT THE ROLE

Handshake is hiring a Senior LLM Platform Engineer to join our Data and ML Platform team. This team supports Handshake’s core career marketplace and Handshake AI (HAI) by building the shared data, ML, and LLM infrastructure behind production workflows.

This infrastructure-heavy role primarily owns our shared LLM control plane: LiteLLM gateways, provider integrations, shared clients, access controls, observability, cost attribution, capacity management, and hosted or self-hosted inference. You’ll also contribute to adjacent platform systems for workflow orchestration, model serving, shared cloud infrastructure, and developer enablement—working closely with Backend Platform, HAI engineering, data science, and FDEs.

What You’ll Do

- Build and operate our LiteLLM-based AI gateways and shared LLM clients.

- Own provider and model onboarding, routing, failover, rate limits, and capacity planning.

- Build self-service virtual-key, model-access, budget, and credential-management workflows.

- Establish SLOs, observability, cost attribution, and alerts for production LLM traffic.

- Safely qualify and roll out new models, providers, SDKs, and gateway configurations.

- Support hosted and self-hosted inference through a consistent platform interface.

- Partner with product, AI, and FDE teams to turn recurring delivery problems into paved-platform capabilities.

- Contribute to the broader Data and ML Platform, including workflow orchestration, model serving, shared cloud infrastructure, and developer tooling.

- Participate in team on-call and support, improving runbooks, automation, and reliability across owned platform services.

What We’re Looking For

- Strong production software engineering experience in Python, TypeScript, Go, or a similar language.

- Experience building or operating high-throughput API gateways, proxies, or multi-tenant platform services.

- Hands-on experience with Kubernetes, Terraform, CI/CD, and production service ownership.

- Practical experience with LLM provider APIs, including streaming, long-running requests, retries, timeouts, cancellation, and rate limits.

- Experience with authentication, quotas, credential management, tenant isolation, and auditability.

- Experience building observability for distributed systems and leading production incident response.

- Experience with usage metering, cost attribution, capacity planning, or FinOps.

- Experience with data or ML platform systems such as BigQuery, Airflow, streamin …

New infrastructure roles at AI companies, every Monday. The week's openings in this function across 286 companies, plus the weekly index. Free.

More infrastructure roles at Handshake

See also: Software Engineer jobs · AI jobs in San Francisco Bay Area · Handshake salaries · Python jobs · TypeScript jobs · PyTorch jobs.

This listing is reproduced from Handshake's public careers feed and links to the original. AI Hiring Index is not the employer and does not accept applications. All Handshake roles · AI salaries.