logo
AI Resume Tailoring Sign in to use this AI AI Cover Letter Sign in to use this AI

Job Description

Our Mission

Reflection is a research lab making intelligence open and accessible for everyone to use, customize, and build on. We build open models that let anyone control their intelligence and help shape the future of AI. Our mission: make intelligence open and accessible to all.

Foundations

Vision:

Build and operate a company-wide foundations platform that accelerates every team by providing reliable, scalable developer infrastructure, SRE capabilities, and high-throughput data ingestion tooling enabling Reflection to move faster as we scale.

What This Team Does

Build and operate the core shared services that power our research, training, and production environments. These systems form the foundational platform that multiple teams depend on for model development, deployment, and evaluation, unifying data, compute, and workflow management across the stack while enabling rapid experimentation and reliable production systems.

  • Build and operate shared services that multiple teams rely on across research and production workflows.

  • Define and uphold reliability targets through SLIs, SLOs, and healthy on-call practices.

  • Maintain strong operational readiness with runbooks, incident playbooks, and capacity planning.

  • Ensure correctness and performance under load, addressing consistency, tail latency, and failure modes.

  • Develop APIs, SDKs, and internal platforms that enable high-velocity experimentation and iteration.

  • Reduce operational burden through better tooling, standardization, and platform patterns that scale across teams.

What You'll Work With

  • Container Abstractions: Containers-as-a-Service, Kubernetes abstraction layers, container orchestration, reproducible environments, multi-tenant isolation.

  • Distributed Systems Architecture: Sharding, replication, coordination services, high-concurrency systems, concurrency control.

  • Service Development Stack: gRPC, Protobuf, Go, Rust, C++.

  • Reliability & Performance: Idempotency, retries, backpressure, SLI/SLO design, tail latency optimization, service reliability engineering.

About You

  • Strong software engineering background with experience shipping production-grade systems.

  • Experience designing APIs, services, or developer platforms that handle large-scale data or compute.

  • Comfortable navigating complex codebases, debugging hard problems, and optimizing for reliability and speed.

  • Thrive in a high-agency, fast-paced startup environment; bias toward action and impact.

  • Excited about zero to one challenges, building new systems rather than maintaining legacy ones.

  • Collaborative, clear communicator, and comfortable working across research and infra boundaries.

  • Motivated by creating the software backbone for the world’s most capable open-weight AI systems.

What We Offer:

We believe that to make intelligence open and accessible to all, you need to start at the foundation. Joining Reflection means building from the ground up as part of a talent-dense team. You will help define our future as a company, and help define the future of open foundational models.

We want you to do the most impactful work of your career with the confidence that you and the people you care about most are supported.

  • Top-tier compensation: Salary and equity structured to recognize and retain our talent globally.

  • Stock options: Everyone who joins and contributes to Reflection's success gets to share in the upside through stock options.

  • Health & wellness: Comprehensive medical, dental, vision, and life, with an annual wellness allowance.

  • Meals: Lunch and dinner are provided in the office daily.

  • Life & family: 22 weeks paid parental leave for all new birthing and non-birthing parents, including adoptive and surrogate journeys.

  • Vacation days: Unlimited paid time off in the U.S. and 30 days in the U.K.

  • Sponsorship support: We sponsor visas to help exceptional talent join our team and support long-term immigration pathways where applicable.

  • Team building: We have regular off-sites, happy hours, and team celebrations.

Export Control Notice: This position may require access to technology or source code subject to the U.S. Export Administration Regulations. Any offer of employment for this role may be conditioned on the Company's ability to provide the candidate with access to such technology or source code in compliance with applicable U.S. export control laws, which may require the Company to seek government authorization.

Do you match this job?

Here is what this employer asked for. Sign in and we will fill in your half.

  • Role Member of Technical Staff - Distributed Systems Engineer you: —
  • Experience 8-9 years you: —
  • Education Any you: —
  • Work type On-site you: —
  • Location United States you: —
Check my match
Reflection AI
AI Research Lab · 50-100 Members · San Francisco, CA, United States

Reflection AI is a frontier research lab founded by researchers from Google DeepMind, with the stated aim of making advanced intelligence open and accessible.

It hires across reinforcement learning, large-scale training and the systems engineering that frontier training runs depend on.

All jobs at Reflection AI
Job Overview

Approx. salary range

213K – 297K

Our estimate — this employer did not publish a salary

Our estimate, not the employer’s. Worked out from the middle half of 15 comparable roles on Neural Jobs that did publish a salary, in the same field, country and experience band. The real figure for this job may be different.

Eligibility
United States Right to work in the United States required.
Workplace
On-site
Job Posted:
5 months ago
Job Type
Full Time
Education
Any
Experience
8-9 years

Share This Job: