Hello everyone!
We're currently hiring a Senior Software Engineer, Infrastructure & Platform for a research lab investigating the boundaries and capabilities of artificial intelligence through novel datasets and experimentation. Our customers are the frontier AI labs. We're backed by top-tier venture investors and advised by senior leadership from leading AI research organizations, and we were one of the fastest-growing companies in our YC batch.
Our founding team brings backgrounds from top quantitative trading firms, Big Tech, leading investment banks, and world-class AI research labs.
This is a full-time role, highly technical, with broad ownership from day one.
Tasks:
— Architect and develop the shared infrastructure powering data generation platforms, human-in-the-loop systems, and evaluation pipelines;
— Build systems capable of processing large-scale datasets and high-throughput workloads with strong reliability guarantees;
— Create reusable infrastructure and APIs that enable product engineers and researchers to build quickly and reliably on top of core systems;
— Design systems with strong observability, monitoring, and fault tolerance to support production workloads at scale;
— Help define long-term system architecture across data pipelines, compute infrastructure, task orchestration, and storage systems;
— Work closely with engineers and researchers to support new AI experimentation workflows and platform capabilities;
— Define standards for system design, deployment, reliability, and infrastructure operations;
What professional skills are important for us?
— 5+ years of experience building production distributed systems or platform infrastructure;
— Proficiency in Python and/or JavaScript (Node.js/Next.js) or similar backend technologies;
— Experience designing and operating systems in cloud environments (GCP or AWS);
— Experience with message queues and event-driven systems (Kafka, RabbitMQ, Pub/Sub, or equivalent);
— Experience working with high-throughput data pipelines and asynchronous processing systems;
— Strong understanding of system scalability, performance, and reliability;
— Experience owning systems running in production environments at scale— Background at a top-tier startup or FAANG-equivalent is a strong signal of fit;
Nice to Have:
— Experience building internal developer platforms or shared infrastructure;
— Experience supporting large-scale data processing pipelines;
— Experience with AI infrastructure, LLM evaluation systems, or ML pipelines— Experience working at high-growth startups or scaling early infrastructure;
— Experience designing human-in-the-loop or workflow orchestration systems;
We are especially interested in candidates from top-tier tech companies, high-growth startups with strong engineering cultures, and teams building large-scale distributed or AI infrastructure.
Interview Process:
— Intro call with recruiter;
— Initial screening;
— Technical interview (backend engineering & infrastructure fundamentals);
— System design interview (architecture and scalability focus);
— Final interview with engineering leadership;
Location & Work Setup
San Francisco, CA
Full-time position