Member of Technical Staff, Distributed Systems
Psi
Tailor my CV for this job, freeView job and applyYour CV rewritten for this role, from your real experience. Sign in with Google, nothing to install.
Got this interview? Our apps help you get the job.
Skills named in this job
Read from the description itself, not inferred.
This role on the market
156 open distributed roles across 52 companies are on ApplySarthi right now, most of them in Bengaluru (13), Hyderabad (2), Pune (1).
- Software Development Engineer - Distributed ServicesAmazon
- Staff Engineer - Distributed SystemsJobgether
- Senior Software Engineer, Distributed Systems Engineer - DGX CloudNvidia
- Software Engineering LMTS - Backend Distributed SystemsSalesforce · bengaluru
- Full-stack Engineer 4 (Distributed data & AWS) (Cloud Operations Resilience Engineering)Capitalone
What distributed roles keep asking for: System design (62%), Observability (49%), Java (46%), C++ (33%), Python (30%), AWS (29%), Go (26%), Kubernetes (23%) — counted across their open postings here.
Member of Technical Staff jobs in the United States · Remote Member of Technical Staff jobs · Kubernetes jobs · Observability jobs
Psi has 19 open roles listed here.
- Contract Recruiter
- Technical Program Manager, Datacenter Optimization
- Member of Technical Staff, Data Systems
- Member of Technical Staff, Performance Engineering
- Member of Technical Staff, Engineering (General Application)
Counted across 14 company job boards, updated as roles open and close.
Preparing for this interview
Interviews for distributed roles keep coming back to System design, Observability, Java, C++. Practise those questions before you sit with Psi.
Questions you are likely to be asked
- Why do you want to join Psi?
- What is your experience with Observability? Tell me one thing you learned the hard way.
- Tell me about a problem you solved at work that you are proud of.
- Tell me about a time you disagreed with your manager. What happened?
- Where do you want to be in three years?
Prep Sarthi gives you a free mock interview: an AI interviewer asks you questions like these out loud, from your own CV and this job, and shows your score and your weakest answer.
Practise the Member of Technical Staff, Distributed Systems at Psi interview free →Overview Physical Superintelligence is a startup with roots at Google, NVIDIA, Harvard, Meta, MIT, Oxford, Johns Hopkins, Cambridge, and the Perimeter Institute building AI systems to discover new physics at scale. We are seeking engineers to build platform infrastructure at the intersection of computational science, AI systems, and software engineering. Our mission is to discover and commercialize transformative physics breakthroughs at scale with artificial superintelligence, safely, verifiably, and for broad public benefit. The last century's golden age of physics gave us transistors, lasers, and nuclear energy. We believe artificial superintelligence will unlock the next one. We're creating the infrastructure to industrialize scientific discovery and usher in this new era. We have one product: new physics, at scale. We are seeking a Member of Technical Staff, Distributed Systems to build the execution layer everything at PSI runs on: the systems that decide what runs where, keep it alive through failure, and make every run replayable and priced. A research campaign is a great many jobs that have to survive machines dying and still add up to a result somebody can reproduce a year from now. That is this layer's problem. Role and Responsibilities Design the runtime primitives researchers and engineers compose into agentic workflows: sequential pipelines, tree-search agents, and whatever pattern the science asks for next. Treat the platform as a library product. Clear layers, explicit API contracts, surfaces other engineers extend instead of fork. Build the durable execution layer. A researcher's request becomes work that finds the right cluster, runs there, survives partial failure, and comes back replayable. Correctness under retries and replays. Idempotency as a design default. Admission and placement when the cluster is full, and gang scheduling for the jobs that need it. Make every run priceable and replayable end to end. Every call traced, cost recorded at call time, one correlation ID from request to result. Operate what you build. Set SLOs and meet them, build the instrumentation, plan capacity with the infrastructure team, and take your turn in incident response for the systems you own. What We're Looking For Five or more years building and operating distributed systems in production at companies known for engineering rigor, on major cloud platforms with Kubernetes, Slurm, or comparable orchestration. You have written code that paying customers or internal teams depend on every day. You have implemented or substantially extended a durable scheduler, workflow engine, dataflow system, or agent runtime. You know the bugs that come out of retries, replays, and non-deterministic execution. You have the failure stories that prove you operated it, not only wrote it. You have shipped a library or internal framework other engineers extend rather than only consume. API ergonomics, composability, and backward-compatible evolution are first-order concerns for you. You have made a build-versus-adopt call on an orchestration engine and can argue both sides. You favor the boring, durable answer. You can explain a workflow-engine tradeoff in two minutes. Nice to Have Hands-on with a durable workflow system at scale: Temporal, Cadence, Step Functions, or Argo Workflows. Data-infrastructure depth alongside the systems work: content addressing, catalogs, lineage. Our execution and data work sit next to each other, and the engineer strong in both is the one we most want to meet. Systems where tracing and metering were product requirements, not afterthoughts; production observability on OpenTelemetry or comparable. Background in scientific computing, HPC environments, or research infrastructure. How We Work We hold a high technical bar and give people full ownership of their work, from spec to ship to on-call. We write contracts before logic, test against real systems instead of mocks, and favor simple designs that ship over clever ones that do not. Our development process is AI-native: we work with agentic coding tools daily, write specs that are legible to humans and agents alike, and lead with leverage. Location and Compensation This role is based in Boston. We will consider remote candidates on a case-by-case basis. We offer competitive compensation including salary, benefits, and meaningful early-stage equity. We evaluate on technical breadth, systems thinking, scientific curiosity, and shipping velocity. We are an equal opportunity employer and value diverse perspectives in building platforms for AI-driven discovery.
Match this job to your CV
ApplySarthi scores your CV against this role, shows the skills you are missing, and writes a tailored version for the application.
Check my match →Similar open roles
Need answers during your interview? Try Live Sarthi.
Live Sarthi, an Interview Sarthi app, shows answer suggestions during the call.
- Hidden from supported screen sharingThe overlay stays out of supported Windows screen captures.
- Answers start in about 1.5 secondsResponse time varies with your connection and model.
- From your own CVYour projects and your experience, not a generic script.
- 30 minutes freeThen ₹99 for a 2-day pass with unlimited calls — you pay for the days you are interviewing, not a subscription.
A Windows app, from the same team as ApplySarthi.
Listed on ashby · posted 2026-05-13. ApplySarthi collects openings and links to application pages; the role is advertised by Psi, not by us.