Member of Technical Staff, India
- India
- On-site
- Posted 15 days ago
- Software Engineering
- FullTime
About the job
ABOUT HUMAN ARCHIVE
Human Archive is a research lab backed by Y Combinator focused on modeling human embodied intelligence.
Humans are the most sophisticated biological systems we have ever observed, yet we still do not fully understand ourselves. Research into human physical intelligence — including the human hand, proprioception, and vision — remains largely unsolved. Our mission is to recover human embodied intelligence as a learned model. To achieve this, we build custom hardware products, deploy them globally at scale, and publish research. Today, our data is used for robotics and world modeling, but the broader opportunity is advancing scientific research into intelligence itself.
Founded by Stanford and UC Berkeley researchers, we are lean, deeply technical, and operate at extreme speed, taking on unglamorous and conventionally impossible problems that directly unlock step-function gains in model capability.
The deployment of capable humanoids at scale will permanently redefine human labor. Undesirable physical work will disappear, and human effort will shift toward a new era of abundant creativity.
We are building the infrastructure to accelerate that transition by assembling the Human Archive mafia. You will own meaningful systems from day one and see your work directly impact model capabilities. This is a once-in-a-generation inflection point. If you want to help reshape physical labor and work on problems that matter at civilizational scale, join us.
About the Role
We build and run the software platform behind a large-scale data collection and annotation operation: the system of record that tracks what was collected and by whom, the tooling annotators work in, the ingest integrations that move data from physical media into cloud storage, and the reporting that tells us whether any of it is working.
These are internal systems with real users — operators in the field, annotators working through millions of items, ingest stations moving terabytes, and machine clients writing back automatically. They are small enough that you will own whole systems and serious enough that correctness matters: a duplicated record or a dropped identifier means data that can no longer be traced to what produced it.
We want a strong generalist. You will design schemas, build APIs, write the interfaces on top of them, and deploy and debug the whole thing in production. Nobody here hands you a spec.
WHAT YOU’LL DO
WHAT WE’RE LOOKING FOR