Steadyworks

Part-time Research Collaboration

Remote or SF

For domain experts and researchers who want to shape the evaluations frontier labs actually use.

About Steadyworks

Steadyworks is a data research lab powering the next generation of AI. We partner with frontier AI labs to build expert-grade datasets, reinforcement learning environments, and evaluation infrastructure that improve state-of-the-art models. Our team previously built frontier AI systems at Meta, Google Research, Autodesk, and other leading organizations.

What this looks like

A part-time fellowship for experts who already do the work we're trying to measure — competitive programmers, ML researchers, CAD practitioners, and other frontier-domain specialists. You'll help design tasks, author environments, and grade agent behavior.

  • Authoring realistic tasks in your domain of expertise
  • Grading agent trajectories and writing rubrics that scale
  • Advising on what a good evaluation looks like in your field
  • Optional: co-authoring writeups or papers where the collaboration warrants it

Good fits include

  • Competitive programmers, ICPC / IOI alumni, or top-rated Codeforces / AtCoder contestants
  • Working ML researchers who want to shape the evals for the field
  • Practicing engineers in CAD, mechanical, or other technical domains
  • PhD students looking for a paid, high-signal side collaboration

Logistics

  • Fully remote or hybrid in San Francisco
  • Flexible hours; usually 5–15 hours per week
  • Paid hourly or per-task depending on the engagement

How to apply

Email careers@steadyworks.ai with anything that would help us understand your work.

Introduce yourself