Litmus

Evals for humans

YC Summer 2026RecruitingData EngineeringAI

About Litmus

Litmus is building the most accurate framework for evaluating and benchmarking human capability, starting with software. AI will compound small differences in human capability into increasingly large differences in what people can accomplish, while making existing static benchmarks obsolete. Software is already there: AI can hill-climb any output-based evaluation, while the ability to direct it is becoming the defining advantage. Litmus applies the same approach we already use for model evals to humans – creating world-like environments, and inspecting trajectory instead of just output. Every knowledge industry will soon face the same problem. We build Litmus to tell you what humans are capable of. We're already helping build frontier technical teams at Mercor, Composio, Neo Scholars, and more.

Founders

  • Shaivi Rau

    Founder

    Co-founder, CEO of Litmus. Hiring sucks, we're fixing it. Prev @ Morgan Stanley, early stage teams, and in VC. CS & Film @ Columbia.

    LinkedIn ↗X ↗

  • Elena Zhao

    Founder

    Building Litmus

    LinkedIn ↗X ↗

Discussion

Posting anonymously

No comments yet — start the discussion.