Member of Technical Staff - Research
Patronus AI · Patronus AI develops simulation research and infrastructure to accelerate progress toward human-aligned AGI.
San Francisco11-50 employeesPosted 5 days ago
Series B · $50Mraised 3 months agoled by Lightspeed Venture Partners, Samsung NEXT, Notable Capital
This board only lists companies whose most recent round closed in the last 180 days.
<div class="content-intro"><h2 style="text-align: justify;"><span style="font-family: arial, helvetica, sans-serif;"><strong>About Patronus AI</strong></span></h2>
<p style="text-align: justify;"> </p>
<p style="text-align: justify;"><span style="font-family: arial, helvetica, sans-serif;">Patronus AI is a frontier lab developing simulation research and infrastructure to accelerate progress toward human-aligned AGI. We are on a mission to simulate all of the world’s intelligence.</span></p>
<p style="text-align: justify;"> </p>
<p style="text-align: justify;"><span style="font-family: arial, helvetica, sans-serif;">We are the team behind some of the earliest and most influential research in AI evaluation like<a class="css-173makr-linkStyle" href="https://arxiv.org/pdf/2311.11944" target="_blank"> <u>FinanceBench</u></a>,<a class="css-173makr-linkStyle" href="https://www.forbes.com/sites/rashishrivastava/2024/07/11/this-ai-powered-coach-catches-hallucinations-in-other-ai-models/" target="_blank"> Lynx</a>,<a class="css-173makr-linkStyle" href="https://arxiv.org/abs/2311.08370" target="_blank"> <u>SimpleSafetyTests</u></a>,<a class="css-173makr-linkStyle" href="https://www.cnbc.com/2024/03/06/gpt-4-researchers-tested-leading-ai-models-for-copyright-infringement.html" target="_blank"> <u>CopyrightCatcher</u></a>,<a class="css-173makr-linkStyle" href="https://arxiv.org/abs/2501.14249" target="_blank"> <u>Humanity’s Last Exam</u></a>, and more. We are formerly AI researchers and engineers from companies like Meta AI, Amazon AGI, and Google. Our customers include foundation model labs and Fortune 500 enterprises like Adobe. We are backed by top-tier investors like Lightspeed Venture Partners, Notable Capital, Stanford University, Noam Brown, Gokul Rajaram, and more.</span></p></div><h2><strong>Responsibilities</strong></h2>
<p> </p>
<p>As a Researcher at Patronus AI, you will own and drive foundational research that defines how agentic AI systems are trained, evaluated, and improved. You will work at the intersection of reinforcement learning, simulations, and scalable oversight, building systems that directly influence how frontier models are developed, stress-tested, and deployed.</p>
<p>This is a highly autonomous role. You will tackle open-ended research questions surrounding agent simulations and translate them into rigorous experiments, benchmarks, environments, and production systems. You will work across areas including reward design, tool simulations, agent cognition, behavior analysis, and scalable oversight, helping shape the industry standard for robust, high-quality environments.</p>
<p>Your work will inform how frontier labs design, train, evaluate, and improve the next generation of agents for complex, long-horizon tasks, advancing our path toward safe, human-aligned general intelligence.</p>
<p> </p>
<p><strong>In this role, you will: </strong></p>
<p> </p>
<ul data-pattern="discCircleSquare" data-depth="1">
<li><strong>Own ambitious research projects end-to-end</strong>, from identifying and formulating open-ended problems through experiment design, execution, analysis, and production impact.</li>
<li><strong>Advance research in agent simulation, reinforcement learning, and scalable oversight</strong>, including agent cognition, behavior analysis, reward design, and new training methods.</li>
<li><strong>Design state-of-the-art simulation and RL environments</strong> for training and evaluating frontier agents, spanning tools and actions, observations and state, trajectories, curricula, and reward systems.</li>
<li><strong>Train models and experiment with post-training algorithms</strong>, including GRPO and SFT. Run ablations with open source models and understand the impact of distillation, COT reasoning, sparse and dense rewards and hyperparameters.</li>
<li><strong>Develop methods to understand and improve agent behavior</strong> across complex, long-horizon tasks, including reasoning, planning, adaptation, generalization, and reward hacking.</li>
<li><strong>Run rigorous experiments and turn findings into measurable outcomes</strong>, including new techniques, benchmarks, datasets, environments, platform capabilities, and research publications.</li>
<li><strong>Build high-quality, reproducible research systems</strong>, writing production-level code and partnering closely with engineering and product to translate research into real-world systems.</li>
<li><strong>Contribute to Patronus AI’s research direction and thought leadership</strong>, staying at the frontier of the field, collaborating with the research community, and publishing or open sourcing our work.</li>
</ul>
<h2><strong>Qualifications</strong></h2>
<p> </p>
<p><em>"The number one qualification to succeed in this machine learning course is gumption” - John Lafferty, CS Professor at Yale</em></p>
<p> </p>
<p>Above all, we look for an eagerness to learn, passion for research, creativity in problem solving and a proactive mindset. You are a great fit if you have a background in the following:</p>
<p> </p>
<ul data-pattern="discCircleSquare" data-depth="1">
<li>An MS or PhD in Computer Science, Machine Learning, Statistics, Mathematics, or a related quantitative field.</li>
<li>Experience conducting independent research in reinforcement learning, NLP, agentic systems, evaluation, alignment, or related areas.</li>
<li>Demonstrated ability to take open-ended research problems from 0→1 and deliver high-impact outcomes.</li>
<li>Strong experimental skills, including experiment design, analysis, and interpretation of results.</li>
<li>Experience writing clean, reproducible research code in Python and modern machine learning frameworks.</li>
<li>Ability to execute quickly and independently with minimal guidance while maintaining a high bar for research quality.</li>
<li>Experience collaborating cross-functionally with research, engineering, and product teams.</li>
<li>Clear written and verbal communication skills, including the ability to explain complex technical ideas succinctly.</li>
<li>Strong integrity, good judgment, and respect for others.</li>
</ul>
<div id="message-list_1780521026.163069">
<div>
<div>
<div>
<div>
<div>
<div>
<div>
<div>
<div>
<div>
<div> </div>
<div><strong>To support close collaboration, this role is based in our San Francisco headquarters and requires in-office attendance five days a week.</strong></div>
<div> </div>
<div><em>The expected base salary range for this role is </em><em>$175,000 - $300,000 USD</em><em>. In addition to base salary, we offer equity and benefits. Actual compensation will be determined based on experience, qualifications, skills, and location.</em></div>
<div> </div>
<div>
<h2><strong>Benefits</strong></h2>
<p> </p>
<ul data-pattern="discCircleSquare" data-depth="1">
<li>Competitive salary and equity packages</li>
<li>15 days of paid vacation per annum</li>
<li>Parental & sick leave</li>
<li>Health, dental, and vision insurance plans</li>
<li>401(k) plan + matching</li>
<li>In-office lunch & dinner</li>
<li>Whoop band</li>
<li>Monthly meal stipend</li>
<li>Monthly health and wellness stipend</li>
<li>Equinox membership</li>
<li>Fun global offsites!</li>
</ul>
</div>
</div>
</div>
</div>
</div>
</div>
</div>
</div>
</div>
</div>
</div>
</div><div class="content-conclusion"><p> </p>
<p><span style="font-family: arial, helvetica, sans-serif;"><em>Patronus AI is an equal opportunity employer. We celebrate diversity in our workplace, and all qualified applicants will receive consideration for employment without regard to age, ancestry, color, family or medical care leave, gender identity or expression, genetic information, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran status, race, religion, sex (including pregnancy), sexual orientation, or other legally protected</em></span><span style="font-family: arial, helvetica, sans-serif;"><em> characteristics.</em></span></p>
<p> </p>
<p><span style="font-family: arial, helvetica, sans-serif;">By clicking ‘Apply’, you agree to Greenhouse's <a href="https://www.greenhouse.com/legal">Terms of Service </a>and <a class="css-1d2k4kf e10ya71h0" href="https://www.greenhouse.com/privacy-policy" target="_blank">Privacy Policy</a><a class="css-1d2k4kf e10ya71h0" href="https://app.rippling.com/legal/privacy" target="_blank">.</a></span></p>
<p><span style="font-family: arial, helvetica, sans-serif;">By clicking 'Apply', you agree to Patronus AI, Inc. <a href="https://www.patronus.ai/privacy-policy">Privacy Policy</a>.</span></p></div>
Apply on Patronus AI’s site
Applications go straight to the employer. VCBacked does not sit between you and the company.