Research Scientist
Quick Summary
About idler idler is a frontier data research lab. We build the evals and environments that the world's leading frontier labs use to measure and train their models.
idler is a frontier data research lab. We build the evals and environments that the world's leading frontier labs use to measure and train their models.
After raising a $9m seed round led by Paradigm, we spent the last year developing coding evals for top coding models you know and love. At the same time, we've expanded into other domains besides coding: RSI & Auto-Research, Law, Enterprise Business, Cybersecurity, and others. Now, we are facing more lab demand for our data than we can serve, and are rapidly scaling the team to grow the business.
Our approach to creating training data scales using technology, and all of our data products are built on a unified self-reinforcing platform that learns through experience.
You would be joining a close-knit team that has reached product market fit, and your work would directly help to multiply our revenue.
You can see some of our work here: https://idler.ai/collections
About the Role
~1 min readAs a Research Scientist at idler, you'll own measuring and improving how models learn from our tasks. The job is to maximize learning signal we produce per unit time. You'll draw on your own experience and collaborate with researchers at the frontier to validate our data quality, identify where improvements are needed, and create new datasets. To succeed, you'll need to have extensive experience doing this work in production at a frontier lab.
Responsibilities
~1 min read- →
Work with our customers — researchers at frontier labs — to design novel post-training recipes and data quality measurement techniques
- →
Design and run our in-house post-training stack to measure model lift on our tasks
- →
Develop new data products based on datasets and experts available to us
- →
Identify opportunities to take advantage of self-reinforcing exponential feedback loops
- →
Create agents to analyze thousands of environments and millions of trajectories
- →
Help curate and specify task distributions for new corpora
- →
Work with procurement to ensure external data we acquire is suitable for refinement
- →
Create scaleable systems for ingesting & evaluating data we are considering buying
- →
Develop new techniques for mining data for signal
1+ years of experience doing RL in production at a frontier lab
Track record of post-training an LLM end to end
Desire to drive the research roadmap and implementation on a fast-moving team
Deep curiosity about how machines learn from data and how to extract the maximum learning signal from our tasks
Typescript, React, NodeJS, Postgres, Redis, Vercel, Cursor/Claude Code/Codex, Tinker, Modal, AWS, Daytona, GRPO
In-person in San Francisco
Competitive salary + meaningful equity
Free meals in office
Healthcare, 401(k), 15 days of PTO per year
Relocation assistance
Small, ambitious team
This is an in-person role in San Francisco. We're a tight-knit founding team and we play to win. Join us if you like to win too.
Location & Eligibility
Listing Details
- Posted
- August 26, 2026
- First seen
- September 25, 2026
- Last seen
- October 5, 2026
Posting Health
- Days active
- 9
- Repost count
- 0
- Trust Level
- 32%
- Scored at
- October 5, 2026
Signal breakdown
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.