OpenAI

AI Research and Deployment

Researcher,Connectors-AgentPost-Training

$250–380k San Francisco, California, United States FULL TIME

Market Sentiment

HIGH DEMAND

Neural analysis suggests this role is
optimal for Mid+ candidates.

The Brief

“Researcher, Connectors - Agent Post-Training at OpenAI. Skills: Agent Post-Training, LLMs, RL. Design and run experiments. improve agentic model behavior”

What You'll Achieve.

improve agentic model behavior; ship improvements into products; make agents genuinely useful

Industry & Context.

AI Research and Deployment

Problems you'll solve

open-ended problems; vague behavioral problem to a concrete experiment; define the hypothesis; build the pipeline; run the model; analyze the result; decide what to do next; messy qualitative behavior into concrete hypotheses

What They're Looking For.

Must Have

technical fundamentals in machine learning, software engineering, systems, statistics, hands-on experience with LLMs, RL, RLHF/RLAIF, post-training, evals, graders, synthetic data, model training, coding agents, tool-using agents, production ML systems

Nice to Have

experience with Python and machine learning frameworks, experience with Slack, experience with Google Workspace, experience with GitHub, experience with Notion, experience with Linear, experience with Salesforce

What You'll Do.

Design and run experiments

improve agentic model behavior

Own end-to-end improvements

Build evals and environments

expose model failures

Partner with product teams

translate product signal

Work on training and alignment interventions

Help decide integrations

Improve machinery for training

Take on cross-functional projects

turn qualitative behavior into hypotheses

How You'll Work.

Team & Collaboration

Partner with researchers; engineers; product teams; infrastructure teams; safety/alignment partners; Work across research; product; infrastructure; data; evals; safety boundaries

Communication Scope

communicate clearly with each group

Full Job Description

About the Team The Agent Post-Training team creates the frontier agents OpenAI ships to the world. We are training the models behind our agents in Codex, ChatGPT, the API, and other frontier products: persistent, proactive intelligence that can operate computers, collaborate with people and other agents, and expand what people and organizations can imagine, attempt, and achieve. We define what the next generation of agents should be able to do, build the training signal that teaches those abilities, and run the experiments that make them real. Our work spans coding, tool use, computer use, multi-agent coordination, long-horizon execution, factuality, instruction following, calibrated reasoning, and taste. Our team is where new model capabilities get made. We build the data, environments, graders, training methods, and feedback loops that shape what OpenAI's next agents can do, then carry those capabilities through major training runs and into the products people use. About the Role As a member of Agent Post-Training, Connectors, you will teach models how to interface with the top professional software using code. You will help train agents to use code, APIs, tools, and structured integrations to operate across applications like Slack, Google Workspace, GitHub, Notion, Linear, Salesforce, and other core systems of work. You will help enable models to take useful actions across a user’s digital context: finding information, updating systems, coordinating work, generating artifacts, and completing multi-step workflows through the tools teams already use. You will train models to be supercharged by the world’s most important productivity and enterprise software, turning connected tools into a powerful action surface for our agents. You will work with researchers, engineers, product teams, infrastructure teams, and safety/alignment partners to decide what should go into major model runs, measure whether it worked, and ship improvements into products used by real people.

Free ATS check

Applying for this Researcher, Connectors - Agent Post-Training role?

Most applicants get filtered before a human reads their resume. See if yours makes the cut.

Should you apply? AI reads your resume vs this job — match score, gaps to address, ATS keywords.

SKILL SIGNAL 49 detected · ranked by frequency

LLMs ×3

RL ×3

coding ×3

tool use ×3

computer use ×3

multi-agent coordination ×3

long-horizon execution ×3

factuality ×3

instruction following ×3

calibrated reasoning ×3

data pipelines ×3

reward signals ×3

model-behavior analysis ×3

synthetic data generation ×3

training interventions ×3

large-scale training ×3

experiment velocity ×3

training reliability ×3

training observability ×3

training reproducibility ×3

training cost ×3

training latency ×3

production readiness ×3

multi-agent systems ×3

production-like environments ×3

qualitative behavior analysis ×3

Agent Post-Training ×2

Codex

ChatGPT

API

RLHF

RLAIF

Role Details

Type FULL TIME

Category agents

Salary Band 200k+

AI-Extracted Insights

Domain Areas

agent-post-trainingconnectorsprofessional-softwarecodeapistoolsstructured-integrationsdigital-context

How to Apply on Ashby

Ashby is a fast modern ATS — most applications take under 3 minutes.
The resume parser is strong; verify parsed experience dates and job titles.
Custom screening questions are often scored algorithmically — answer completely.
Location field affects geo-based screening; use your actual metro area.

ANONYMOUS · UNFILTERED

What do employees actually say about OpenAI?

Real rants from real employees. Read before you apply.

Read Company Rants →