Agent Post-Training, Artifacts Research

OpenAI
San Francisco
Workplace: OnsiteFull timeUSD 295,000 - 445,000 annuallyFunction: Education & TrainingSkills: ["Communication","Collaboration","Problem-solving","Analytical thinking","Project management"]

Join OpenAI’s Agent Post-Training, Artifacts team to train frontier models that generate polished artifacts (documents, dashboards, reports, analyses) and advance post-training capabilities across RL, data pipelines, graders, reward signals, and evals. Collaborate with researchers, engineers, product, and safety teams to ship improvements that shape OpenAI’s next-generation agents and their real-world products.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
2 months ago

Agent Post-Training, Artifacts Research

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Join OpenAI’s Agent Post-Training, Artifacts team to train frontier models that generate polished artifacts (documents, dashboards, reports, analyses) and advance post-training capabilities across RL, data pipelines, graders, reward signals, and evals. Collaborate with researchers, engineers, product, and safety teams to ship improvements that shape OpenAI’s next-generation agents and their real-world products.
Location: San Francisco
Workplace: Onsite
Employment Type: Full time
Job Function: Education & Training

Key Responsibilities

  • •Design and run experiments that improve agentic model behavior for complex software and plugins.
  • •Own end-to-end improvements to the post-training stack, including RL, data pipelines, graders, reward signals, evals, diagnostics, and model-behavior analysis.
  • •Build evals and environments that expose the next set of model failures, then turn those failures into training data, product fixes, or new research directions.
  • •Partner with Codex and ChatGPT product teams to understand user needs and translate product signal into model improvements.
  • •Work on early-training and alignment interventions, including data mixtures, objectives, synthetic data, and eval loops that shape downstream agent behavior.

Pay and Benefits

Salary: USD 295,000 - 445,000 annually

Key Requirements

  • •Strong fundamentals in machine learning, software engineering, systems, statistics, or a related field and ability to learn quickly across new areas
Experience:AIMLLLMs
Skills:CommunicationCollaborationProblem-solvingAnalytical thinkingProject management
Tech Stack:RLMachine learningData pipelinesEvaluationsGradersProduction MLMulti-agent systemsCodexChatGPT

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALL·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor