Data Scientist, Inference Capacity Optimization

OpenAI · San Francisco

Company
OpenAI
Location
San Francisco
Salary
$293,000–$325,000 USD /yr
Employment
Full-time
Workplace
Hybrid
Last updated
2026-07-27
Source
Official Ashby feed

About this role

About the Role

OpenAI’s Industrial Compute organization is responsible for ensuring our compute infrastructure scales efficiently to support millions of users and increasingly sophisticated AI models.

We’re looking for a Data Scientist to partner closely with Capacity Systems Engineering, Infrastructure, Product, and Research to optimize inference capacity across our global GPU fleet. This role combines statistical modeling, large-scale data analysis, forecasting, and systems thinking to drive critical decisions around infrastructure investments, performance-efficiency trade-offs, and customer experience.

You’ll transform complex operational data into actionable insights that directly influence how OpenAI allocates and scales one of the world’s largest AI compute environments.

Key Responsibilities

- Build statistical and machine learning models to profile and improve GPU utilization, latency, throughput, and overall fleet efficiency.

- Develop forecasting models for inference demand across products, regions, and model families.

- Analyze production workloads to identify latency bottlenecks and capacity constraints, highlighting optimization opportunities.…

Summary from OpenAI's official Ashby career feed — read the full description on the original posting

Apply on OpenAI's site

The salary and details above come straight from OpenAI's official Ashby feed. The full job description lives on OpenAI's careers page — AI Stack Jobs links you straight to it, we never sit between you and the employer.

More AI/ML roles at OpenAI

  1. 2026-07-27 San Francisco
  2. 2026-07-27 San Francisco
  3. 2026-07-25 San Francisco
  4. 2026-07-24 San Francisco
  5. 2026-07-24 San Francisco
  6. 2026-07-24 New York City