RefolkCandidates
Open nowScalingCompute

Performance Modeling Lead

OpenAI · San Francisco

Location
San Francisco
Workplace
Remote
Employment
Full time
Level
Mid level
Posted
5 months ago

About this role

About the Team

OpenAI’s Hardware organization develops system and infrastructure solutions designed for the unique demands of advanced AI workloads. We work closely with research, software, and external hardware partners to shape the next generation of AI systems, from silicon through full-scale deployments.

Our team focuses on understanding and optimizing performance across the full system stack - ensuring that architectural decisions are grounded in rigorous, quantitative analysis of real-world workloads.

About the Role

We are seeking a Performance Modeling Lead to build and lead a small, high-impact team responsible for answering forward-looking architectural questions across AI infrastructure systems.

You will develop modeling frameworks and methodologies to evaluate system-level tradeoffs and guide key design decisions. Your work will directly influence reference architectures, vendor designs, and long-term infrastructure strategy.

This role sits at the intersection of AI workloads, system architecture, and quantitative modeling, and requires strong technical judgment, ownership, and the ability to translate complex analysis into clear, actionable guidance.

This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance.

Key Responsibilities

  • Build and own a performance modeling framework/toolchain to evaluate AI systems across multiple levels of abstraction.

  • Analyze and quantify architectural tradeoffs across compute, memory, networking, storage, and system topology.

  • Develop performance models to guide decisions on:

    • scale-up vs. scale-out architectures

    • interconnect and network design

    • memory hierarchy and system balance.

  • Translate modeling outputs into clear recommendations for internal teams and external hardware vendors.

  • Influence reference designs and vendor roadmaps through data-driven insights.

  • Partner closely with machine learning, systems, and hardware teams to understand workload characteristics and requirements.

  • Lead and grow a small team (2 - 3 engineers), setting technical direction and maintaining high standards for modeling rigor.

  • Continuously improve modeling fidelity by validating against real system behavior and measurements.

  • Qualifications

    • Have experience owning or building performance modeling frameworks used to drive real system design decisions.

    • Have deep knowledge of AI/ML workloads, including training and/or inference at scale.

    • Understand system-level tradeoffs across compute, memory, and networking in large-scale distributed systems.

    • Are comfortable working across abstraction layers - from workload behavior to hardware implementation.

    • Have experience using modeling (analytical or simulation) to inform architectural decisions.

    • Can operate in ambiguous problem spaces and turn open-ended questions into structured analysis.

    • Communicate clearly and influence both internal teams and external partners.

    Preferred Skills

    • Experience working with hardware vendors (ODM/JDM, silicon, networking).

    • Background in data center infrastructure or hyperscale systems.

    • Familiarity with accelerators (GPUs/ASICs) and interconnects (e.g., NVLink, InfiniBand, Ethernet).

    • Experience influencing hardware roadmaps or reference architectures.

    • Prior experience leading or mentoring engineers.

    About OpenAI

    OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.

    We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

    For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement.

    Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.

    To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance.

    We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link.

    OpenAI Global Applicant Privacy Policy

    At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

    As published by OpenAI. Applications are handled on their site.

    Skills this posting mentions

    Artificial IntelligenceODMMentorship

    About OpenAI

    OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. AI is an extremely powerful tool that must be created with safety and human needs at its core. OpenAI is dedicated to putting that alignment of interests first - ahead of profit. To achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. Our investment in diversity, equity, and inclusion is ongoing, executed through a wide range of initiatives, and championed and supported by leadership. At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

    All 877 openings at OpenAI

    One click, then it is written

    Apply to OpenAI with a resume written for this role.

    Queue Performance Modeling Lead and I read the posting, rewrite your resume against it, draft the cover letter, and score the fit. Then you press send, or press one button and I fill in OpenAI’s form for you.

    1. 01Drop your resume

      A PDF or a LinkedIn URL. About a minute, once.

    2. 02I rank the openings

      Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.

    3. 03Each one is written up

      Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.

    • 25 sent a week, free
    • No card
    • Nothing sent until you say so

    More roles at OpenAI

    See all

    Put this to work

    Paste your career in once. Every application after that is written for you.

    Drop a resume or a LinkedIn URL. I rank the live openings against it, rewrite the resume and write a cover letter for the best of them, and fill in the employer's form when you press the button. You read, you decide what goes out.

    1. 01Drop your resume

      A PDF or a LinkedIn URL. About a minute, once.

    2. 02I rank the openings

      Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.

    3. 03Each one is written up

      Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.

    • New matches ranked and written before you are up.
    • Every bullet stays inside what your history supports. Nothing invented.
    • Queued, submitted, interviewing, offer: one screen, not a spreadsheet.

    500 free credits on sign-up. No card. Nothing is sent until you say so.

    Listed from the job board OpenAI publishes. Refolk is not the employer and does not handle their hiring. Applications go to OpenAI directly.