RefolkCandidates

ML Research Engineer - Hardware Codesign

OpenAI · San Francisco

$185k - $455k/yrMid levelEngineeringScalingHardwarePosted 6 months ago

About this role

ABOUT THE TEAM OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. ABOUT THE ROLE We’re seeking a Research-Hardware Codesign Engineer to operate at the boundary between model research and silicon/system architecture. You’ll help shape the numerics, architecture, and technology bets of future OpenAI silicon in collaboration with both Research and Hardware. Your work will include debugging gaps between rooflines and reality, writing quantization kernels, derisking numerics via model evals, quantifying system architecture tradeoffs, and implementing novel numeric RTL. This is a hands-on role for people who go looking for hard problems, get to ground truth, and drive it to production. Strong prioritization and clear, honest communication are essential. Location: San Francisco, CA (Hybrid: 3 days/week onsite) Relocation assistance available. IN THIS ROLE YOU WILL: - Build on our roofline simulator to track evolving workloads, and deliver analyses that quantify the impact of system architecture decisions and support technology pathfinding. - Debug gaps between performance simulation and real measurements; clearly communicate root cause, bottlenecks, and invalid assumptions. - Write emulation kernels for low-precision numerics and lossy compression schemes, and get Research the information they need to trade efficiency with model quality. - Prototype numerics modules by pushing RTL through synthesis; hand off novel numerics cleanly, or occasionally own an RTL module end-to-end. - Proactively pull in new ML workloads, prototype them with rooflines a

Excerpt from the posting OpenAI published. Read the full description on their site before applying.

Skills this posting mentions

Machine LearningCUDASystem Architecture

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. AI is an extremely powerful tool that must be created with safety and human needs at its core. OpenAI is dedicated to putting that alignment of interests first - ahead of profit. To achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. Our investment in diversity, equity, and inclusion is ongoing, executed through a wide range of initiatives, and championed and supported by leadership. At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

All 755 openings at OpenAI

Apply to this role, tailored

Queue ML Research Engineer - Hardware Codesign at OpenAI and I will read the posting, rewrite your resume against it, draft the cover letter, and score the fit before you send anything.

Applications finish as drafts. Nothing is sent until you read it and press send. New accounts start with 500 free credits.

More roles at OpenAI

See all

Similar roles elsewhere

See more

Listed from the job board OpenAI publishes. Refolk is not the employer and does not handle their hiring. Applications go to OpenAI directly.