Software Engineer, AI accelerator Runtime
OpenAIGenerative AI company
San Francisco, United StatesSenior
Microsoft
Nvidia
Thrive Capital
Sequoia Capital
SoftBank
Amazon
Software EngineeringNew
About the role
TL;DR
Build low-level device runtime for custom AI accelerators.
- •You will build the low-level device runtime that turns compiled programs into efficient, functional and performant execution on OpenAI’s custom AI accelerator.
- •This software will schedule kernel launches, manage device memory and address spaces, coordinate synchronization, and expose reliable abstractions to higher-level runtimes and frameworks.
- •Key Responsibilities Design and implement the low-level device runtime for OpenAI custom silicon.
- •Build kernel-launch scheduling, command submission, queueing, dependency tracking, and completion handling.
- •Manage device memory spaces, allocation, virtual-to-physical mappings, data movement, and lifetime across concurrent workloads.
- •Implement synchronization primitives, events, barriers, streams, and ordering guarantees that are correct and efficient.
- •Use event-based, cycle-accurate simulators to develop, validate, debug, and performance-tune runtime behavior.
- •Requirements Strong low-level systems programming experience in C, C++, Rust, or comparable environments.
- •Experience building runtimes, drivers, firmware, operating-system components, accelerator software, or adjacent systems infrastructure.
- •Understanding of concurrency, synchronization, asynchronous execution, queues, events, and memory-ordering semantics.
- •Hands-on experience with event-based, cycle-accurate simulators or closely related architectural and performance models.
- •Skilled at debugging failures that span software abstractions, device interfaces, and hardware behavior.
Required skills
CC++RustLinuxGit
Domain expertise
deeptechai
Tech stack
CC++RustLinuxGit