Build low-level device runtime for custom AI accelerator silicon.
Design and implement device runtime software that executes compiled programs efficiently on OpenAI's custom AI accelerator hardware. This involves kernel scheduling, memory management, synchronization primitives, and interfaces between drivers, firmware, and higher-level frameworks. The role requires deep systems programming expertise, particularly in concurrency and memory ordering, with hands-on experience using cycle-accurate simulators for validation before silicon availability.
Membership is €29/month, cancel anytime: every rate, every original listing link, and a daily alert for roles matching your filters.
Found at a specialist agency · listed 25 August 2026 · InsideJobs links you to the original posting.