Build automation and monitoring systems for OpenAI's large-scale compute fleet.
You'll build automation and monitoring systems for large-scale compute infrastructure, working with Python, Go, and Linux in a full-time onsite role based in San Francisco. The position involves developing monitoring solutions using Prometheus, Grafana, and PromQL, along with data analysis using SQL and Pandas. This is a permanent, mid-level engineering position requiring fluency in English.
Membership is €29/month, cancel anytime: every rate, every original listing link, and a daily alert for roles matching your filters.
Found at a specialist agency · listed 27 July 2026 · InsideJobs links you to the original posting.