Build and optimize AI inference stack for AWS Trainium accelerators.
Develop high-performance kernels, compilers, and runtime systems to execute cutting-edge models efficiently on AWS Trainium hardware. This cross-stack systems role requires deep expertise in performance optimization across hardware, kernels, compilers, and ML frameworks. You'll collaborate with inference and ML teams to identify bottlenecks and deploy production solutions that unlock Trainium's full capabilities.
Membership is €29/month, cancel anytime: every rate, every original listing link, and a daily alert for roles matching your filters.
Found at a specialist agency · listed 20 August 2026 · InsideJobs links you to the original posting.