We are looking for a Senior/Lead AI/ML Performance Engineer to benchmark and profile AI/ML workloads on TPU hardware to generate the empirical performance data that feeds the optimization solver's cost model.
Responsibilities
Conduct JAX/XLA benchmarking on physical TPU slices
Capture and analyze TPU Profiler traces
Build hardware coefficient matrices for use in the optimizer's cost model