The role Youll own the cost and performance of our inference stack. Your work will determine how efficiently we serve models as workloads, traffic, and hardware change. Youll work closely with the engineers operating the serving
Adaption is seeking a senior ML systems engineer to own the cost and performance of our inference stack in a rapidly evolving environment. You will shape throughput, latency, and reliability by tuning caching, batching, and kernels while