Inference Performance Engineer
adaption· San Francisco
hybrid seen live 1 day ago
The role You'll own the cost and performance of our inference stack. Your work will determine how efficiently we serve models as workloads, traffic, and hardware change. You'll work closely with the engineers operating the serving fleet while owning the core performance levers:
This link goes straight to the employer's ashby page. Posted 10 days ago.We last confirmed it was open 1 day ago.