Search / adaption

Inference Performance Engineer

adaption· San Francisco

hybrid seen live 1 day ago

The role You'll own the cost and performance of our inference stack. Your work will determine how efficiently we serve models as workloads, traffic, and hardware change. You'll work closely with the engineers operating the serving fleet while owning the core performance levers:

Apply on adaption's site

This link goes straight to the employer's ashby page. Posted 10 days ago.We last confirmed it was open 1 day ago.