Applied AI Engineer, Inference

AI EngineerFull-timeMid-level · 4+ yearsBellevue, San Francisco

You will build and maintain benchmarking workflows to measure and improve the performance of the inference platform, focusing on latency, throughput, and cost. This role requires 4+ years of experience in machine learning, systems, or performance engineering, along with strong Python programming skills. You must have familiarity with LLM inference systems and the ability to translate experimental results into concrete engineering decisions.

Be the first to hear about new roles at CoreWeave

Follow this company and we'll notify you as soon as new roles are posted.

Start free
Open