Mediabistro logo
job logo

LLM Inference Architect & Systems Optimizer

Together AI · San Francisco, CA, USA ·

Pay:
$160,000-$230,000/yr
Job type:
Full Time

Togetherai in San Francisco seeks an Inference Frameworks and Optimization Engineer to design scalable inference engines for large language models. As part of a dynamic team, you will work on optimizing performance and collaborating closely with hardware teams.
The role requires a strong background in deep learning frameworks and programming with Python and C++. Competitive compensation includes a salary between $160,000 - $230,000 plus equity.

#J-18808-Ljbffr