Machine Learning Engineer - Inference

Together AI • San Francisco

Posted 2w agoMid-Level Machine learning engineer 📍 San francisco

Apply Now →

Skills & Technologies

Python PyTorch

Overview

Together AI is seeking a Machine Learning Engineer to optimize AI inference systems. You'll work with large language models and collaborate with researchers and engineers. This role requires proficiency in Python and PyTorch.

Job Description

Who you are

You have 3+ years of experience writing high-performance, well-tested, production-quality code — you've developed systems that are reliable and efficient at scale. Your proficiency with Python and PyTorch allows you to build and optimize runtime inference services for large-scale AI applications. You possess a strong understanding of low-level operating systems concepts including multi-threading, memory management, and performance optimization — these skills enable you to create robust and fault-tolerant systems for data ingestion and processing. You thrive in collaborative environments, working closely with AI researchers and engineers to bring innovative features and capabilities to life. You are detail-oriented and conduct design and code reviews to ensure high standards of quality in your work. You are passionate about AI inference and are eager to contribute to cutting-edge AI solutions.

Desirable

Knowledge of additional AI frameworks or libraries would be a plus, as would experience in building high-performance libraries and tooling. Familiarity with large language models and their applications in real-world scenarios can enhance your contributions to the team.

What you'll do

In this role, you will design and build the production systems that power the Together AI inference engine, ensuring they run efficiently and effectively at scale. You will develop and optimize runtime inference services, collaborating with a diverse team of researchers, engineers, product managers, and designers to implement new features and research capabilities. Your responsibilities will include conducting design and code reviews, creating services and tools to support the inference engine, and documenting your work for developers. You will implement robust systems for data ingestion and processing, ensuring that the AI inference systems are reliable and performant. Your contributions will directly impact the performance of AI applications, shaping the future of AI solutions at Together AI.

What we offer

Together AI provides an opportunity to work in a dynamic environment where innovation is encouraged. You will have the chance to collaborate with talented professionals in the AI field, contributing to projects that push the boundaries of technology. We offer competitive compensation and a supportive culture that values your contributions and growth. Join us in shaping the future of AI and making a meaningful impact in the industry.

Interested in this role?

Apply now or save it for later. Get alerts for similar jobs at Together AI.

Apply Now →Get Job Alerts

About Together AI

Key Highlights

🎁 Benefits

🌟 Culture