
Empowering corporate mentorship for effective learning
Together is a corporate mentorship management platform founded in 2018, headquartered in CityPlace, Toronto, ON. The platform streamlines the mentorship lifecycle, facilitating connections among employees at companies like Heineken, Reddit, and 7-Eleven. With $1.7 million in seed funding, Together a...
Together offers competitive salaries and equity packages, 4 weeks of paid vacation, and a comprehensive health, dental, and vision plan through Honeyb...
Together fosters a culture of autonomy and impact, allowing employees to take on significant responsibilities without bureaucratic constraints. The fo...

Together AI • San Francisco
Together AI is seeking a Machine Learning Engineer to optimize AI inference systems. You'll work with large language models and collaborate with researchers and engineers. This role requires proficiency in Python and PyTorch.
You have 3+ years of experience writing high-performance, well-tested, production-quality code — you've developed systems that are reliable and efficient at scale. Your proficiency with Python and PyTorch allows you to build and optimize runtime inference services for large-scale AI applications. You possess a strong understanding of low-level operating systems concepts including multi-threading, memory management, and performance optimization — these skills enable you to create robust and fault-tolerant systems for data ingestion and processing. You thrive in collaborative environments, working closely with AI researchers and engineers to bring innovative features and capabilities to life. You are detail-oriented and conduct design and code reviews to ensure high standards of quality in your work. You are passionate about AI inference and are eager to contribute to cutting-edge AI solutions.
Knowledge of additional AI frameworks or libraries would be a plus, as would experience in building high-performance libraries and tooling. Familiarity with large language models and their applications in real-world scenarios can enhance your contributions to the team.
In this role, you will design and build the production systems that power the Together AI inference engine, ensuring they run efficiently and effectively at scale. You will develop and optimize runtime inference services, collaborating with a diverse team of researchers, engineers, product managers, and designers to implement new features and research capabilities. Your responsibilities will include conducting design and code reviews, creating services and tools to support the inference engine, and documenting your work for developers. You will implement robust systems for data ingestion and processing, ensuring that the AI inference systems are reliable and performant. Your contributions will directly impact the performance of AI applications, shaping the future of AI solutions at Together AI.
Together AI provides an opportunity to work in a dynamic environment where innovation is encouraged. You will have the chance to collaborate with talented professionals in the AI field, contributing to projects that push the boundaries of technology. We offer competitive compensation and a supportive culture that values your contributions and growth. Join us in shaping the future of AI and making a meaningful impact in the industry.
Apply now or save it for later. Get alerts for similar jobs at Together AI.