Machine Learning Engineer - Inference
Together AI is an AI-native cloud company building infrastructure for AI engineers and application developers. The Machine Learning Engineer will optimize and enhance large-scale AI inference systems, developing production services, tools, and fault-tolerant data processing systems while collaborating with researchers and engineers.ResponsibilitiesDesign and build the production systems that power the Together AI inference engine, enabling reliability and performance at scaleDevelop and optimize runtime inference services for large-scale AI applicationsCollaborate with researchers, engineers, product managers, and designers to bring new features and research capabilities to the worldConduct design and code reviews to ensure high standards of qualityCreate services, tools, and developer documentation to support the inference engineImplement robust and fault-tolerant systems for data ingestion and processing