About this role
About the Role Moveworks is seeking a Machine Learning Engineer to help build cutting edge ML infrastructure for building and serving LLMs at Moveworks. This role will be critical in building, optimizing and scaling end-to-end machine learning systems. The ML infra team covers a variety of responsibilities including distributed training and inference pipeline for large language models (LLMs), model evaluation and monitoring framework, LLM latency optimization, etc. These frameworks serve as a strong foundation for our hundreds of ML and NLP models in production. What You'll Do
- Build, optimize, and scale end-to-end ML systems and infrastructure for LLMs at Moveworks.
- Develop distributed training and inference pipelines for large language models.
- Create and maintain model evaluation and monitoring frameworks.
- Optimize LLM latency and throughput to meet production SLAs.
- Collaborate with the ML Infra team to support hundreds of ML and NLP models in production.
- Work with cross-functional teams to deploy reliable AI-powered workflows across the platform. What We're Looking For
- Experience building and maintaining ML infrastructure for LLMs and NLP models.
- Strong background in distributed training and scalable inference pipelines.
- Experience with model evaluation, monitoring, and latency optimization.
- Familiarity with production ML systems and end-to-end ML lifecycle.
- Ability to collaborate across teams and communicate effectively.