About this role
Work within the Applied AI team to build and deploy scalable AI and LLM-driven features (including Ask iManage), owning the end-to-end ML lifecycle from model development and evaluation to production serving. Deploy and optimize ML/AI systems on GPUs and Kubernetes-based cloud infrastructure, and design production-ready LLM applications and APIs with monitoring, observability, and integration testing.
Skills for this role
PythonPyTorchHugging FaceLarge Language Models (LLMs)Generative AILanguage model fine-tuningGPU optimizationKubernetesAKSAzureAWSGCPContainerizationCI/CDModel versioningMonitoringObservabilityAPIsIntegration testingDistributed trainingPyTorch DistributedRayvLLMSGLangLangChainLlamaIndexKnowledge graphsMultimodal LLMsModel lifecycle managementAgentic engineering