About this role
NYU Langone Health seeks an Engineer II, Gen AI, for NYU Grossman School of Medicine’s Remote Patient Monitoring (RPM) initiatives. The role designs, deploys, operates, and improves production generative AI and machine learning solutions for clinical workflows, patient engagement, and operational use. The position is listed in New York, NY, as full-time, with a day shift scheduled 9am–5pm.
Responsibilities
- Build MLOps and LLMOps pipelines for production large language models, and collaborate with data scientists on models for uses such as summarization, triage, and patient messaging.
- Develop data and ML pipelines covering ingestion, preprocessing, validation, training, evaluation, and deployment. Monitor performance, latency, drift, safety, and data quality; maintain model versioning and experiment tracking.
- Evaluate RAG, vector databases, embeddings, and LLM providers against clinical needs, compliance, performance, and cost. Improve inference efficiency using quantization, batching, caching, and resource allocation; use Docker and Kubernetes for scalable deployments.
- Apply security and data-protection standards, including HIPAA requirements for protected health information. Implement CI/CD, automated testing, deployment, and rollback; prepare design documents, runbooks, and operational procedures.
- Work with clinicians, care teams, product stakeholders, data scientists, and IT on requirements, backlog items, reviews, and delivery. Participate throughout the AI software development lifecycle and coach less experienced colleagues.
Required: A bachelor’s degree in computer science, data science, biomedical informatics, software engineering, or a related field; at least 1–3 years of AI solution development experience; strong Python or comparable AI programming skills; knowledge of AI, machine learning, and deep learning; experience with PyTorch or TensorFlow and large-scale or compute-intensive applications using clusters such as HPC, Spark, or Kubernetes. Software engineering, problem-solving, teamwork, and communication skills are also required.
Preferred
A relevant master’s degree; the posting states '35 years' of hands-on production AI experience, including LLM applications; Azure or AWS, cloud-native AI/ML tools, CI/CD, RAG, semantic search, embeddings, vector databases, and MLOps/LLMOps tools. Healthcare AI, EHR integration, HIPAA-regulated environments, prompt engineering, fine-tuning, and LLM provider APIs are also preferred. The stated salary range is $94,577.93–$110,000 annually, excluding bonuses, incentives, differential pay, and other compensation or benefits.