About this role
Humana seeks a Senior AI Engineer to build and operate production AI systems that turn millions of clinical documents into trusted, actionable data. These systems use LLMs to extract structured facts from medical records, answer questions with source citations, and route ambiguous cases to human experts. The role owns solutions end-to-end, with particular emphasis on accuracy, reliability, traceability and scalability in a regulated healthcare environment.
Responsibilities
- Design and deploy full-stack AI applications, including frontend experiences, APIs, backend services and data pipelines. Build LLM workflows for extraction, classification, retrieval, summarization and agentic processes, using prompts, structured outputs, tool calling and fallback mechanisms.
- Implement RAG with embeddings, vector databases and semantic search. Build document ingestion, OCR, metadata extraction, indexing and search pipelines.
- Create benchmark datasets, regression tests, quality metrics and human-review workflows so clinicians and domain experts can validate outputs and provide feedback. Improve accuracy, latency, reliability, scalability and cost efficiency.
- Operate cloud-native services with monitoring, logging, alerting and incident response. Collaborate with product, engineering, clinical and operational teams while meeting privacy, security, compliance and auditability requirements.
Required qualifications
A bachelor's degree in Computer Science, Engineering or a related technical field, or equivalent practical experience; 5+ years building and operating production software systems; experience with backend services, APIs, distributed systems or data-intensive applications; and hands-on integration of LLMs into production applications beyond simple chat interfaces. Candidates need experience with at least one major AI platform or model provider, strong Python and/or TypeScript/JavaScript skills, production debugging and support experience, and the ability to work independently through ambiguity.
Preferred qualifications include production-critical LLM systems; evaluation, testing, prompt versioning, tracing and model monitoring; agentic architecture or MCP; React, Next.js and full-stack development; Kubernetes, Docker, cloud deployment and CI/CD; regulated-industry experience; and familiarity with data protection and AI-assisted development. The listed stack also includes PostgreSQL and Gemini on Vertex AI.
This full-time, 40-hour-per-week role is hybrid: three office days weekly, with remaining days remote. Candidates must live within, or be willing to relocate within, commuting distance of a listed talent market. Starting base pay is estimated at $141,100–$194,000 annually and may vary by location and qualifications. The role is eligible for a performance-based bonus incentive plan; listed benefits include health coverage, a 401(k), paid leave, disability coverage and life insurance.