New opportunity

AI Inference Engineer

Baseten · San Francisco, United States

About this role

Partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten’s platform, owning the end-to-end journey from exploration to production deployment. This hands-on engineering role involves coding (preferably Python), optimizing AI/ML projects, and working across product, performance, and customer-facing implementations.

Skills for this role

PythonMachine LearningAI/ML pipelinesModel deploymentSoftware developmentProduction monitoringPerformance engineeringProduct managementTechnical customer successSolution engineeringDockerComfyUIWhisperModel evaluationObservabilityCommunicationProject management