About this role
The role involves developing and deploying Generative AI applications using LLMs, building LLM pipelines, and implementing LLMOps practices. You will design end-to-end AI/ML workflows, work with vector databases for RAG architectures, and optimize model performance and latency.
Skills for this role
PythonGenerative AILLMsGPTLlamaPrompt EngineeringRAGVector DatabasesFAISSPineconeWeaviateChromaPyTorchTensorFlowFastAPIFlaskRESTful servicesMicroservices