About this role
Qualcomm is hiring an engineer in Hyderabad to help build its enterprise-grade generative AI platform. The team develops agentic applications for use across Qualcomm and exposes model SDKs, frameworks, and application APIs so other teams can build AI solutions. Its systems address low latency, high-throughput scalability, and orchestration of on-premises and cloud-based data.
Responsibilities
- Design and implement retrieval-augmented generation (RAG) solutions that enhance large language model capabilities.
- Develop and optimize LLM fine-tuning for domain-specific use cases, such as wireless and 5G.
- Build and maintain autonomous, multi-step agentic workflows.
- Create evaluation frameworks to measure model performance and guide improvements.
- Collaborate with systems, hardware, architecture, test engineering, and other teams on software solutions and performance requirements.
Qualifications and skills:
- Minimum education and software engineering experience are either a bachelor's degree in Engineering, Information Systems, Computer Science, or a related field with 3+ years of relevant experience; a master's degree in one of those fields with 2+ years; or a PhD in one of those fields with 1+ year.
- The posting also specifies 1–5 years of AI or machine learning engineering experience and 2+ years of academic or work experience with a programming language such as C, C++, Java, or Python.
- The technical stack includes Python, PyTorch or TensorFlow, and vector databases such as Milvus or Pinecone. Experience with LangChain, LlamaIndex, and Transformers is sought.
- A master's degree with an ML specialization is preferred. Knowledge of prompt engineering and reinforcement learning is highly valued.
Skills for this role
Generative AIRetrieval-Augmented GenerationLarge Language ModelsModel fine-tuningAI agentsModel evaluationPythonPyTorchTensorFlowVector databasesMilvusPineconeLangChainLlamaIndexTransformersPrompt engineeringReinforcement learningCC++JavaSoftware engineering