About this role
Research Fellow to integrate multimodal LLMs with embodied AI by designing and implementing Vision-Language-Action and world models, training and fine-tuning models with PyTorch on multimodal datasets, and deploying these models on robotic platforms (particularly robotic arms) for autonomous task execution and human-robot interaction. The role involves collaboration with a multidisciplinary team and publishing high-quality research in top-tier conferences and journals.
Skills for this role
Vision-Language-Action modelsWorld modelsDeep learningMultimodal LLMsPyTorchPythonRobotic systemsRobotic armsHuman-robot interactionReinforcement learningMultimodal datasetsAcademic publishing