About this role
Contribute to development of a next-generation multimodal LLM stack that combines speech, text, tools, and real-time reasoning, building conversational AI models from research through production serving millions of calls per day.
Skills for this role
Large Language Models (LLMs)Multimodal modelsSpeech-language systemsPromptingFine-tuningAlignment techniquesNeural audio codecsStreaming audioReal-time reasoningTool executionDataset creationExperiment designConversational AIAgent frameworksMultimodal datasetsProduction deploymentLatency optimizationSystems thinking