About this role
Research and develop novel methods to improve agent reasoning, planning, and multi-step exploitation for autonomous LLM-based agents; design and run experiments, build evaluation harnesses and benchmarks, and translate research into production improvements for an AI-powered penetration testing platform.
Skills for this role
Large Language Models (LLMs)Agentic systemsPromptingTool useFine-tuningReinforcement Learning (RL)Evaluation designExperimental designBenchmarkingPythonML/AI researchCybersecurityWeb vulnerability discoveryOffensive securityServing LLMs at scaleProductionizing research prototypes