What you’ll work on
- Design and build agents that can plan, use tools and complete multi-step tasks in production security workflows.
- Make agent workflows reliable with error handling, retries and human approval for sensitive actions.
- Build tool-calling workflows that choose and configure security tools using the context available.
- Develop prompts and agent workflows that can work across OpenAI, Google Vertex AI and Anthropic Claude.
- Build and maintain the context layer that gives agents awareness of users, assets, past incidents, typical behavior, and environmental state.
- Define practical evaluations for reliability, accuracy and safety, especially for exploit execution and incident response recommendations.
- Collaborate with security engineers to translate offensive and defensive domain expertise into agent behavior, tool profiles, and decision logic.
- Optimize for latency, cost, and token efficiency in production agent workloads.
- Support on-premise deployments using self-hosted open-source models (DeepSeek, Llama) for air-gapped enterprise customers.
