Lead AI Engineer, Armada
2023 — presentSeattle, WA · first AI hire
I own the AI stack for Armada's edge platform — model post-training and agent-harness engineering, from NVIDIA H100/A100 GPUs down to CPU-only nodes. The work has moved through three generations:
- Conversational AI assistant
- Built the platform's core assistant from scratch on open-weight Llama models with an LLM-as-router design; deployed end-to-end on AKS with vLLM, and quantized for on-device edge inference.
- Multi-agent systems
- Designed a production supervisor–sub-agent system with modular inference routing; hardened against prompt injection, with observability and evaluation built in.
- Skill-agent harness & coding agents
- Re-architected the assistant as a composable skill + coding agent — deep agents, code sandboxing, and Armada-specific asset skills — alongside AI-safety research on backdoors in tool-using LLMs.