Sr Staff AI Engineer - Veza

ServiceNow · Minneapolis, Minnesota, us

remotefull-time6-10 years

posted 1d

The Access AI team is the customer-obsessed engineering group building the agentic AI and enterprise-scale harness platform that powers agents for Access Management across all Identity Security Products. We build AI as foundational platform infrastructure — prioritizing robustness, performance, safety, and real-world customer impact at scale.  Types of problems you’ll get to work on  You will architect, build, and operate production-grade agentic AI systems—autonomous agents that reason over enterprise data and execute mission-critical identity security actions at Fortune 500 scale. As a tier-one technical leader at the intersection of Agent research and large-scale backend systems, you will shape the architectural patterns, scalability guardrails, and strategic vision for autonomous enterprise security. Your core focus areas:  Agentic architecture. Design and ship multi-agent systems — orchestration, tool use, planning loops, memory, and failure recovery — that operate reliably in production, not in notebooks.  Enterprise-grounded reasoning. Build agents that leverage Access Graph, Access Reviews, and permission and risk data — to make decisions with context no frontier model has on its own.  Trust, safety, and governance. Own the guardrails: observability, human-in-the-loop controls, and compliance infrastructure that make autonomous systems safe to deploy at scale.  Retrieval and grounding. Work closely with our different product teams, platform and graph teams to ensure agents are grounded in accurate, low-latency retrieval — RAG pipelines, semantic search, re-ranking, and evaluation — as a critical dependency of agentic quality.  Model integration and evaluation. Integrate frontier models, evaluate trade-offs across cost, latency, and capability for production use cases.  Engineering leadership. Raise the technical bar through architecture decisions, code reviews, and coaching — particularly on agentic design patterns and production AI discipline. Qualifications To be successful in this role you have: 8+ years of software engineering with strong fundamentals in data structures, algorithms, and distributed systems.  Hands-on depth designing, shipping, and operating agentic systems in production — multi-agent orchestration, tool calling, planning loops, memory, and failure recovery. Not prototypes.  Production-grade Python. Systems language (Go, Java, or C++) is a plus.  Working experience with frontier AI SDKs (Anthropic, Google, or OpenAI) — prompt engineering, structured outputs, and model evaluation in production settings.  Familiarity with RAG and retrieval patterns in production — vector stores, hybrid search, and retrieval evaluation metrics.  Track record of technical leadership: architecture ownership, code quality bar-raising, and mentoring engineers on production AI practices.  Nice to Have  Deeper specialization in search and retrieval at scale or MLOps/model observability.  Published work or open-source contributions in agentic systems or retrieval.  Exposure to LLM fine-tuning or inference optimization in production. Why join us  Intelligence is commoditizing. Context and execution are not. With over 30 billion access permissions under management, global enterprises in Fortune 500 trust Veza to manage privileged access monitoring, non-human identity security, access entitlement management, and next-generation identity governance, we are building the system that makes AI actually work inside the enterprise to secure Enterprises. Redefining Agentic Identity Governance: Agentic Search: Provides natural language, context-aware discovery capabilities across complex permission graphs, enabling rapid identification of access risks and opportunities. Agentic Access Reviews: Streamlines certification processes by using autonomous agents to analyze risk and suggest remediation, significantly reducing reviewer fatigue and preventing