Senior AI/ML Engineer - Agentic AI
SAP
- Location
- Palo Alto, CA, US, 94304
- Work model
- On-Site
- Level
- Senior
Skills
About this role
We help the world run better At SAP, we keep it simple: you bring your best to us, and we'll bring out the best in you. We're builders touching over 20 industries and 80% of global commerce, and we need your unique talents to help shape what's next. The work is challenging – but it matters. You'll find a place where you can be yourself, prioritize your wellbeing, and truly belong. What's in it for you? Constant learning, skill growth, great benefits, and a team that wants you to grow and succeed.
Summary
We're building specialized foundation models and AI agents that accelerate SAP customers' data transformation journeys.
The agents you build will directly power SAP's Autonomous Enterprise, where AI runs core business processes end-to-end across finance, supply chain, HR, and procurement at global scale.
You'll set technical direction, define how we architect and scale multi-agent systems, and raise the engineering bar across a global team. You'll work directly with pretraining and fine-tuning team leads in Europe, India, and early-adopter customers.
We want someone who has shipped agentic systems in production and knows where they break.
What you'll do
Architect and lead multi-agent systems: design, orchestration patterns, failure modes, memory, planning, and human-in-the-loop
Own the path from prototype to production: containerization, guardrails, cost and latency optimization, scalable serving
Define the team's evaluation strategy: offline/online harnesses, trajectory quality, tool-call accuracy, regression testing, CI/CD eval gates
Lead instrumentation and observability: tracing, span capture, automated scoring, closing the trace → eval → fix loop
Drive tool integration architecture via MCP across multiple product teams
Mentor junior and mid-level engineers through code and architecture reviews; set engineering standards
What you bring
Education: BS, MS, or PhD in Computer Science, ML, or a related field.
Experience: 6+ years building and shipping ML systems, with 3+ years hands-on with LLMs and agents in production.
Core technical skills
Expert Python; strong fundamentals: system design, testing, modularity, async, API design
PyTorch; working knowledge of fine-tuning and PEFT methods (LoRA, QLoRA)
LLM application development: prompting, structured outputs, tool calling, context management
Inference optimization: vLLM, TensorRT-LLM, quantization (int8, int4, GPTQ, AWQ)
Human-in-the-loop annotation workflows at scale
Agent frameworks and orchestration (production experience with at least three)
LangGraph / LangChain
CrewAI, AutoGen/AG2, or equivalent
MCP (Model Context Protocol)
Coding agents: Claude Code, OpenCode, or similar
Evaluation and observability (production experience with at least two)
Langfuse, LangSmith, Arize, or equivalent
LLM-as-judge evaluators, CI/CD eval gates
Leadership
Demonstrated track record mentoring engineers and raising team technical quality
Drives decisions in ambiguous, fast-moving environments
Writes design docs that earn buy-in across teams
Nice to have
Continued pretraining or fine-tuning pipelines (SFT, DPO, RLHF)
Meet your team
The Foundation Model team is part of the Generative AI foundation within Business AI at SAP. We build domain-specialized foundation models on SAP business data through continued pretraining and fine-tuning, and we build the agentic layer that brings them into SAP products. Bring out your best SAP innovations help more than four hundred thousand customers worldwide work together more efficiently and use business insight more effectively. Originally known for leadership in enterprise resource planning (ERP) software, SAP has evolved to become a market leader in end-to-end business application software and related services for database, analytics, intelligent technologies, and experience management. As a cloud company with two hundred million users and more than one hundred thousand employees