Sr. Engineering Manager, Inference
CoreWeave
- Location
- Livingston, NJ / New York, NY / San Francisco, CA / Sunnyvale, CA / Bellevue, WA / Remote - US
- Work model
- Remote
- Level
- Senior
- Salary
- $188k – $303k/yr
- Posted
- 2h ago
Skills
About this role
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at www.coreweave.com.
What You'll Do
As a Senior Engineering Manager for the AI/ML Platform team, you will lead the group responsible for productizing and operating our inference offering. In partnership with CoreWeave AI platform team, you will ensure the service becomes a polished, reliable, developer-friendly product within the AI/ML Platform.
You will guide engineers working on areas such as service reliability, observability, operational excellence, packaging, developer-facing tooling, and application-layer enhancements. You will partner closely with Product, CoreWeave engineering teams, Design, Support, and GTM stakeholders to ensure the inference experience meets the needs of practitioners deploying and scaling real-world AI workloads. Your work will directly support end-users, as well as other products powered by this service, like W&B Training.
You will combine strong engineering leadership with clarity in execution, helping the team deliver improvements that raise reliability, accelerate development workflows, and strengthen the overall inference experience for W&B users.
Responsibilities
• Lead and grow the engineering team responsible for evolving and operating the W&B Inference product, focusing on service reliability, orchestration, operational maturity, and developer experience.
• Drive execution on roadmap initiatives in close partnership with Product, ensuring that platform capabilities are delivered predictably, robustly, and with measurable customer impact.
• Own the engineering processes for the inference productization layer, including incident response, operational readiness, release management, observability, and engineering quality.
• Partner with CoreWeave’s infrastructure teams to integrate capabilities from the underlying inference platform into a cohesive, user-facing product.
• Guide the design and delivery of application-layer enhancements such as tracing and tool call handling.
• Ensure the inference offering meets high standards of reliability, usability, compliance, and performance, balancing trade-offs across cost, operational load, and architectural constraints.
• Create clarity in complex, cross-functional environments, ensuring strong communication, aligned priorities, and smooth execution across multiple teams.
• <span