JobBobsReal-time global job discoveryLive

Technical Program Manager, Inference

Coreweave · Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA · 2026-07-27

executive
Apply on the employer's site

About this role

<div class="content-intro"><div> <div> <div class="gmail_quote"> <div> <div><span id="m_1770241969069985273m_-2746164444908759431gmail-docs-internal-guid-131e4fb0-7fff-b4e9-ff50-e8cf32449b1b">CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at <a href="http://www.coreweave.com/" target="_blank" data-saferedirecturl="https://www.google.com/url?q=http://www.coreweave.com&source=gmail&ust=1762613132717000&usg=AOvVaw3D-UOhNaqEvF5BEWxjYyAU">www.coreweave.com</a>.</span></div> </div> </div> </div> </div></div><h3 class="font-claude-response-body break-words whitespace-normal"><strong>What You'll Do:</strong></h3> <p class="font-claude-response-body break-words whitespace-normal">The <strong>AI/ML TPM</strong> team owns delivery and execution across CoreWeave's AI/ML Platform Services organization. The team partners closely with Product, Engineering, Research, Infrastructure, and Go-to-Market teams to deliver scalable, reliable, and high-performance platforms that support the full AI lifecycle. AI/ML TPMs drive alignment and execution across highly technical, cross-functional teams to ensure the successful delivery of customer-facing infrastructure and platform capabilities used by researchers, engineers, and enterprise customers.</p> <p class="font-claude-response-body break-words whitespace-normal">As a Technical Program Manager focused on inference, you will lead complex, cross-functional programs spanning inference platform delivery, customer onboarding, launch readiness, and runtime optimization. The Inference team is responsible for building and operating highly scalable, reliable production inference services for both serverless and dedicated inference use cases. The work spans platform reliability, operational excellence, customer onboarding, release validation, runtime performance, and the launch of new inference capabilities that help customers run models at scale with strong price-performance and low operational friction.</p> <p class="font-claude-response-body break-words whitespace-normal">In this role, you will partner with engineering, product, infrastructure, and go-to-market teams to drive programs that improve how inference services are launched, onboarded, operated, and optimized. This includes managing the execution of customer-facing onboarding programs, launch readiness for dedicated inference capabilities, and the platform improvements needed to support scale, reliability, and predictable delivery.</p> <h3 class="font-claude-response-body break-words whitespace-normal"><strong>In this role, you will:</strong></h3> <ul class="[li_&]:mb-0 [li_&]:mt-1 [li_&]:gap-1 [&:not(:last-child)_ul]:pb-1 [&:not(:last-child)_ol]:pb-1 list-disc flex flex-col gap-1 pl-8 mb-3"> <li class="font-claude-response-body whitespace-normal break-words pl-2">Drive end-to-end program management for inference platform initiatives spanning reliability, customer onboarding, launch readiness, and runtime optimization</li> <li class="font-claude-response-body whitespace-normal break-words pl-2">Lead cross-functional programs for customer onboarding across dedicated and serverless inference offerings, ensuring clear ownership, launch criteria, and readiness for strategic customer use cases</li> <li class="font-claude-response-body whitespace-normal break-words pl-2">Drive launch readiness for new inference capabilities by aligning teams around real customer outcomes, supportability, and end-to-end validation</li> <li class="font-claude-response-body whitespace-normal break-words pl-2">Partner with engineering and product to define and deliver roadmap outcomes for latency, throughput, uptime, operational quality, and price-performance</li> <li class="font-claude-response-body whitespace-normal break-words pl-2">Coordinate multi-team execution across platform, infrastructure, and customer-facing teams to deliver reliable and scalable inference services</li> <li class="font-claude-response-body whitespace-normal break-words pl-2">Build and operationalize success metrics, dashboards, launch gates, and review cadences to measure service reliability, onboarding readiness, efficiency, and quality across the inference stack</li> <li class="font-claude-response-body whitespace-normal break-words pl-2">Establish repeatable processes for release validation, performance regression tracking, launch management, and postmortem follow-through</li> <li class="font-claude-response-body whitespace-normal break-words pl-2">Help unify operational processes, support mechanisms, and execution visibility across inference deployment models and customer onboarding paths</li> <li class="font-claude-response-body whitespace-normal break-words pl-2">Create strong communication channels between Engineering, Product, Infrastructure, and Go-to-Market teams to align priorities and deliver predictable, high-impact outcomes</li> </ul> <h3 class="font-claude-response-body break-words…

Skills asked for

Similar jobs

Apply on the employer's site

Your next role is already in here.

Search live openings from thousands of employers, save the ones worth a second look, and let JobBob keep watch for the rest.