仕事内容
<div class="content-intro"><div>
<div>
<div class="gmail_quote">
<div>
<div><span id="m_1770241969069985273m_-2746164444908759431gmail-docs-internal-guid-131e4fb0-7fff-b4e9-ff50-e8cf32449b1b">CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at <a href="http://www.coreweave.com/" target="_blank" data-saferedirecturl="https://www.google.com/url?q=http://www.coreweave.com&source=gmail&ust=1762613132717000&usg=AOvVaw3D-UOhNaqEvF5BEWxjYyAU">www.coreweave.com</a>.</span></div>
</div>
</div>
</div>
</div></div><h3><strong>What You'll Do:</strong></h3>
<p>CoreWeave’s Inference team builds and operates the core cloud platform powering massive-scale GPU workloads for AI/ML, VFX, rendering, and real-time inference. Our stack is engineered for speed, scale, and cost efficiency—providing a powerful alternative to traditional hyperscalers. Infrastructure is our product, and we operate it at massive scale to power the most demanding workloads.</p>
<p><strong>About the role:</strong> </p>
<p>As a Staff Software Engineer (IC5) on the Inference team, you will operate as a technical leader across multiple teams and services, driving architecture, performance, and reliability for CoreWeave’s Kubernetes-native inference platform. You will define and lead complex, cross-cutting design initiatives spanning request routing, adaptive scheduling, GPU resource management, and cost-per-token optimization under strict P99 SLAs. This high-impact role requires you to implement advanced inference optimizations—such as specul
求めるスキル
Python
CUDA
LLM
Kubernetes
AWS
Rust
C++