Snowflake Logo

Snowflake

Senior Software Engineer - Capacity

Posted 23 Days Ago
Be an Early Applicant
Hybrid
Bellevue, WA, USA
200K-288K Annually
Senior level
Hybrid
Bellevue, WA, USA
200K-288K Annually
Senior level
Design and build Snowflake’s multi-cloud capacity platform for CPU and GPU resources across AWS, Azure, and GCP. Own canonical capacity data, demand forecasting, procurement, reservations, allocation, utilization, supply-risk visibility, and cost optimization systems. Collaborate with cloud providers and core services, AI/ML, warehouse, and finance teams. Ensure high availability, reliability, and performance through production support, troubleshooting, on-call participation, and incident management.
The summary above was generated by AI

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done.

Senior Software Engineer, Capacity Engineering

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic, fast-moving environments and approach challenges with an experimental mindset, rapidly testing emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done.

Snowflake’s infrastructure is expanding rapidly across AWS, Azure, and GCP. The Capacity team plays a pivotal role in provisioning the cloud resources essential for Snowflake's operations and ongoing growth. Capacity Engineering accurately models demand, forecasts requirements, and delivers optimal CPU and GPU capacity on schedule. We drive hardware cost-efficiency and price/performance while continually maximizing fleet utilization. To achieve this across all major cloud providers, the team is building a centralized, self-serve internal capacity platform. This software-driven system provides early visibility into supply risks, ensures sufficient lead time for capacity deployment, and maintains high utilization across committed cloud resources.

The technical problem spans the full lifecycle. We model demand and supply as first-class data, reconcile heterogeneous provider telemetry and commitments into a single canonical capacity layer, and make the live state of the fleet legible and actionable in real time. That includes CPU and GPU procurement and reservation lifecycle for AI/ML workloads (training, fine-tuning, and model serving), demand forecasting, cloud resource and cost optimization, hardware evolution analysis as new generations become available (price/performance, cross-family flexibility, migration paths), and the allocation and efficiency systems that close utilization gaps with the teams that own those workloads.

We are actively looking for a senior software engineer. If you love solving problems at scale, prefer to write scalable, reliable, and testable software, are an ace troubleshooter, and are deeply technical, then this is the role for you! Snowflake’s growth and multi-cloud footprint in a constrained capacity environment demand real engineering maturity in the systems that plan and land compute. While the domains below describe the shape of our current goals, the engineer will drive the strategy and deliverables for clear company impact.

AS A SENIOR SOFTWARE ENGINEER AT SNOWFLAKE YOU WILL:
  • Design and build the capacity platform that unifies CPU and GPU allocation, procurement, reservation lifecycle, and utilization across all three clouds.

  • Own the canonical capacity data layer: ingest and reconcile demand forecasts, provider supply signals, commitments, and fleet utilization into a single, trustworthy model consumed across the company.

  • Serve as the liaison to Cloud Service Providers managing and integrating vendor relationships into the capacity planning and procurement workflow.

  • Build planning and allocation systems that translate demand into hardware requirements (shape, quantity, region, timing) and surface supply risk early, with real-time visibility into fleet and reservation health.

  • Drive efficiency: instrument utilization across CPU and GPU accelerator workloads, establish price/performance baselines, and build the tooling that recovers stranded capacity and right-sizes commitments.

  • Integrate hardware evolution into the platform: evaluate new CPU and GPU generations and their price/performance, and build the flexibility (backup and cross-family fallbacks) that keeps plans aligned to the hardware roadmap.

  • Partner with core services, warehouse, AI/ML, and finance teams to forecast and procure capacity ahead of launches, support AI/ML workloads reliably, and turn insights into procurement and allocation decisions.

  • Ensure high availability, reliability, and performance of capacity systems by participating in on-call rotations and incident management.

WHAT WE LOOK FOR:
  • 7+ years of industry experience designing, building, and supporting large-scale systems in production.

  • Hands-on experience working with cloud providers on compute cluster and cloud services provisioning (CPU and/or GPU fleets).

  • Experience with capacity planning, procurement, resource management, or efficiency work on systems built on large private clouds or public cloud providers.

  • Deep system and architectural analysis experience to identify actionable performance, availability, and efficiency insights across CPU and GPU accelerator fleets.

  • Proficiency in programming languages such as Go, Python, or Java.

  • Excellent problem-solving skills and ability to troubleshoot complex issues in a production environment.

  • Strong communication skills and the ability to collaborate effectively in a team environment.

  • BS / MS in Computer Science, Engineering or related fields.

  • Experience developing or using observability infrastructure such as OpenTelemetry or Prometheus is a plus.

  • Familiarity with accelerator/GPU fleets, hardware price/performance analysis, or Kubernetes-based compute at scale is a plus.

  • Prior background working with Modeling, Forecasting, Cloud Spend Optimization, and LLMs is a plus.

Snowflake is growing fast, and we’re scaling our team to help enable and accelerate our growth. We are looking for people who share our values, challenge ordinary thinking, and push the pace of innovation while building a future for themselves and Snowflake.

Snowflake is growing fast, and we’re scaling our team to help enable and accelerate our growth. We are looking for people who share our values, challenge ordinary thinking, and push the pace of innovation while building a future for themselves and Snowflake.

How do you want to make your impact?

For jobs located in the United States, please visit the job posting on the Snowflake Careers Site for salary and benefits information: careers.snowflake.com

Snowflake Bellevue, Washington, USA Office

In the heart of Silicon Valley, you'll find our 4-story, 2-tower San Mateo hub, which actually emerged from the very spot Snowflake started in 2012 – it all began in one of our founder's humble San Mateo apartments.

Similar Jobs

One Month Ago
In-Office
Seattle, WA, USA
75K-215K Annually
Senior level
75K-215K Annually
Senior level
Insurance
Design, build, and maintain full-stack applications and backend services to support capacity management. Develop React front-ends, RESTful APIs, and cloud-based scalable systems. Collaborate with stakeholders, improve performance and observability, participate in code reviews, troubleshoot production issues, and promote engineering best practices.
Top Skills: Ci/CdCloud PlatformsGoIaasJavaScriptMicroservicesObservabilityReactRestful ApisVersion Control
5 Hours Ago
In-Office or Remote
Seattle, WA, USA
200K-260K Annually
Expert/Leader
200K-260K Annually
Expert/Leader
Blockchain • Fintech • Payments • Financial Services • Cryptocurrency • Web3
Lead regional growth strategy for USDC by building partnerships with exchanges and OTC desks, delivering regionally compliant product initiatives, and creating AI-driven, automated growth platforms. Own experimentation frameworks, dashboards, and metrics to scale liquidity, adoption, and measurable share shift versus competing stablecoins through cross-functional execution and data-driven prioritization.
Top Skills: Agent-Driven SystemsAIAnalyticsAutomation PlatformsBlockchainDashboardsExperimentation FrameworksStablecoins
5 Hours Ago
In-Office or Remote
Seattle, WA, USA
200K-258K Annually
Expert/Leader
200K-258K Annually
Expert/Leader
Blockchain • Fintech • Payments • Financial Services • Cryptocurrency • Web3
Leads Circle’s internal communications and employer brand strategy across a global, distributed workforce. Partners with executives, HR, Talent Acquisition, Marketing, and Corporate Communications to shape company messaging, employee engagement, EVP, talent storytelling, candidate experience, alumni programs, and employer reputation. Oversees content, communication infrastructure, change communications, measurement, and AI-enabled workflows while managing and developing a communications team.
Top Skills: AILlm

What you need to know about the Seattle Tech Scene

Home to tech titans like Microsoft and Amazon, Seattle punches far above its weight in innovation. But its surrounding mountains, sprinkled with world-famous hiking trails and climbing routes, make the city a destination for outdoorsy types as well. Established as a logging town before shifting to shipbuilding and logistics, the Emerald City is now known for its contributions to aerospace, software, biotech and cloud computing. And its status as a thriving tech ecosystem is attracting out-of-town companies looking to establish new tech and engineering hubs.

Key Facts About Seattle Tech

  • Number of Tech Workers: 287,000; 13% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Amazon, Microsoft, Meta, Google
  • Key Industries: Artificial intelligence, cloud computing, software, biotechnology, game development
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Madrona, Fuse, Tola, Maveron
  • Research Centers and Universities: University of Washington, Seattle University, Seattle Pacific University, Allen Institute for Brain Science, Bill & Melinda Gates Foundation, Seattle Children’s Research Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account