Designworks Talent LLC Logo

Designworks Talent LLC

Virtualization & Orchestration Engineer

Posted 20 Days Ago
Be an Early Applicant
Hybrid
Bellevue, WA, USA
Senior level
Hybrid
Bellevue, WA, USA
Senior level
Design, build, and operate GPU-focused virtualization and Kubernetes orchestration systems for multi-tenant AI/HPC clusters. Implement automated provisioning, workload scheduling, resource management, and tooling to scale, secure, and run GPU compute reliably across production environments while partnering with hardware, networking, and platform teams.
The summary above was generated by AI
Virtualization & Orchestration Engineer

Location: Hybrid | Bellevue, WA Area
Titles: Intermediate, Senior and Staff (multiple roles available)

Build the Platform Layer Powering Next-Generation AI Infrastructure
About the Opportunity

A well-funded, rapidly growing AI infrastructure company is building a next-generation cloud platform designed to power the full lifecycle of artificial intelligence. The organization is developing a comprehensive AI infrastructure, platform, and services portfolio that supports the full spectrum of AI workloads—including large-scale compute, model training, fine-tuning, inference, and emerging agentic AI applications.

Backed by significant long-term investment, the company combines the speed, ownership, and innovation of a startup with the stability and resources of an established parent organization. Engineering teams are intentionally lean, highly collaborative, and AI-native, leveraging modern automation and tooling to build infrastructure capable of supporting the industry's most demanding AI workloads.

We're seeking Virtualization & Orchestration Engineers to build the platform layer that enables customers to reliably consume GPU compute at scale. This team is responsible for designing and operating the virtualization, Kubernetes, provisioning, and orchestration systems that power large-scale AI workloads across next-generation data center infrastructure.

 
 
The Opportunity

This is a foundational engineering role within the company's largest infrastructure engineering organization. You'll help design and build the systems that make GPU capacity available, scalable, secure, and reliable across a multi-tenant AI cloud platform.

You'll work at the intersection of virtualization, Kubernetes, distributed systems, GPU infrastructure, and high-performance computing—solving complex challenges around workload scheduling, resource allocation, cluster management, and infrastructure automation.

This opportunity is ideal for engineers who enjoy building large-scale platforms from the ground up and owning critical infrastructure systems end-to-end.

 

What You'll Do
  • Design and build virtualization infrastructure supporting GPU-intensive AI and HPC workloads.

  • Develop and operate Kubernetes-based orchestration systems for GPU cluster provisioning and workload scheduling.

  • Build automated provisioning systems that enable GPU capacity to be allocated, scaled, and reclaimed efficiently across multiple tenants.

  • Design solutions for workload placement, resource management, and cluster lifecycle operations.

  • Partner closely with hardware, networking, infrastructure, and AI platform teams to ensure orchestration systems align with real-world cluster architectures and constraints.

  • Improve the reliability, security, scalability, and operational maturity of the orchestration platform.

  • Build tooling and automation that simplifies infrastructure management and improves developer and customer experiences.

  • Contribute to architectural decisions, engineering standards, and best practices as the platform evolves.

 
What We're Looking For
  • Strong hands-on experience with Kubernetes and container orchestration in production environments.

  • Experience designing, building, and operating large-scale infrastructure platforms.

  • Background with virtualization technologies supporting cloud, HPC, GPU, or distributed computing environments.

  • Understanding of GPU cluster provisioning, workload scheduling, and resource management.

  • Experience with Linux-based infrastructure and distributed systems concepts.

  • Ability to independently own complex systems from design through production operation.

  • Comfortable working in a fast-moving environment where architecture and processes are being established.

 
Preferred Qualifications
  • Experience with GPU scheduling technologies such as Slurm, Kubernetes device plugins, NVIDIA GPU Operator, or similar frameworks.

  • Experience supporting AI infrastructure, machine learning platforms, HPC environments, or GPU cloud providers.

  • Background building multi-tenant infrastructure platforms for cloud providers or large-scale compute environments.

  • Experience with infrastructure automation, Infrastructure as Code, and platform engineering practices.

  • Familiarity with high-performance networking and GPU cluster architectures.

 
Compensation
  • Competitive base pay for Bellevue market

  • Certain roles are eligible for additional rewards, including merit increases, annual bonus, and long term incentives. These awards are allocated based on individual performance

  • U.S. based employees have access to medical, dental, and vision insurance, a 401(k) plan and company match, employees also receive per calendar year, paid holidays

 
Location
  • Hybrid role based in the Bellevue, WA area.

  • Approximately three days per week in the office.

  • Candidates elsewhere in the U.S. who are open to relocation are encouraged to apply.

  • U.S. work authorization is required. Visa sponsorship is not currently available.

 
Why Join?
  • Build the orchestration platform powering one of the industry's most advanced AI infrastructure environments.

  • Work directly on GPU clusters, Kubernetes platforms, and large-scale distributed systems.

  • Solve complex infrastructure challenges across virtualization, scheduling, resource management, and automation.

  • Join early enough to influence architecture, engineering practices, and platform strategy.

  • Collaborate with a highly experienced team building the foundation for next-generation AI applications.

  • Enjoy the ownership and technical impact of a startup environment backed by significant long-term investment.

Similar Jobs

27 Minutes Ago
In-Office
Bellevue, WA, USA
182K-242K Annually
Senior level
182K-242K Annually
Senior level
Cloud • Information Technology • Machine Learning
Lead product strategy and roadmap for CoreWeave's cloud storage offerings, focusing on high-performance file and dedicated storage for AI workloads. Conduct market research, collaborate with engineering to design and launch scalable storage features, advocate for customers, coordinate cross-functional launches, and monitor product performance using data-driven insights to guide enhancements.
Top Skills: Cloud StorageFile System StorageHigh-Performance StorageObject Storage
27 Minutes Ago
In-Office
Seattle, WA, USA
157K-210K Annually
Senior level
157K-210K Annually
Senior level
Cloud • Information Technology • Machine Learning
Lead and grow a Bare Metal Infrastructure Support team to maintain data center hardware, triage incidents, manage escalations, and improve processes. Collaborate with product and engineering, manage ticket workflows and shift coverage, mentor engineers, and ensure high availability and performance for GPU-heavy client workloads.
Top Skills: Bare MetalCliJIRALinuxLiquid CoolingNvidia A100Nvidia H100NvlinkPcieZendesk
47 Minutes Ago
Easy Apply
Remote or Hybrid
United States
Easy Apply
79K-120K Annually
Senior level
79K-120K Annually
Senior level
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Support partner sales by implementing and managing scalable partner processes, driving program design, owning projects end-to-end, improving systems (including Salesforce), performing quantitative analysis to track KPIs, and coordinating cross-functional stakeholders to enable partner-facing sales teams.
Top Skills: CRMPartner Relationship Management (Prm) ToolsSalesforce

What you need to know about the Seattle Tech Scene

Home to tech titans like Microsoft and Amazon, Seattle punches far above its weight in innovation. But its surrounding mountains, sprinkled with world-famous hiking trails and climbing routes, make the city a destination for outdoorsy types as well. Established as a logging town before shifting to shipbuilding and logistics, the Emerald City is now known for its contributions to aerospace, software, biotech and cloud computing. And its status as a thriving tech ecosystem is attracting out-of-town companies looking to establish new tech and engineering hubs.

Key Facts About Seattle Tech

  • Number of Tech Workers: 287,000; 13% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Amazon, Microsoft, Meta, Google
  • Key Industries: Artificial intelligence, cloud computing, software, biotechnology, game development
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Madrona, Fuse, Tola, Maveron
  • Research Centers and Universities: University of Washington, Seattle University, Seattle Pacific University, Allen Institute for Brain Science, Bill & Melinda Gates Foundation, Seattle Children’s Research Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account