Top SRE Engineer Jobs in Seattle, WA

3 Days AgoSaved
Hybrid
Seattle, WA
140K-215K Annually
Senior level
140K-215K Annually
Senior level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Lead architecture and implementation of reliability improvements across CrowdStrike's cloud-native platform. Build shared libraries and services, drive observability and SLO practices, perform performance and cost optimization, run resilience engineering and chaos experiments, automate infrastructure-as-code, mentor engineers, and embed with product teams to deliver scalable, highly reliable distributed systems at organizational scale.
Top Skills: AIAlertingAWSCassandraElasticsearchGCPGoInfrastructure-As-CodeJavaKafkaKotlinKubernetesNode.jsObservability (TracingOciOpensearchProfilingProtobufPythonScalaSlos)
Reposted 11 Days AgoSaved
In-Office
Seattle, WA
167K-204K Annually
Senior level
167K-204K Annually
Senior level
Fintech • Financial Services
As a Site Reliability Engineer focusing on frontend performance, you build infrastructure for self-service performance monitoring, optimize AI operations, and architect resilient systems for high-traffic applications.
Top Skills: Apm InstrumentationAWSCli ToolsDatadogJavaKubernetesRumTypescript
Reposted 4 Days AgoSaved
Remote or Hybrid
Seattle, WA
200K-250K Annually
Senior level
200K-250K Annually
Senior level
Digital Media • Gaming • Information Technology • Software • Sports • Esports • Big Data Analytics
Lead long-term strategy and architecture for cloud and on‑prem platform infrastructure, driving Kubernetes and multi‑cloud reliability, IaC/GitOps automation, observability, SLO/SLI/error‑budget practices, incident leadership, AI‑augmented tooling adoption, and mentorship of senior engineers to improve platform resilience and developer experience.
Top Skills: Amazon Elastic Kubernetes Service (Eks)AutoscalingAWSCapacity PlanningCi/CdGitopsGoGoogle Cloud PlatformGoogle Kubernetes Engine (Gke)Identity And Access ManagementInfrastructure As CodeKubernetesLinuxNetworkingObservabilityOperatorsPulumiPythonRke2StorageTerraform
Reposted 15 Days AgoSaved
Easy Apply
Remote or Hybrid
Seattle, WA
Easy Apply
119K-170K Annually
Senior level
119K-170K Annually
Senior level
Cloud • Information Technology • Security • Software • Cybersecurity
As a Staff Site Reliability Engineer, you'll oversee Zscaler production data center services, optimize code, and ensure cloud service availability and performance. Collaborate with cross-functional teams to improve processes and resolve escalated issues.
Top Skills: BashDnsFirewallsGrafanaHTTPIcmpLoad BalancingNagiosOsi ModelPrometheusPythonTcp/Ip
Reposted 6 Days AgoSaved
Remote
Seattle, WA
150K-200K Annually
Senior level
150K-200K Annually
Senior level
Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
Ensure stability and resilience of Runpod's distributed AI platform by defining SLIs/SLOs, leading incident response, building observability and reliability tooling, automating operational workflows, and partnering with engineering teams to reduce toil and improve production readiness.
Top Skills: BashCi/CdContainerized Production SystemsGoGpu Observability ToolingGrafanaInfrastructure As CodeLinuxPrometheusPython
Reposted 7 Days AgoSaved
Easy Apply
Remote or Hybrid
Seattle, WA
Easy Apply
Internship
Internship
Cloud • Information Technology • Security • Software • Cybersecurity
This internship role focuses on SRE skills, requiring collaboration and problem-solving in dynamic environments for Zscaler's Zero Trust Exchange team.
Top Skills: AnsibleAws EcsKubernetesLinuxPythonTerraform
Reposted 7 Days AgoSaved
Easy Apply
Remote
Seattle, WA
Easy Apply
Mid level
Mid level
Cloud • Security • Software • Cybersecurity • Automation
As a Cloud Cost Utilization SRE at GitLab, you'll manage cloud spending, improve tracking and optimization of cloud usage, and collaborate with finance and engineering teams to enhance cost efficiency across AWS and GCP.
Top Skills: AnsibleAWSElkGCPGrafanaLokiMimirPrometheusTempoTerraform
Reposted 17 Days AgoSaved
Easy Apply
Remote or Hybrid
Seattle, WA
Easy Apply
127K-249K Annually
Senior level
127K-249K Annually
Senior level
Big Data • Cloud • Software • Database
Develop and maintain Kubernetes runtime environments, support developers, resolve critical issues, and participate in on-call rotations for production systems.
Top Skills: AWSAzureCert-ManagerCorednsCrdsCriCsiGatekeeperGCPGoHelmKubernetesKustomizeOperatorsPythonTerraform
Reposted 10 Days AgoSaved
Easy Apply
Remote
Seattle, WA
Easy Apply
100K-110K Annually
Mid level
100K-110K Annually
Mid level
Healthtech • Software
Operate and maintain AWS-hosted MERN applications and large-scale data workflows. Manage serverless and Spark-based pipelines, perform incident response and on-call duties, engineer automation to eliminate operational toil, ensure HIPAA/SOC2/HITRUST compliance, build observability and lead blameless post-mortems.
Top Skills: Amazon EcsAmazon EksAmazon EmrAthenaAws GlueAws LambdaAws SnsAws SqsCloudwatchEc2IamJavaScriptMernMySQLNode.jsOpentofuPysparkPythonRabbitMQTerraformTypescriptVpc
20 Days AgoSaved
Hybrid
Seattle, WA
Mid level
Mid level
Financial Services
Design, deploy, and operate secure, highly available cloud infrastructure and container platforms. Build and maintain IaC (Terraform), CI/CD pipelines, monitoring and logging, security integrations, disaster recovery, and on-call support. Collaborate with developers to automate deployments and improve reliability, scalability, and operational efficiency.
Top Skills: Aqua SecurityAWSAws CloudwatchAws CodepipelineAzureBashCircleCICloud FoundryDynatraceEc2EcsEksElastic StackElasticsearchElk StackGitlab CiGrafanaIamJavaJenkinsKibanaKubernetesLogstashNode.jsPrometheusPythonRdsS3ShellSnykSonarqubeSpinnakerSplunkSpring BootTerraformTrivyVpc
Reposted 13 Days AgoSaved
Easy Apply
Remote or Hybrid
Seattle, WA
Easy Apply
200K-230K Annually
Senior level
200K-230K Annually
Senior level
Artificial Intelligence • Machine Learning
Lead development of AI-assisted reliability tooling, own incident response end-to-end, improve observability and SLO/SLI frameworks, scale single-tenant SaaS operations, mentor engineers, and reduce recurring operational toil through engineering and automation.
Top Skills: Cloud PlatformsGoKubernetesLinuxLlm/Ai ToolingLogs And TracingObservability ToolingPythonSlo/Sli Frameworks
14 Days AgoSaved
Remote
Seattle, WA
180K-220K Annually
Senior level
180K-220K Annually
Senior level
Software • Defense
Work as an SRE embedded with product teams to improve reliability by fixing application code (primarily TypeScript), building observability (Prometheus, Loki, Grafana, Alloy), defining SLIs/SLOs, leading incident response and postmortems, automating toil, and supporting deployments across on‑prem DoD and AWS environments.
Top Skills: AlloyAWSBashContainersDockerGithub ActionsGitlab Ci/CdGoGrafanaJenkinsKubectlKubernetesLokiNode.jsPrometheusPythonTypescript
New

Track Smarter, Apply Better.

Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.

Use For Free
Application Tracker Preview
Reposted 14 Days AgoSaved
Easy Apply
Remote or Hybrid
Seattle, WA
Easy Apply
127K-249K Annually
Senior level
127K-249K Annually
Senior level
Big Data • Cloud • Software • Database
As a Senior Site Reliability Engineer, you'll design and build complex systems, support Atlas platform operations, automate processes, and ensure high availability of services.
Top Skills: AWSAzureDnsGCPGoHTTPLinuxPythonRubyTls
Reposted 8 Hours AgoSaved
In-Office
Seattle, WA
102K-219K Annually
Mid level
102K-219K Annually
Mid level
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Own reliability and operational health for Substrate services in regulated environments. Respond to incidents as on-call engineer, diagnose and fix production issues, implement automation, develop monitoring and telemetry for SLOs, lead post-incident reviews, and collaborate with engineering teams to embed reliability and security.
Top Skills: Department Of Defense EnvironmentsExchange OnlineGcc High (Gcch)Gcc Moderate (Gccm)M365 CopilotMicrosoft CloudMicrosoft SubstrateMonitoringOn-Call AutomationSlosTelemetry
19 Days AgoSaved
Remote or Hybrid
Seattle, WA
Senior level
Senior level
Fintech • Software
Lead SRE efforts for DFIN SaaS: ensure availability, performance, scalability, and automation. Implement monitoring, CI/CD, IaC, container orchestration, AI-enhanced observability, incident response, RCA, and runbook automation while collaborating across engineering teams.
Top Skills: .NetAiopsAksAnsibleAppdynamicsAWSAzureAzure DevopsBashC#Ci/CdCloud Ai ServicesContainersCosmosDatadogDynatraceEksFirewallHarnessIdera Sql Diagnostic ManagerInfrastructure As Code (Iac)JavaJenkinsKubernetesLinuxLoad BalancingNew RelicPowershellPythonRedgate Sql MonitorSolarwinds Database Performance AnalyzerSQLTerraformWindows
Reposted YesterdaySaved
Hybrid
Seattle, WA
170K-220K Annually
Senior level
170K-220K Annually
Senior level
Artificial Intelligence • Legal Tech • Software • Generative AI
Lead and own the release and deployment process, manage GitHub workflows and Actions, build and maintain AWS infrastructure and observability, automate deployments and internal tooling, respond to incidents and be on-call, contribute code to reliability tooling, and support global/offshore teams across time zones.
Top Skills: AWSBashChatgptCi/CdClaudeEc2GitGithub ActionsIamLambdaMetricsObservability (LogsPostgres SqlPythonRdsTraces)TypescriptVpc
Reposted 3 Days AgoSaved
In-Office or Remote
Seattle, WA
120K-261K Annually
Senior level
120K-261K Annually
Senior level
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Lead qualification, performance validation, and production readiness for new Azure Storage hardware and firmware. Drive test planning, automation frameworks, large-scale telemetry and benchmark analysis, root-cause investigations across software/hardware/firmware, and partner with engineering and vendors to resolve reliability and performance issues.
Top Skills: Automation FrameworksAzureAzure StorageBenchmarkingFirmwareNetworkingSsdsTelemetry
Reposted 3 Days AgoSaved
In-Office
Seattle, WA
165K-230K Annually
Senior level
165K-230K Annually
Senior level
Aerospace • Other
Design, deploy, and automate infrastructure for on‑prem and cloud compute. Manage core services (databases, monitoring, storage), collaborate with software teams to build scalable, operable systems, and own the service lifecycle from design through deployment, operation, and refinement to ensure secure, reliable, and autonomous satellite software services.
Top Skills: AnsibleBashBazelCloudDatabasesKubernetesLinuxMakefilesMonitoringPythonTcp/IpTerraform
Reposted 3 Days AgoSaved
In-Office
Seattle, WA
125K-175K Annually
Junior
125K-175K Annually
Junior
Aerospace • Other
Design, deploy, and operate on-premises Kubernetes clusters and core infrastructure (databases, monitoring, distributed storage). Build automation, troubleshoot across the Starshield stack, collaborate with software teams to ensure scalable, highly available services, and improve lifecycle processes.
Top Skills: AnsibleBashBazelC++DatabasesDistributed StorageGoKubernetesLinuxMakefilesMonitoringOci ContainersPythonTcp/IpTerraform
4 Days AgoSaved
In-Office or Remote
Seattle, WA
120K-261K Annually
Senior level
120K-261K Annually
Senior level
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Lead reliability and availability for large-scale Azure services: act as DRI, run on-call, design and implement improvements, automate operations, develop design docs and code, collaborate across teams, and drive service lifecycle and incident restoration for cloud-hosted services.
Top Skills: AzureJavaScriptPythonService FabricShell Scripting
Reposted 4 Days AgoSaved
Hybrid
Seattle, WA
145K-200K Annually
Senior level
145K-200K Annually
Senior level
Blockchain • Energy • Cryptocurrency
Hands-on role to assess, implement, test, and document backup, restore, failover, and recovery capabilities. Inventory critical systems, design and automate backup and restoration, run recovery exercises, produce runbooks, validate recoverability, measure RTO/RPO, and train system owners. Collaborate with Security, SRE, DevOps, QA, and application teams to harden shared recovery capabilities and transfer operational ownership.
Reposted 5 Days AgoSaved
In-Office
Seattle, WA
120K-261K Annually
Senior level
120K-261K Annually
Senior level
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Lead and develop SRE practices for secure, highly regulated Microsoft CISO services. Architect and automate hybrid/cloud infrastructure, manage petabyte-scale data platforms and pipelines, implement IaC and disaster recovery, deliver telemetry and automation, participate in on-call rotations, and mentor engineers to improve reliability, diagnosability, security, and compliance.
Top Skills: AciAksArm TemplatesAWSAzureAzure BicepAzure Container AppsAzure Event HubsAzure Key VaultAzure SynapseAzure VmsBashBicepC#DockerGCPHadoopIaasJavaKubernetesMicrosoft 365 (ExchangePowershellPythonSharepointSkypeSparkTeams)Terraform
5 Days AgoSaved
In-Office or Remote
Seattle, WA
102K-219K Annually
Junior
102K-219K Annually
Junior
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Design, operate, and improve large-scale Microsoft 365 and Purview services. Automate operational processes, build telemetry and monitoring pipelines, develop scripts/code, troubleshoot and optimize systems, participate in on-call incident response, and collaborate with engineering teams to improve availability, performance, security, and customer experience.
Top Skills: CC#C++JavaJavaScriptMicrosoft 365Microsoft CloudPurviewPython
Reposted 23 Days AgoSaved
Easy Apply
Remote or Hybrid
Seattle, WA
Easy Apply
126K-248K Annually
Senior level
126K-248K Annually
Senior level
Big Data • Cloud • Software • Database
The Senior Site Reliability Engineer will develop and support distributed storage services, ensuring reliability and operational safety, with a focus on automation and efficiency.
Top Skills: AWSAzureDnsGoGoogle Cloud PlatformKubernetesLinuxPythonTcp/IpTls
Reposted 24 Days AgoSaved
Remote or Hybrid
Seattle, WA
175K-200K Annually
Senior level
175K-200K Annually
Senior level
eCommerce • Fintech • Payments • Software
The role involves ensuring software reliability and performance, managing incidents, developing infrastructure automation, and mentoring junior engineers within a platform team.
Top Skills: AWSCloudFormationDatadogKubernetesOpentelemetryRubyRuby On RailsTerraform
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account