Peraton Logo

Peraton

Senior Monitoring and Observability Engineer

Posted Yesterday
Be an Early Applicant
Remote
Hiring Remotely in United States
104K-166K Annually
Senior level
Remote
Hiring Remotely in United States
104K-166K Annually
Senior level
Build and operate enterprise observability capabilities across AWS and GovCloud, including telemetry pipelines, instrumentation, dashboards, alerting, incident support, and platform integrations. Automate agent and configuration deployment using Ansible, GitLab CI/CD, and infrastructure-as-code. Manage telemetry reliability, cost, cardinality, governance, compliance, and documentation while supporting on-call operations and continuous improvement.
The summary above was generated by AI
Responsibilities

Peraton is seeking a Monitoring and Observability Engineer with a strong observability background to build and operate the telemetry, monitoring, and alerting capability supporting an enterprise platform running in AWS and AWS GovCloud within a FedRAMP-authorized boundary. This is a hands-on engineering role, reporting to the Senior SRE / Observability Lead, focused on instrumentation, dashboarding, alert quality, and telemetry pipeline reliability across metrics, logs, and traces. 

This person will work closely with platform engineering and the broader SRE team to ensure every service is properly instrumented, every alert is actionable, and incident responders have the data they need to diagnose issues quickly — while keeping telemetry cost and cardinality under control. 


Location: Remote


Shift Schedule: 8am – 5pm Eastern Standard Time (EST)


What you will do:

Observability and Monitoring Engineering

  • Build and maintain telemetry pipelines end to end — collection, enrichment, routing, sampling, storage, and retention — across Dynatrace, Datadog, and Splunk.
  • Instrument applications and infrastructure for metrics, logs, and traces; support OpenTelemetry-based, vendor-neutral instrumentation standards.
  • Design observability for distributed systems: microservices, containers, and cloud-native workloads in AWS/GovCloud.
  • Build dashboards and golden-signal monitoring aligned to service criticality; enable correlation across metrics, events, logs, and traces so anomalies can be diagnosed without manual pivoting.
  • Monitor the health of the observability platforms themselves and remediate collector, agent, and pipeline issues.

Alerting and Incident Support

  • Build and tune alert definitions with clear ownership, severity, and runbook links for every production alert; reduce false positives and alert flapping.
  • Support on-call rotations and participate in incident bridges, providing telemetry-driven triage and root-cause data.
  • Integrate alert routing, deduplication, and maintenance-window handling with ServiceNow and paging tools.
  • Contribute to post-incident reviews with detection-source analysis and MTTD/MTTA trend reporting.

Automation and Platform Integration

  • Deploy and configure observability agents/collectors using Ansible, Ansible Tower, and AAP for consistent, automated rollout across environments.
  • Build observability-as-code pipelines in GitLab (GitLab CI/CD), including automated dashboard, alert, and collector-configuration deployment; support migration of remaining Jenkins jobs to GitLab.
  • Partner with platform engineering to instrument new Terraform-provisioned infrastructure as part of standard build patterns.
  • Advance configuration-as-code for pipelines, dashboards, and alert definitions so changes are versioned and peer-reviewed.

Governance, Capacity, and Compliance

  • Support tagging conventions, dashboard/alert lifecycle management, and retention-tier governance.
  • Help manage telemetry cost, ingest volume, and cardinality growth across the observability platforms.
  • Support sensitive-data masking, audit, and continuous-monitoring requirements for observability data.
  • Maintain observability architecture documentation and runbooks; operate within SAFe using ServiceNow, Jira, and Confluence.
Qualifications

Basic Qualifications

  • Must be a U.S. Citizen with the ability to obtain and maintain the required Public Trust level clearance.
  • Bachelor's Degree and 8 years of experience, a Master's Degree and 6 years of experience, or a High School diploma or equivalent and 12 years.
  • 4+ years hands-on experience with enterprise observability, monitoring, and alerting platforms.
  • Hands-on experience with Dynatrace, Datadog, and/or Splunk, including instrumentation, dashboarding, and alert design.
  • Working knowledge of metrics, logs, and traces, and experience with OpenTelemetry-based instrumentation.
  • Experience operating in AWS and/or AWS GovCloud; familiarity with containerized and cloud-native workloads.
  • Experience with Ansible / Ansible Tower / AAP for automated agent and configuration deployment.
  • Experience with GitLab CI/CD (and Jenkins) for building observability-as-code pipelines.
  • Scripting/automation proficiency in Python, Bash, or Go.

Preferred Qualifications

  • Observability platform certifications (Dynatrace, Datadog, or Splunk) or cloud architecture certifications.
  • Production experience with OpenTelemetry Collector deployment at scale.
  • Kubernetes and container observability, service-mesh telemetry, or eBPF-based collection experience.
  • Experience with AIOps or AI-assisted anomaly detection, event correlation, or incident summarization.
  • Experience in federal or regulated environments (FISMA, FedRAMP, NIST 800-53).
Peraton Overview

Peraton is a next-generation national security company that drives missions of consequence spanning the globe and extending to the farthest reaches of the galaxy. As the world’s leading mission capability integrator and transformative enterprise IT provider, we deliver trusted, highly differentiated solutions and technologies to protect our nation and allies. Peraton operates at the critical nexus between traditional and nontraditional threats across all domains: land, sea, space, air, and cyberspace. The company serves as a valued partner to essential government agencies and supports every branch of the U.S. armed forces. Each day, our employees do the can’t be done by solving the most daunting challenges facing our customers. Visit peraton.com to learn how we’re keeping people around the world safe and secure.

Target Salary Range$104,000 - $166,000. This represents the typical salary range for this position. Salary is determined by various factors, including but not limited to, the scope and responsibilities of the position, the individual’s experience, education, knowledge, skills, and competencies, as well as geographic location and business and contract considerations. Depending on the position, employees may be eligible for overtime, shift differential, and a discretionary bonus in addition to base pay. EEOEEO: Equal opportunity employer, including disability and protected veterans, or other characteristics protected by law.

Similar Jobs

25 Minutes Ago
Easy Apply
Remote or Hybrid
USA
Easy Apply
105K-127K Annually
Mid level
105K-127K Annually
Mid level
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Software • Big Data Analytics • Automation
Owns a defined territory of mid-sized customer accounts, driving revenue expansion through upselling, cross-selling, consultative selling, and executive engagement. Develops territory and account plans, identifies growth opportunities, manages Salesforce pipeline forecasts, promotes PagerDuty Operations Cloud adoption, and collaborates with solution consultants, customer success, renewals, and product teams. Requires B2B SaaS sales experience, quota attainment, C-level selling skills, and familiarity with Salesforce and complex sales methodologies.
Top Skills: Pagerduty Operations CloudSalesforce
25 Minutes Ago
Easy Apply
Remote or Hybrid
Easy Apply
105K-127K Annually
Mid level
105K-127K Annually
Mid level
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Software • Big Data Analytics • Automation
Owns a territory of mid-sized SaaS customer accounts, driving expansion through upselling, cross-selling, consultative selling, and executive engagement. Responsibilities include developing territory and account plans, identifying growth opportunities, managing Salesforce pipeline forecasts using MEDDICC, increasing Operations Cloud adoption, and collaborating with solution consultants, customer success, renewals, and product teams. The role requires 3–5 years of B2B SaaS or enterprise software sales experience and is based in a hybrid San Francisco location.
Top Skills: DevOpsPagerduty Operations CloudSalesforce
25 Minutes Ago
Easy Apply
Remote or Hybrid
USA
Easy Apply
131K-171K Annually
Senior level
131K-171K Annually
Senior level
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Software • Big Data Analytics • Automation
Own and expand a portfolio of enterprise accounts through upselling, cross-selling, consultative selling, and complex multi-product SaaS deals. Build executive relationships, create strategic account plans, articulate business value and ROI, manage forecasts and Salesforce pipeline using MEDDICC, and collaborate with solution consultants, customer success, product, and renewals teams to drive adoption, retention, and growth.
Top Skills: Cloud SoftwareDevOpsOperations CloudSaaSSalesforce

What you need to know about the Seattle Tech Scene

Home to tech titans like Microsoft and Amazon, Seattle punches far above its weight in innovation. But its surrounding mountains, sprinkled with world-famous hiking trails and climbing routes, make the city a destination for outdoorsy types as well. Established as a logging town before shifting to shipbuilding and logistics, the Emerald City is now known for its contributions to aerospace, software, biotech and cloud computing. And its status as a thriving tech ecosystem is attracting out-of-town companies looking to establish new tech and engineering hubs.

Key Facts About Seattle Tech

  • Number of Tech Workers: 287,000; 13% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Amazon, Microsoft, Meta, Google
  • Key Industries: Artificial intelligence, cloud computing, software, biotechnology, game development
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Madrona, Fuse, Tola, Maveron
  • Research Centers and Universities: University of Washington, Seattle University, Seattle Pacific University, Allen Institute for Brain Science, Bill & Melinda Gates Foundation, Seattle Children’s Research Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account