Ciklum Logo

Ciklum

Senior Site Reliability Engineer

Posted 4 Days Ago
Be an Early Applicant
Remote
Hiring Remotely in Ukraine
Senior level
Remote
Hiring Remotely in Ukraine
Senior level
Designs and automates scalable cloud infrastructure for a SaaS platform using IaC and GitOps. The role manages observability, monitoring, alerting, incident response, post-mortems, capacity planning, performance optimization, and production reliability. It also involves Linux systems administration, large-scale data platform architecture, automation, on-call support, and collaboration with development and research teams.
The summary above was generated by AI

Ciklum is looking for a Senior Site Reliability Engineer to join our team full-time in Ukraine.

We are a custom product engineering company that supports both multinational organizations and scaling startups to solve their most complex business challenges. With a global team of over 4,000 highly skilled developers, consultants, analysts and product owners, we engineer technology that redefines industries and shapes the way people live.

About the role:

As a Senior Site Reliability Engineer, become a part of the R&D team. In this role, you'll be a key contributor to our high-performance SaaS Cloud Platform.

As a Senior Site Reliability Engineer, you will design and automate scalable infrastructure with infrastructure-as-code, implement and tune monitoring and observability systems to meet defined SLIs and SLOs, lead on-call incident response and post-mortem reviews to continuously improve system resilience, collaborate with development teams on performance optimization and capacity planning, and drive the automation of routine operational tasks to minimize toil and maximize uptime.

Join our innovative, high-performing team, where the convergence of reliable cloud infrastructure and advanced data processes drives our success in a fast-paced, agile environment.

Responsibilities:

  • You’ll develop, improve, and maintain Guardicore's Cyber Security SaaS cloud platform
  • Lead problem-solving efforts for the entire technology stack in collaboration with other teams in the R&D
  • Establish scalable, efficient, automated processes for large-scale data analyses
  • Work closely with other R&D to develop a strategy for long-term data platform architecture
  • Practice infrastructure as a code (IaC) and GitOps using technologies like Terraform and ArgoCD
  • Collaborate with Guardicore's development and research groups to constantly improve our platform and infrastructure
  • Participate in the on-call rotation supporting the applications and infrastructure
  • Developed and evolved our tooling, logging, monitoring, and alerting mechanisms to increase observability and transparency

Requirements:

  • Extensive experience with containerization and orchestration technologies (e.g., Docker, Kubernetes)
  • Excellent problem-solving skills and ability to think critically about complex technical challenges and optimizing production systems
  • Experience in managing and troubleshooting Linux systems
  • experience with observability systems such as Datadog/Splunk/New Relic/Grafana, or similar
  • Experience in Shell scripting and/or high-level Programming like Python and Go
  • Experience working with cloud environments like GCP, Linode, AWS, and Azure
  • Experience in a SaaS environment managing large-scale data sets - Advantage
  • Excellent verbal and written English communication and presentation skills

What’s in it for you?

  • Strong community: Work alongside top professionals in a friendly, open-door environment
  • Growth focus: Take on large-scale projects with a global impact and expand your expertise
  • Tailored learning: Boost your skills with internal events (meetups, conferences, workshops), Udemy access, language courses, and company-paid certifications
  • Endless opportunities: Explore diverse domains through internal mobility, finding the best fit to gain hands-on experience with cutting-edge technologies
  • Flexibility: Enjoy radical flexibility – work remotely or from an office, your choice
  • Care: We’ve got you covered with company-paid medical insurance, mental health support, and financial & legal consultations

About us:

At Ciklum, we are always exploring innovations, empowering each other to achieve more, and engineering solutions that matter. With us, you’ll work with cutting-edge technologies, contribute to impactful projects, and be part of a One Team culture that values collaboration and progress.

As one of Ukraine’s largest IT companies and a top employer recognized by Forbes, we’ve spent over 20 years delivering meaningful tech solutions. We proudly support diverse talent and military veterans, recognizing their unique skills and perspectives they bring to shaping the future.

Explore, empower, engineer with Ciklum!

Interested already? We would love to get to know you! Submit your application. We can’t wait to see you at Ciklum.

#LI-NV1

Similar Jobs

One Month Ago
Remote or Hybrid
Senior level
Senior level
Digital Media • Gaming • Information Technology • Software • Sports • Esports • Big Data Analytics
Build, scale, and maintain cloud and on-premises infrastructure using IaC and configuration management. Implement observability, deployment platforms, and tooling; define infrastructure patterns and standards. Lead incident response, mentor engineers, and collaborate with Security and Networking to improve reliability and performance.
Top Skills: Amazon EksAnsibleApi GatewayCdnChefDnsGoKubernetesLoad BalancingPythonRancherReverse ProxyRubyTerraformVpc
10 Days Ago
Remote
United States
Senior level
Senior level
Edtech • Kids + Family • Sports
Audit infrastructure, deployment pipelines, monitoring, alerting, incident response, on-call practices, and internal tools. Produce actionable audit reports, implement code and configuration fixes, improve SLOs and reliability practices, advise on scalable architecture, and partner with engineers on implementation and handoff. The role is a fully remote, part-time consulting engagement with potential for full-time conversion.
Top Skills: Ai Coding ToolsAWSCi/CdDatadogGrafanaPrometheus
29 Days Ago
In-Office or Remote
Senior level
Senior level
eCommerce • On-Demand • Software • Manufacturing
Lead architecture and operation of cloud infrastructure and Kubernetes (EKS). Drive large-scale automation with Terraform and GitOps, improve reliability and observability (Grafana/Prometheus/Loki/Tempo), participate in on-call incident response, enforce security and cost-optimization practices, and mentor mid-level SREs while partnering with product teams.
Top Skills: ArgocdAtlantisAuroraAWSCiliumEcrEksGitGithub ActionsGrafanaHelmHelm ChartIamImage ScanningJenkinsKubernetesLinuxLokiMimirMongoDBMySQLPostgresPrometheusPythonRdsRedisS3SqsTempoTerraformVpc

What you need to know about the Seattle Tech Scene

Home to tech titans like Microsoft and Amazon, Seattle punches far above its weight in innovation. But its surrounding mountains, sprinkled with world-famous hiking trails and climbing routes, make the city a destination for outdoorsy types as well. Established as a logging town before shifting to shipbuilding and logistics, the Emerald City is now known for its contributions to aerospace, software, biotech and cloud computing. And its status as a thriving tech ecosystem is attracting out-of-town companies looking to establish new tech and engineering hubs.

Key Facts About Seattle Tech

  • Number of Tech Workers: 287,000; 13% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Amazon, Microsoft, Meta, Google
  • Key Industries: Artificial intelligence, cloud computing, software, biotechnology, game development
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Madrona, Fuse, Tola, Maveron
  • Research Centers and Universities: University of Washington, Seattle University, Seattle Pacific University, Allen Institute for Brain Science, Bill & Melinda Gates Foundation, Seattle Children’s Research Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account