Very Good Security Logo

Very Good Security

Sr. Infrastructure Engineer

Reposted 21 Days Ago
Remote
Hiring Remotely in United States
145K-185K Annually
Senior level
Remote
Hiring Remotely in United States
145K-185K Annually
Senior level
The Sr. Infrastructure Engineer will architect and maintain scalable infrastructure, lead incident management, improve operational processes, and mentor junior engineers.
The summary above was generated by AI

VGS is the world's leader in payment tokenization.  Large banks, aspiring fintechs, and growing merchants embed our universal token vault into their technology stack to manage the complexities of payment data tokenization across processors and networks, open banking, card issuance, omnichannel loyalty, PCI compliance, payment orchestration, and more. We empower our clients and partners by tokenizing sensitive payment data, limiting compliance scope, and consolidating payments to unlock revenue and business opportunities. 

VGS provides processor-agnostic tokenization solutions via secure universal token vaults, iframes, mobile SDKs, tokenization proxies, APIs, and data orchestration tooling to support payment acceptance, card issuance, PII and bank account tokenization, and other payments value-added services. Some of the use cases we enable include multi-processor Network Tokenization, Account Updater, payment orchestration, secure settlement file processing, 3DS, and Risk provider connectivity.

We are looking for a well-versed, passionate Engineer who wants to play a key role in site reliability engineering and cloud operations of our global cloud infrastructure.

We’re seeking individuals with creative problem-solving, enthusiasm for new technologies, and a desire to contribute to our product. You will likely be successful in this role if you identify with the following traits: attention to detail, problem solver, customer-oriented, versatile, resilient, and confident.  If all of this sounds interesting to you, we’d love to hear from you.

What you will be doing at VGS (Responsibilities)...

  • Architect and maintain scalable, reliable infrastructure: Design and optimize infrastructure for high availability, fault tolerance, and performance across distributed systems.

  • Lead incident management and root cause analysis: Own incident response processes, ensure swift resolution of issues, and drive post-incident improvements to prevent recurrences.

  • Service monitoring and automation: Build and maintain automated monitoring, alerting, and healing systems that improve system health, reduce manual intervention, and minimize downtime.

  • Performance tuning and capacity planning: Identify bottlenecks and optimization opportunities, and implement scaling strategies to handle traffic spikes and growing workloads efficiently.

  • Collaborate with cross-functional teams: Work closely with software engineers, product teams, and DevOps to enhance system reliability and delivery pipelines.

  • Improve operational processes: Champion continuous improvement initiatives in deployment, scaling, and performance testing, while advocating for the adoption of SRE best practices across the organization.

  • Mentorship and leadership: Provide technical mentorship to junior engineers, contribute to strategic decisions around infrastructure, and ensure best practices are implemented at scale.

  • Be proactive and innovative:  we rely on your feedback to build a world-class product.

  • Be a part of a team that believes in the core values of transparency, collaboration, grit, and humility; in going above and beyond what is required to do the right thing for our customers and the company; and in having fun while doing all this!

What we are looking for from you (Requirements)...

  • On-call support:  Go on call for the services owned and operated by the team.
  • Proven experience in Infrastructure/SRE roles, with a track record of managing production systems in complex, large-scale environments.

  • Strong proficiency in AWS, including infrastructure-as-code (Terraform, CloudFormation, etc.).

  • Solid understanding of cloud-native architecture, Linux Systems, microservices, Infrastructure-as-code (Terraform, CloudFormation, CDK), CI/CD (CircleCI, GitHub Actions, Argo), GitOps, Authentication and Authorization,  APIs and API Gateway, Docker, Kubernetes (EKS), Kafka (MSK), Java, Spring Framework, Python,  and AWS services.

  • Strong plus if you are a database wiz.

  • Expertise in monitoring and observability tools like Prometheus, Grafana, Open Telemetry or similar tools to measure system health and performance.

  • Programming and scripting experience in languages such as Python, Go, Bash, or other relevant languages used in automating infrastructure.

  • Solid understanding of networking, security, and load balancing in cloud-native environments.

  • Strong communication and collaboration skills, with the ability to lead cross-functional initiatives and mentor junior team members.

  • Experience with incident management and disaster recovery best practices.

  • Strong written and verbal communication skills.

What you get from us...
 
• Flexible work hours and flexible PTO
• Competitive health benefits
• VGS stock options
• 401k plan, with employer matching 4% and immediate vesting (available only for US employees)
• Life & disability insurance
• Pre-tax flexible spending accounts, dependent and healthcare FSA (available only for US employees)
• Global parental leave program
• Employee Assistance Program
• Home Internet reimbursement
• New hire home office set-up allowance
• Professional learning reimbursement
 
At VGS, we have a remote-first philosophy because we believe flexibility leads to great work and a healthy work-life balance. That said, if you live within 30 miles of one of our office locations, you’ll be on a hybrid schedule with some in-person time—because we know there’s real value in coming together.
 
We’re not about being in the office every day—but we are about connection, collaboration, and the energy that comes from a great brainstorm, a team lunch, or celebrating a big win in person.
 
We consider applicants without regard to race, color, national origin, sex, age, religion, sexual orientation, gender identity, veteran status, marital status, physical or mental disability, or other protected classes under all local, state, and federal laws and ordinances (AA/EOE/W/M/Vet/Disabled).
 
Qualified applicants with arrest and conviction records will be considered for the position in accordance with the San Francisco Fair Chance Ordinance.
 
 
 
Please note we are currently only hiring in the following states...
 
Arizona, California, Colorado, Connecticut, Florida, Idaho, Illinois, Iowa, Michigan, Minnesota, New York, North Carolina, Ohio, Oregon, Pennsylvania, Texas, Utah, Virginia, Washington, Ontario (Canada), Alberta (Canada), or British Columbia (Canada).

Similar Jobs

11 Days Ago
Easy Apply
Remote
Easy Apply
126K-314K Annually
Senior level
126K-314K Annually
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Maintain and improve reliability, scalability, and automation for user-facing production systems. Build infrastructure tooling, operate Kubernetes-based services, write IaC, participate in on-call and incident response, and advance observability and runbooks to reduce toil and improve platform reliability.
Top Skills: AWSCi/CdGCPGitopsGoInfrastructure As Code (Iac)KubernetesKubernetes Operators/ControllersLoggingMetricsRubySlos/SlisTerraform
Yesterday
Remote
USA
175K-200K Annually
Senior level
175K-200K Annually
Senior level
Blockchain • Fintech • Financial Services
Design, build, and operate scalable, low-latency trading infrastructure processing millions of orders daily. Implement exchange integrations, trading algorithms, and performance improvements. Collaborate with product, execute system architecture roadmap, write tests, and troubleshoot complex distributed-system issues.
Top Skills: C#C++GoJavaSnowflakeSQL
5 Days Ago
Remote
190K-258K Annually
Senior level
190K-258K Annually
Senior level
Artificial Intelligence • Cloud • Consumer Web • Productivity • Software • App development • Data Privacy
Design, build, and operate large-scale distributed storage systems ensuring durability, availability, and performance. Implement replication, erasure coding, and lifecycle management; write high-quality Go and Rust code; participate in on-call rotations; troubleshoot production issues; collaborate across networking, hardware, and capacity teams; own scoped projects and drive infrastructure architecture and reliability improvements.
Top Skills: C++CephDistributed Storage SystemsErasure CodingFile SystemsGfs/ColossusGoObservabilityProduction MonitoringReplication ProtocolsRustS3

What you need to know about the Seattle Tech Scene

Home to tech titans like Microsoft and Amazon, Seattle punches far above its weight in innovation. But its surrounding mountains, sprinkled with world-famous hiking trails and climbing routes, make the city a destination for outdoorsy types as well. Established as a logging town before shifting to shipbuilding and logistics, the Emerald City is now known for its contributions to aerospace, software, biotech and cloud computing. And its status as a thriving tech ecosystem is attracting out-of-town companies looking to establish new tech and engineering hubs.

Key Facts About Seattle Tech

  • Number of Tech Workers: 287,000; 13% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Amazon, Microsoft, Meta, Google
  • Key Industries: Artificial intelligence, cloud computing, software, biotechnology, game development
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Madrona, Fuse, Tola, Maveron
  • Research Centers and Universities: University of Washington, Seattle University, Seattle Pacific University, Allen Institute for Brain Science, Bill & Melinda Gates Foundation, Seattle Children’s Research Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account