Very Good Security Logo

Very Good Security

Sr. Staff Infrastructure Engineer

Posted An Hour Ago
Be an Early Applicant
Remote
Hiring Remotely in United States
145K-220K Annually
Senior level
Remote
Hiring Remotely in United States
145K-220K Annually
Senior level
Lead platform engineering for high-scale, multi-region AWS payment infrastructure. Architect immutable, highly available, self-healing systems; automate infrastructure through Terraform, GitOps, CI/CD, and Kubernetes; establish observability with Prometheus, Grafana, and OpenTelemetry; own incident management and resiliency improvements; design private connectivity; and mentor engineers while driving organization-wide platform standards.
The summary above was generated by AI
About VGS

VGS is the world's leader in payment tokenization, trusted by the most innovative AI and Fortune 500 companies, merchants, banks, and fintechs to power modern payments and agentic commerce.

We tokenize over 9 billion tokens and process more than 10 billion monthly interactions, touching a third of all e-commerce. That scale isn't incidental. It's why leading companies embed our universal token vault and credential management platform directly into their stack, taming the complexity of payment data so they can move faster.

We’re helping businesses unlock new possibilities in an industry that never stops moving. And that takes exceptional people.

 


Trusted Engineering That Matters

Engineering at VGS

Our vision is clear: VGS is building the platform powering agentic commerce and global payments infrastructure. VGS Engineering is the pillar that supports that vision. We create a platform that is an indispensable, ubiquitous component of our partners' payments and security infrastructure. To achieve this, we build and operate high-scale, reliable, and resilient systems that manage mission-critical payment data across the globe. We don't just participate in the payments ecosystem; we power it.

 

Our Environment & Culture

  • Tooling as Leverage for Impact: We value speed and impact. To ensure nothing slows you down, we arm you with a highly consistent modern stack (e.g. AWS, Kubernetes, Java, TypeScript, Python) and world-class AI-augmented SDLC tooling. We strip away the friction so you can focus on building simple things that work, pushing boundaries, and delivering value faster.

  • Impact Over Input: Scaling global payments infrastructure requires individuals who take personal ownership of outcomes from day one. We look for builders who combine high agency with deep strategic alignment; channeling their urgency and relentless iteration into what matters most. Leading from the front means multiplying the impact of those around you so we win collectively.

  • Mission-Critical Accountability: We power payments infrastructure for global market leaders, making system trust and integrity our core deliverable. We don't treat compliance (PCI, SOC2, ISO27001), security patching, and platform maintenance as administrative overhead—we engineer them with first-class product discipline. We apply the same rigor, roadmap thinking, and speed to system health as we do to new features. In high-scale payments, reliability isn't background work; it is our product.

 

Learn More: See how our engineering practices drive trust and innovation in our official blog and our reliability practices.

 


About the Role

As a Senior Infrastructure Engineer (IC6), you will serve as a technical leader on our Platform Engineering team. You will architect, scale, and fortify global cloud infrastructure designed to handle mission-critical, high-throughput payments applications with zero downtime.

You will take ownership of key platform foundations powering our core payments infrastructure. Rather than executing against a rigid task list or holding blanket ownership over the entire platform, you will drive the technical strategy, architecture, and reliability standards for your designated domains within our multi-region AWS environment. We are looking for high-agency engineers who want to own complex distributed system challenges end-to-end and elevate how the entire engineering team operates.

If you thrive on solving complex distributed systems problems, building automated resiliency, and driving modern SRE practices, we want to build the future with you.

 


What You'll Do
  • Build Immutable, Self-Healing Systems: Design, build, and optimize multi-region, high-availability AWS infrastructure. You will drive our evolution from hand-crafted environments to a standardized, globally scalable fleet managed entirely through code.

  • Drive Resiliency & Automation: Replace manual toil with self-healing, automated infrastructure using GitOps, modern CI/CD pipelines, and IaC.

  • Deep Observability & Resiliency: Build end-to-end telemetry (Prometheus, Grafana, OpenTelemetry) to proactively spot bottlenecks. You will own incident management and conduct blameless post-mortems to continuously harden our reliability baseline.

  • Force-Multiply Engineering Velocity: Partner closely with Product, Security, and Core Engineering teams and lead from the front by designing "Golden Paths" that strip away friction for feature teams. You will influence company-wide engineering practices and mentor the organization on how to move fast with high alignment.

  • Customer Impact: Architect and operate high-performance, low-latency private connectivity to optimize the experience for external customers. Partner strategically with internal engineering teams at the design and architectural level for platform enablement and adoption.

 


What You Bring
  • Ownership at Scale: 10+ years of experience taking personal ownership of outcomes in complex, large-scale distributed systems within mission-critical environments.

  • AWS & Infrastructure-as-Code: Advanced proficiency in AWS ecosystems leveraging Terraform to build reproducible environments.

  • Containerization & Orchestration: Strong, hands-on experience with Kubernetes (EKS), Docker, and GitOps workflows (Flux, Argo, GitHub Actions).

  • Automation & Scripting: Strong coding skills in Python, Go, or Bash to automate infrastructure and build operational tools.

  • Observability Expertise: Deep experience implementing Prometheus, Grafana, or OpenTelemetry at scale.

  • Security & Networking Foundations: Solid understanding of cloud security, API Gateways, load balancing, and network isolation, viewing security as a fundamental engineering constraint, not an afterthought.

Nice to Have:

  • Experience with tokenization, payment processing, cryptology, or security products

  • BA/BS degree

  • A knack for out-of-the-box thinking that thrives in a fast-paced startup environment

  • Experience managing distributed data streaming platforms like Kafka (MSK).

  • Database performance tuning and query optimization skills.

  • Familiarity with Java / Spring Framework services.

 

Similar Jobs

27 Days Ago
Easy Apply
Remote
Easy Apply
126K-314K Annually
Senior level
126K-314K Annually
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Maintain and improve reliability, scalability, and automation for user-facing production systems. Build infrastructure tooling, operate Kubernetes-based services, write IaC, participate in on-call and incident response, and advance observability and runbooks to reduce toil and improve platform reliability.
Top Skills: AWSCi/CdGCPGitopsGoInfrastructure As Code (Iac)KubernetesKubernetes Operators/ControllersLoggingMetricsRubySlos/SlisTerraform
11 Days Ago
Remote or Hybrid
157K-234K Annually
Mid level
157K-234K Annually
Mid level
Transportation
The role involves designing and implementing MLOps pipelines, optimizing machine learning processes, and ensuring model performance and compliance in a collaborative environment.
Top Skills: SparkAWSAzureDockerGitGCPHadoopKubernetesPythonPyTorchTensorFlowTerraform
2 Hours Ago
Remote
145K-183K Annually
Senior level
145K-183K Annually
Senior level
Fintech • Financial Services
Conduct manual penetration testing, secure code reviews, and SAST triage across web applications, APIs, and AWS infrastructure. Develop AI-assisted security tooling, define testing approaches for LLM applications and agentic systems, and review AI-generated software before production. Communicate vulnerabilities, risks, and remediation priorities to engineering teams while improving secure SDLC practices.
Top Skills: Ai AgentsAWSGoLlmsPythonRubySastTypescript

What you need to know about the Boston Tech Scene

Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.

Key Facts About Boston Tech

  • Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
  • Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
  • Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
  • Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account