VP Engineering - Infrastructure & SRE

Jobgether· Brussels (Firmensitz, recherchiert)· lever· offentliggjort 31.07.2026
Skal:AWSKubernetesCloudDevOpsCI/CDAIFinTechSecuritySeniorLeadRemote

Accountabilities: As the VP Engineering - Infrastructure & SRE, you will own the technical vision and execution strategy for infrastructure, platform engineering, and site reliability operations. You will ensure the reliability, scalability, security, and efficiency of mission-critical systems while building a culture of engineering excellence and continuous improvement.

Define and execute the infrastructure vision, strategy, and multi-year roadmap to support platform growth, operational resilience, and compliance requirements.

Lead, mentor, and scale infrastructure, platform, and SRE teams by developing talent, improving processes, and fostering a culture of ownership and operational excellence.

Own platform reliability by establishing service level objectives (SLOs), error budgets, incident management practices, and measurable reliability improvements.

Participate in and oversee the on-call rotation, acting as a senior incident leader during critical events by driving investigation, communication, mitigation, and long-term resolution.

Direct cloud infrastructure strategy, including AWS architecture, account structures, networking, identity management, multi-region capabilities, and cost optimization initiatives.

Own Kubernetes platform strategy, including architecture, upgrades, workload management, scaling, developer experience, and operational best practices.

Manage database infrastructure strategy, including availability, performance, migrations, backup processes, recovery planning, and scalability.

Design and continuously improve disaster recovery and business continuity programs, including testing, failover exercises, and recovery objectives.

Drive AI-enabled improvements across infrastructure operations by implementing tools for incident response, automation, observability, and reduction of operational workload.

Partner with security and compliance teams to strengthen infrastructure controls, audit readiness, vulnerability management, and regulatory requirements.

Establish infrastructure-as-code and automation as standard practices across engineering workflows.

Own technology strategy, vendor relationships, infrastructure investments, and build-versus-buy decisions.

Communicate infrastructure risks, investments, and operational performance clearly to executives, auditors, and business stakeholders.

Requirements:

The ideal candidate is a seasoned infrastructure and engineering leader with extensive experience operating large-scale, highly available platforms. You should combine deep technical expertise with proven leadership capabilities, strong business judgment, and the ability to drive transformation across engineering organizations.

15+ years of experience across infrastructure, platform engineering, site reliability, software development, or related engineering disciplines.

5+ years of experience leading engineering teams, including hiring, coaching, organizational design, and performance development.

Extensive hands-on expertise designing and operating production AWS environments, including compute, networking, IAM, multi-account architectures, and cloud cost management.

Deep production experience with Kubernetes, including cluster operations, workload architecture, scaling strategies, and platform reliability.

Strong expertise with relational databases at scale, particularly Aurora RDS, MySQL, and/or PostgreSQL, including high availability, replication, performance tuning, and recovery strategies.

Proven experience owning disaster recovery and business continuity programs with measurable recovery objectives and tested failover processes.

Demonstrated ability to operate and improve 24/7 high-availability platforms where reliability directly impacts customers and revenue.

Strong incident management experience, including leading production recovery efforts through data-driven troubleshooting and decisive action under pressure.

Ability to remain technically engaged through architecture reviews, system design discussions, debugging, and engineering decisions.

Experience managing significant cloud infrastructure budgets while improving efficiency without compromising reliability.

Strong knowledge of infrastructure-as-code practices, including Terraform or equivalent technologies.

Experience with modern CI/CD practices, automation, and progressive deployment strategies.

Proven ability to recruit, develop, and retain exceptional infrastructure and SRE talent.

Bachelor’s degree in Computer Science, Engineering, or a related technical field.

Experience with PCI-DSS, SOC 2, fintech, payments, banking, or other regulated environments is highly preferred.

Experience implementing AI-powered operations tools, LLM-based workflows, intelligent monitoring, or automation solutions is a strong advantage.

Familiarity with observability platforms such as Prometheus, Grafana, Loki, Tempo, or similar solutions.

Experience with multi-region architectures, chaos engineering, service mesh, secrets management, and zero-trust networking is preferred.

Benefits:

Competitive total compensation package with an estimated range of $400,000 - $600,000 USD annually .

Equity opportunities and retirement benefits including a 401(k) match.

Comprehensive medical, dental, and vision insurance coverage.

Life, short-term disability, and long-term disability insurance.

Unlimited paid time off, volunteer hours, and sabbatical opportunities.

Opportunity to work remotely from the United States.

Access to professional growth and development opportunities.

Highly discounted gym membership benefits.

Opportunity to join a fast-growing technology company in the fintech space.

Collaborative environment with experienced engineers, innovators, and technology leaders.

Chance to influence the future of infrastructure, reliability engineering, and AI-driven operations at scale

How Jobgether works: We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team. We appreciate your interest and wish you the best!  Why Apply Through Jobgether? 

Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.

#LI-CL1