Sr Lead Network Reliability Engineer
Accountabilities: Design, deploy, automate, and support highly scalable network infrastructure across AWS, Azure, and GCP, applying established cloud architecture and networking standards.
Architect and operate firewall, VPN, DNS, BGP, TLS, routing, security group, and hybrid-cloud networking solutions, including Palo Alto, Fortinet, Check Point, and cloud-native firewall technologies.
Provide network engineering leadership for production Kubernetes platforms and workloads, including Kubernetes networking, CNI implementations, ingress and gateway controllers, service discovery, DNS, and NetworkPolicy.
Develop and execute network infrastructure roadmaps while participating in architecture and design reviews across the broader engineering organization.
Apply an automation-first and development-oriented mindset to network operations, using infrastructure as code to reduce manual work, eliminate operational toil, and improve service reliability.
Build and maintain network performance, capacity, availability, and monitoring solutions, using advanced troubleshooting techniques such as packet capture and traffic-flow analysis.
Lead and support network initiatives from conception through delivery and operational ownership, ensuring projects and issues are driven to successful resolution.
Help grow and mature the global network engineering function by providing technical leadership, mentoring new engineers, sharing knowledge, and contributing to team direction.
Stay current with evolving networking technologies, cloud infrastructure practices, and industry developments, incorporating relevant innovations into the network strategy.
Actively leverage AI tools and methodologies to improve engineering workflows, accelerate problem solving, and drive infrastructure innovation.
Participate in an on-call rotation and provide operational support for critical network services when required.
Travel occasionally as needed to fulfill the responsibilities of the position.
Requirements:
Bachelor’s degree in Computer Science or a related field, or equivalent experience, with 10+ years of professional engineering experience; experience in an enterprise SaaS environment is highly valued.
Deep expertise in enterprise networking, including firewalls, VPNs, DNS, BGP, TCP/IP, TLS, public-cloud networking, VPCs, hybrid routing, and security groups across AWS, Azure, and/or GCP.
Proven experience managing and maintaining production Kubernetes environments, with advanced knowledge of Kubernetes networking and CNI implementations; experience with multiple CNI solutions is a plus.
Strong understanding of ingress and gateway controller architectures, service discovery, DNS within Kubernetes, and NetworkPolicy functionality.
Experience operating across both traditional data center and public-cloud environments, as well as supporting multiple Linux distributions.
Strong programming or scripting capabilities in Golang, Python, Java, Ruby, or an equivalent language.
Hands-on experience deploying and managing cloud network infrastructure using Terraform, CloudFormation, Chef, Ansible, Spacelift, or comparable infrastructure-as-code and configuration-management technologies.
Advanced network troubleshooting skills, including packet capture analysis, traffic-flow diagnostics, and root-cause investigation.
Demonstrated ability to independently own complex projects, make sound technical decisions, and drive issues through resolution.
Strong problem-solving, analytical, and critical-thinking skills, with a willingness to investigate complex technical challenges in depth.
Ability to provide technical leadership, influence engineering teams, mentor colleagues, and communicate effectively across collaborative environments.
A strong service-ownership mindset, attention to reliability and operational excellence, and enthusiasm for continuous improvement.
Comfortable working autonomously in a fast-moving environment while collaborating closely with engineering and infrastructure teams.
Experience incorporating AI tools into engineering workflows and using emerging technologies to improve productivity and innovation.
Benefits:
Annual salary range of approximately $149,000–$208,333 , with the applicable range and starting compensation determined by skills, experience, and geographic location.
Comprehensive, location-specific benefits packages that may include medical, dental, retirement or pension plans, and life or accident protection.
Two company-wide paid wellness days each year.
Paid birthday time off during your birthday month.
40 hours of paid Volunteer Time Off (VTO) annually.
Employee Assistance Program providing confidential, 24/7 support for emotional, financial, legal, and work-life needs.
Business travel protection and assistance.
Employee referral bonus opportunities.
Opportunities for professional development, including internal training and access to learning and certification resources.
An autonomous, collaborative, and innovation-focused working environment.
Opportunity to influence large-scale cloud infrastructure and work with modern networking, Kubernetes, automation, and AI technologies.
How Jobgether works: We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team. We appreciate your interest and wish you the best! Why Apply Through Jobgether?
Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.
#LI-CL1