Job Description
Accountabilities
As a Manager, Cloud Infrastructure Engineering, you will lead the team responsible for building and operating cloud infrastructure that supports a global SaaS platform. You will combine hands-on technical expertise with strong leadership skills to improve reliability, scalability, security, and engineering effectiveness.
- Lead and develop a team of engineers responsible for production cloud infrastructure, fostering a strong engineering culture and effective execution.
- Guide the design, implementation, and evolution of Kubernetes-based, multi-cloud infrastructure environments.
- Ensure infrastructure solutions meet high standards for reliability, scalability, security, and operational excellence.
- Collaborate with product engineering, security, support, and engineering leadership teams to enable safe and efficient software delivery.
- Improve infrastructure automation, developer tooling, and platform capabilities to support growth without unnecessary complexity.
- Establish effective approaches for using AI in infrastructure operations, including automation, incident analysis, documentation, and operational improvements.
- Balance immediate operational needs with long-term investments in platform evolution, cloud expansion, and infrastructure innovation.
- Coach engineers, support professional growth, manage priorities, and create an environment where teams can perform at their best.
- Drive improvements in observability, incident response, reliability practices, and infrastructure decision-making processes.
Requirements
The ideal candidate is an experienced cloud infrastructure leader with a strong background in platform engineering, Kubernetes, and SaaS environments. You should be comfortable managing technical teams while remaining close to engineering challenges and emerging infrastructure practices.
- 3-5 years of experience leading engineers who build and operate production infrastructure for SaaS products.
- Strong technical background in cloud infrastructure engineering, platform engineering, Site Reliability Engineering (SRE), or related disciplines.
- Hands-on production experience with Kubernetes, including scaling, reliability management, and cluster lifecycle operations.
- Experience working with multiple major cloud providers and understanding the challenges and trade-offs of multi-cloud environments.
- Proven ability to collaborate with senior engineers and cross-functional stakeholders to deliver major infrastructure improvements.
- Strong people leadership skills, including coaching, performance management, prioritization, and team development.
- Experience improving infrastructure reliability, scalability, security, and automation practices.
- Clear understanding of how AI can support infrastructure engineering workflows, with good judgment around responsible and effective adoption.
- Experience with infrastructure-as-code, policy-as-code, platform automation, observability, incident management, or cloud cost optimization is a plus.
- Experience managing global or multi-region SaaS infrastructure is considered an advantage.
Are you interested in this position?
Apply by clicking on the āApply Nowā button below!
#GraphicDesignJobsOnline
#WebDesignRemoteJobs
#FreelanceGraphicDesigner
#WorkFromHomeDesignJobs
#OnlineWebDesignWork
#RemoteDesignOpportunities
#HireGraphicDesigners
#DigitalDesignCareers
# Dynamicbrand guru