Job Description
Description
We are seeking a highly motivated Site Reliability Engineer (SRE) to help build, scale, and maintain cloud infrastructure, CI/CD pipelines, and deployment automation. This role is responsible for improving system reliability, operational efficiency, and developer experience through automation, observability, and modern reliability engineering practices. The ideal candidate is passionate about solving complex technical challenges, driving continuous improvement, and enabling high-performing engineering teams. Key Responsibilities:
- Design, build, and maintain scalable cloud infrastructure, CI/CD pipelines, and deployment automation solutions that support secure, reliable, and efficient software delivery
- Implement Infrastructure as Code (IaC), observability, monitoring, and service reliability practices, including SLOs and SLIs, while driving automation and reducing operational toil across the development lifecycle
- Partner closely with engineering, product, and support teams to improve system reliability, streamline release processes, support incident response and post-mortem activities, and eliminate developer bottlenecks
- Continuously evaluate emerging cloud, DevOps, and AIOps technologies, participate in on-call support, and mentor team members by promoting best practices, technical excellence, and a culture of continuous improvement
Requirements
- 2-4 years of relevant work experience
- Experience with private, public and hybrid cloud environments
- Hands-on experience establishing Observability using Prometheus, Grafana, and OpenTelemetry
- Proven software development experience with a focus on building internal tooling and applying AIOps to streamline and accelerate development
- Experience with object-oriented programming (preferably Java), cloud architecture, CI/CD pipelines, and modern design patterns
- Experience with FinOps practices, including cloud cost allocation and resource optimization
- Experience with security frameworks for user and services authorization and authentication
- Experience with destructive and performance test design and execution
- Experience with modern debugging and root cause analysis techniques
- Experience with version control and code repositories, such as GitHub
Are you interested in this position?
Apply by clicking on the “Apply Now” button below!
#GraphicDesignJobsOnline
#WebDesignRemoteJobs
#FreelanceGraphicDesigner
#WorkFromHomeDesignJobs
#OnlineWebDesignWork
#RemoteDesignOpportunities
#HireGraphicDesigners
#DigitalDesignCareers
# Dynamicbrand guru