Site Reliability Engineer / Release Management & Automation Specialist
- Full-time
Job Description
:
We are looking for a highly motivated, experienced and self-driven technical resource to serve as a Site Reliability Engineer / Release Management & Automation Specialist. This role provides critical support for a strategic enterprise release orchestration platform and delivery operations system. You will be responsible for sustaining current enterprise release orchestration operations while supporting our transition to a future enterprise release management solution. This unique opportunity combines immediate platform sustainment with strategic transformation initiatives, offering a chance to work with modern technologies in a highly rewarding environment.
What will you do?
- Sustain current enterprise release orchestration operations, integrations, automation, governance, and support activities while preserving critical platform knowledge
- Design, develop and support technical solutions that automate agent deployments and release orchestration processes
- Provide strategic technical support for vendor evaluations, proof-of-concepts, and the transition to a future enterprise release management platform
- Participate in pipeline planning, design, develop and support release and orchestration processes
- Provide input to automation policy, standards and processes
- Promote re-use and adoption of solutions across teams through demos and coaching
- Ensure all automations are compliant with SLF security directives and policies
- Create automation playbooks and companion documentation as references to help the team understand the logic and contents of automated tasks
- Identify opportunities for automation and guide on effective ways to implement
- Educate others on the use of automation tools and best practices
- Always operate in strict compliance of security and change management directives
- Balance platform reliability, operational support, stakeholder engagement, and incident resolution with an SRE mindset
What will you need to succeed?
- 3-5 years' experience in automation development (Ansible specifically)
- Kubernetes/OpenShift administration (ideally CKA/CKS certified)
- Linux/Unix system administration
- Python, JavaScript and Ansible automation
- Hands-on experience using Broadcom Continuous Delivery Director (CDD) tool or any other types of enterprise release orchestration and management tool such as Harness.io etc.
- Hands-on coding experience in Java, Apache Tomcat, JAX-RS, Maven, and Gradle
- Hands-on experience working in containers/Kubernetes
- Experience and understanding of monitoring and log ingestion tools (Splunk monitoring and query development preferred)
- Networking, firewalls, proxies, and connectivity troubleshooting
- System architecture, REST API integration and integration design
- Passion and thought leadership for innovation and automation
- Strong SRE mindset focused on reliability, automation, and operational excellence
- Emerging capabilities in AI-assisted operations, AI-driven automation, Model Context Protocol (MCP), and AI agent integration patterns
What’s in it for you?
- Competitive salary and bonus structure informed by market range data
- 20+ days of vacation per year and an innovative sabbatical program
- A friendly, collaborative, and inclusive culture
- Flexible benefits from day one to meet the needs of you and your family
- Pension, stock, and savings programs to support long-term financial security
- Fitness and wellness programs that help you balance work and life