- This role involves a deep understanding of infrastructure as code, automated deployments
- We are looking for someone who can drive improvements in operational excellence
- Your responsibilities will include designing and implementing automated solutions for infrastructure provisioning
- strong programming skills, preferably in Python or Go, for developing automation tools
- You will collaborate closely with development teams to embed reliability practices throughout
- Experience with Kubernetes, public cloud platforms (AWS, Azure, GCP), and CI/CD pipelines
- strong programming skills, preferably in Python or Go, for developing automation tools
- Kubernetes
- AWS
- Azure
- GCP
- Python
AI-generated summary of what this role requires — see the full description below for the employer's original text.
Audacity is seeking a Staff Site Reliability Engineer to join our growing SRE team. You will be instrumental in ensuring the reliability, scalability, and performance of our production systems across multiple cloud environments. This role involves a deep understanding of infrastructure as code, automated deployments, monitoring, alerting, and incident response. We are looking for someone who can drive improvements in operational excellence and contribute to a resilient and observable platform.
Your responsibilities will include designing and implementing automated solutions for infrastructure provisioning and management, developing robust monitoring and alerting systems, and participating in on-call rotations to support our services. You will also conduct blameless post-mortems, identify root causes of incidents, and implement preventative measures. Experience with Kubernetes, public cloud platforms (AWS, Azure, GCP), and CI/CD pipelines is highly valued.
This role requires strong programming skills, preferably in Python or Go, for developing automation tools and integrating various systems. You will collaborate closely with development teams to embed reliability practices throughout the software development lifecycle, from design to deployment. A proactive approach to identifying and mitigating potential issues before they impact users is key. Join us to build and maintain the high-availability infrastructure that powers Audacity.

