Director of Cloud Operations

Firstup Remote, United States Publicerat 21 juli 2026
full_timeremotemid
<img src="https://we-work-remotely.imgix.net/logos/0171/5900/logo.gif?ixlib=rails-4.0.0&w=50&h=50&dpr=2&fit=fill&auto=compress" /> <p> <strong>Headquarters:</strong> Remote - US <br /><strong>URL:</strong> <a href="http://firstup.io">http://firstup.io</a> </p> <div class="content"><div class="section-wrapper page-full-width"><div class="section page-centered"><div><b>Who We Are</b></div><div><br></div><div>At Firstup, our mission is to improve the employee experience at every moment that matters, large and small. As the communication pipeline for the world's workforce, we now serve 40 of the Fortune 100 companies, reaching and connecting more than 17 million employees daily.</div><div><br></div><div>Our employees are experts in the employee experience, workforce communications and technology. </div><div>Joining Firstup means joining a movement to make work better for every worker. As the world’s first intelligent communication platform, Firstup meaningfully engages employees at every moment from hire to retire, and delivers engagement insights to help companies support, promote and retain their talent. Our movement has taken root and is evident in our world-class customer base. Now we need your help. Ready to make a difference in the world?</div><div><br></div><div><br></div><div> <p><strong><span>Job Summary: </span></strong></p> <p>We are seeking a Director of Cloud Operations (CloudOps) to lead and evolve our cloud infrastructure and operational practices across a globally distributed SaaS platform. This is a hands-on leadership role responsible for ensuring the reliability, scalability, and efficiency of our systems running across multiple AWS regions in the United States and Europe.</p> <p>As part of the senior leadership <span>team,</span> you will partner closely with Engineering, Security, and Product to strengthen operational excellence, enhance system observability, and drive continuous improvement in how we build and run services. You will lead a distributed team of engineers across the US and UK, fostering a high-performing, collaborative, and growth-oriented environment.</p> <p>This role is ideal for a leader who combines deep technical expertise with a pragmatic approach to improving systems, processes, and team capabilities.</p> </div></div><div class="section page-centered"><div><h3>What You’ll Do</h3><div class="posting-requirements plain-list"><div><ul> <li> <h4><strong>Cloud Platform & Reliability</strong></h4> <ul> <li> <p>Own the availability, performance, and resilience of our multi-region AWS platform.</p> </li> <li> <p>Drive improvements in system reliability through well-defined <strong>SLIs/SLOs</strong>, error budgets, and proactive engineering practices.</p> </li> <li> <p>Lead efforts to reduce <strong>MTTR</strong> and improve incident response effectiveness across the organization.</p> </li> <li> <p>Guide architecture decisions for microservices, Kubernetes (EKS), and serverless workloads to ensure scalability and fault tolerance.</p> </li> </ul> </li> <li> <h4><strong>Observability & Incident Management</strong></h4> <ul> <li> <p>Advance our observability strategy using <strong>Datadog</strong>, ensuring actionable insights across infrastructure and applications.</p> </li> <li> <p>Establish and refine incident management practices, including on-call processes, escalation paths, and post-incident reviews.</p> </li> <li> <p>Act as an incident commander for critical events and contribute to the on-call rotation.</p> </li> </ul> </li> <li> <h4><strong>Operational Excellence & Efficiency</strong></h4> <ul> <li> <p>Elevate operational standards through automation, standardization, and adoption of modern best practices.</p> </li> <li> <p>Drive cost optimization initiatives across AWS environments without compromising performance or reliability.</p> </li> <li> <p>Leverage <strong>AI and automation</strong> to improve operational efficiency, accelerate root cause analysis, and enhance system insights.</p> </li> <li> <p>Continuously improve CI/CD pipelines (CircleCI) and infrastructure-as-code practices (Terraform).</p> </li> </ul> </li> <li> <h4><strong>Team Leadership & Development</strong></h4> <ul> <li> <p>Lead, mentor, and support a distributed team of CloudOps engineers across the US and UK.</p> </li> <li> <p>Foster a culture of accountability, learning, and continuous improvement.</p> </li> <li> <p>Provide technical guidance while enabling the team to grow in ownership and capability.</p> </li> </ul> </li> <li> <h4><strong>Hybrid & Legacy Environment Support</strong></h4> <ul> <li> <p>Ensure stability and support for existing customers while maintaining clear operational boundaries with the cloud platform.</p> </li> </ul> </li> </ul></div></div></div></div><div class="section page-centered"><div><h3>What We’re Looking For</h3><div class="posti

Findigo hittar jobben och fyller i ansökan. Du klickar Skicka.

Visa jobbet och ansök

Ursprunglig annons: weworkremotely.com