As our System Operations Lead, you will be the strategic and operational linchpin for our entire infrastructure ecosystem. You will shape the future of our software products in the cloud, ensuring reliability, scalability, and seamless deployment by guiding and steering three of our core areas: the Platform Team, the upcoming SRE Team, and the Installation Team.
Leadership & Strategy: Provide disciplinary and technical leadership, synchronizing three key teams: the current Platform Team, the newly forming SRE (Site Reliability Engineering) Team, and the Installation Team.
SaaS Transformation & Reliability: Actively drive the SaaS transformation of our business unit while establishing modern SRE practices to guarantee high availability and peak performance.
Team Development: Oversee onboarding, mentoring, and continuous professional growth for team members across different operational disciplines.
Cross-Functional Coordination: Act as the primary interface for cross-team communication and alignment with management and software engineering teams.
Hands-on Operations: Collaborating on the teams' technical and operational tasks, including infrastructure architecture, deployment pipelines, and automation.
Quality Assurance: Conduct code reviews (especially for Infrastructure as Code) and review system deployment procedures.