04 ago - Milano
Coreview
ph3About CoreView /h3 pCoreView is the global leader in Microsoft 365 (M365) tenant resilience, serving over 23 million users worldwide. We empower the world’s leading organizations to master the complexity of Microsoft M365. Through robust security and precise governance, we help ensure that our client’s environments stay cyber-resilient and productive, no matter how complex they are. /p pOur unified, cloud-native platform delivers powerful automation, rapid value, and end-to-end visibility across the entire M365 ecosystem. Backed by world-class support and a collaborative, innovative culture, CoreView is a place where your ideas matter, and your work truly impacts global enterprises. /p h3Job Summary /h3 pTo support our growth, we are looking for an SRE Manager, in either UK or Italy. As our Site Reliability Engineering Manager, you will be responsible for building and leading our Site Reliability Engineering team from the ground up. In this role, you will define the team structure, establish SRE practices and processes, and actively contribute as a hands‑on engineer — especially in the early stages. You will work closely with the Director of Cloud IT Infrastructure and collaborate with software engineering, DevOps, and IT operations teams to ensure the reliability, scalability, and performance of our systems. /p pThe ideal candidate has a strong background in IT operations or infrastructure engineering and has successfully led or coordinated technical teams. We are looking for someone who combines solid operational expertise with a natural inclination toward leadership, process building, and continuous improvement. /p h3Job Responsibilities /h3 ul liBuild the SRE team from scratch: define roles, participate in hiring, onboard and mentor engineers. /li liDefine the SRE technical direction and operational roadmap, aligning reliability initiatives with business objectives and product priorities. /li liEstablish SRE practices, processes, and culture within the organization, including on‑call rotations, incident management, and blameless post‑mortems, fostering a culture of psychological safety and continuous learning. /li liDefine and track SLOs, SLIs, and error budgets in collaboration with engineering and product teams. /li liAct as a hands‑on contributor during the team ramp‑up phase, directly involved in designing and operating critical infrastructure on Azure.
/li liOwn the incident management process end‑to‑end: detection, response, escalation, resolution, and root cause analysis. /li liDrive automation initiatives — including AI‑assisted operations where applicable — to reduce toil and improve the operational efficiency of the team. /liliDesign and oversee monitoring, alerting, and observability solutions to ensure full visibility across systems and services. /li liOwn capacity planning and cloud cost governance, ensuring infrastructure scales efficiently while remaining cost‑effective. /li liBuild and maintain a strong documentation culture: runbooks, operational procedures, incident playbooks, and architectural decisions. /li liCollaborate with software engineering teams to embed reliability and operational readiness into the development lifecycle. /li liDefine and maintain disaster recovery plans, backup strategies, and business continuity procedures. /li liEnsure that security best practices and compliance requirements are applied consistently across infrastructure and operations. /li liReport on team performance, reliability metrics, and operational health to senior leadership. /li /ul h3Job Requirements /h3 ul liA minimum of 3+ years of experience in an SRE role. /li liA minimum of 1+ years of experience leading or coordinating a technical team, with demonstrated ability to hire, mentor, and develop engineers. /li liProven experience in IT operations, infrastructure engineering. /li liSolid hands‑on experience with Azure cloud services and cloud‑based infrastructure management. /li liStrong understanding of IT operations best practices, including incident management, change management, and service continuity. /li liExperience with monitoring and observability tools (e.g., Prometheus, Grafana, Azure Monitor, ELK stack). /li liFamiliarity with Infrastructure as Code (IaC) tools such as Terraform or Ansible. /li liGood knowledge of containerization and orchestration technologies (Docker, Kubernetes / AKS). /li liAbility to define and implement SLOs, SLIs,
and error budgets in a production environment. /li liExperience with capacity planning and cloud cost management. /li liStrong analytical and problem‑solving skills, with the ability to manage complex incidents under pressure. /li liExcellent communication and stakeholder management skills, with the ability to interact effectively at both technical and leadership levels. /li liProactive mindset, ownership attitude, and passion for building reliable, scalable systems. /li liProficient in English. You read and write proficiently and speak at a conversational level in English. /li /ul h3Nice‑to‑have /h3 ul liExperience in a greenfield or scale‑up environment, with a track record of building teams or processes from scratch. /li liBackground in SRE discipline with knowledge of Google SRE principles and practices. /li liExperience with CI/CD pipeline management and DevOps practices. /li liFamiliarity with scripting or automation (Python, Bash, PowerShell). /li liExperience leveraging AI and automation tools to improve operational workflows and reduce manual intervention. /li liKnowledge of chaos engineering practices and fault injection tools (e.g., Azure Chaos Studio). /li liRelevant certifications such as Microsoft Certified: Azure Administrator, Azure DevOps Engineer Expert, or ITIL Foundation. /li /ul h3Coreview Values /h3 pbOwnership Mindset /b: Take ownership. Drive outcomes. /p pbOne Team /b: One team, one goal, embracing diversity - to achieve more together. /p pbVelocity /b: Decide fast. Deliver fast. Repeat. /p pbContinuous Improvement /b: Curiosity drives us. We challenge the status quo. /p pbCustomer First /b: Listen deeply. Solve boldly. /p pbResilience /b: Steady under pressure. /p pCoreView is an organisation which values the strength that diversity brings to the workplace. As an employer, we seek to promote equal opportunity through affirmative action. All qualified applicants will therefore receive consideration for employment and will not be discriminated against based on gender/sex, race/ethnicity, disability, age or any other protected group status (such as protected veteran status) or characteristic that is protected by local legislation. /p pPrivacy Notice: By submitting your application, you acknowledge that CoreView will process your personal data for recruitment purposes in accordance with our Privacy Policy /p /p #J-18808-Ljbffr
06 ago - Torino
ITConsulting
06 ago - Reggio Emilia
K2 Partnering Solutions
06 ago - Torino
ITConsulting
06 ago - Cagliari
K2 Partnering Solutions