Site Reliability Engineer

13 ago - Ro
Altro

Who We Are We are looking for a hands‑on

Site Reliability Engineer (SRE)

to help improve, scale, and operationalize an internal platform that enables engineering teams to ship faster and safer through reliable, automated, and resilient engineering workflows.

You will work on platform reliability, observability, automation, incident management, CI/CD workflows, GitHub‑based engineering automation, and developer tooling that helps teams consistently adopt operational best practices across repositories and services.

What You’ll Be Doing

Design, implement and maintain monitoring, alerting, and observability solutions to ensure platform reliability and performance

Develop and improve automation for infrastructure provisioning, deployment pipelines, and operational processes

Administer and optimize GitHub Enterprise environments, including repository management, access controls, branch protection policies, and enterprise‑wide standards

Partner with engineering teams to define and measure Service Level Objectives (SLOs), Service Level Indicators (SLIs), and reliability metrics

Investigate production incidents, perform root cause analysis,



and drive post‑incident improvements to prevent recurrence

Improve system resilience, scalability, and availability through proactive reliability engineering practices

Build and maintain GitHub‑based automation, CI/CD pipelines, and developer self‑service capabilities

What You’ll Bring Along

BSc/MSc in Computer Science or related field

Minimum 6+ years as a SRE

Strong experience with GitHub Enterprise (repos, orgs, actions, integrations)

Advanced hands‑on knowledge of Terraform (IaC)

Experience building and running CI/CD pipelines (GitHub Actions ideally)

Solid understanding of cloud platforms (Azure/AWS/GCP) and integrations

Experience with identity and access management (RBAC, token/app auth models)

Knowledge of backup, recovery, and cyber resilience principles (Cybervault)

Experience with security best practices and software supply chain risks (due to recent events)

Focus on improving developer experience and self‑service capabilities

Excellent problem‑solving and root cause analysis skills

#J-18808-Ljbffr

CAMERIERA/E AI PIANI

13 ago - Forte dei Marmi
E-Work S. P. A.

Addetto Controllo Accessi – Accoglienza e Coordinazione

13 ago - Trento
Blue Zone

Ricevi nuove offerte di lavoro

Crea una Job Alert gratuita per site reliability engineer / ro

Addetto/a mensa

13 ago - Antey-Saint-André
Jobtech

Ottico Esperto: Crescita in Team

13 ago - Bassano del Grappa
VISION GROUP