06 ago - Milano
Michael Page International
Categoria: Technology & Telecoms Tutti i potenziali candidati sono incoraggiati a leggere l'intera descrizione del lavoro prima di candidarsi.
Luogo di lavoro: Milano
Multinational company specialized in IT Business Consulting is seeking a production support Site Reliability Engineer (SRE) with strong automation skills to join a dynamic team. The ideal candidate will be responsible for ensuring the reliability, availability and performance of the production systems, while driving automation and operational excellence.
- Provide day-to-day operational support for production environments, ensuring high availability and reliability of critical services.
- Develop, maintain and enhance automation scripts and tools using Bash, Python and Ansible to streamline operational tasks and incident response.
- Monitor system performance, proactively identify issues and implement solutions to prevent service disruptions.
- Collaborate with development, QA and infrastructure teams to implement practices for deployment, monitoring and incident management.
- Participate in on-call rotation and respond to production incidents, performing root cause analysis and driving resolution.
- Maintain and improve configuration management, CI/CD pipelines and infrastructure as code practices.
- Document operational processes,
troubleshooting steps and automation workflows.
Requisiti:
- Proven experience in a production support or SRE role within a complex, high-availability environment.
- Strong automation skills with proficiency in Bash, Python and Ansible.
- Experience with monitoring and alerting tools (e.g. Prometheus, Grafana, Elastic stack, Datadog).
- Solid understanding of Linux/Unix systems administration and troubleshooting.
- Familiarity with cloud platforms (e.g. AWS) and containerisation technologies (e.g. Docker, Kubernetes).
- Experience with configuration management and infrastructure as code tools (e.g. Terraform, CloudFormation).
- Knowledge of networking fundamentals, security practices and incident management processes.
- Excellent problem-solving skills, attention to detail and ability to work under pressure.
- Strong communication and collaboration skills.
- Desirable skills:
- Experience with version control systems (e.g. Git).
- Familiarity with Agile methodologies and DevOps culture.
- Exposure to database administration and troubleshooting (e.g. MySQL, PostgreSQL, Oracle). xrdztoy
- Scripting or automation experience with other languages (e.g. Go, Ruby).
08 ago - Italia
Altro
08 ago - Italia
WAICO
08 ago - Italia
Altro
08 ago - Italia
Altro