Pubblica annuncio gratuito
Ricevi nuove offerte di lavoro

Crea una Job Alert gratuita per senior infrastructure engineer — openstack & cloud systems / palermo

Senior Infrastructure Engineer — OpenStack & Cloud Systems

Pubblicato il 15-09-2026 - Azienda Anonima in Palermo

SENIOR INFRASTRUCTURE ENGINEER - OPENSTACK & CLOUD SYSTEMS

We are looking for a Senior Infrastructure Engineer to serve as a technical leader within our Infrastructure Team, driving the design, evolution, and operation of the software-defined infrastructure that powers some of the largest HPC and AI clusters in Europe. OpenStack is the primary production platform today; the role owns the broader Linux, Kubernetes, and bare-metal systems stack around it.

We hire first for depth of systems engineering — the ability to understand why something does not work, not just how to restart it: reading an strace, reading the source of a service, isolating a kernel or network issue, and driving a fix upstream when needed. A strong engineer with this foundation who knows OpenStack — or an equivalent large-scale production platform — is exactly who we are looking for; the specific stack can be learned, the way of reasoning cannot.

The successful candidate will be responsible for the architecture and operations of cloud control planes in large-scale production environments, ensuring reliability, scalability, and operational continuity. They will lead critical technical initiatives, plan and execute infrastructure upgrades and migrations, collaborate with leading technology vendors to manage escalations, and provide technical leadership to the team through mentoring, knowledge sharing, and best practices.

KEY RESPONSIBILITIES

- Own the architecture and day-2 operations of OpenStack control planes for clusters of hundreds of nodes today, with a growth path to 1000+ nodes at upcoming public and private AI Factories: uptime, capacity, performance, and security KPIs.
- Diagnose and resolve complex, cross-layer production issues down to the root cause — kernel, systemd, storage, and network stack — using tools such as strace, perf, and packet capture, reading service source code where needed and contributing fixes upstream.
- Design and operate the advanced networking underpinning high-throughput HPC and AI workloads (Neutron OVN / OVS, SR-IOV, VF-LAG, DPDK, BGP-EVPN), and drive technology and architecture decisions with senior team members and stakeholders.
- Coordinate deployments and upgrades across geographically distributed sites, including cross-site data replication, federated identity, and disaster-recovery posture.
- Develop and maintain advanced Infrastructure-as-Code pipelines (Ansible, OpenTofu / Terraform, Helm) and enforce gitops-style review for production change management.
- Produce and maintain technical documentation,



including operational runbooks — step-by-step procedures with explicit go / no-go decision gates and rollback plans — for control-plane upgrades, security patching, and migrations, as well as RCA reports.
- Mentor mid-level and junior engineers on Linux and OpenStack internals, lifecycle operations, and production best practices; collaborate with the Presales team in designing systems end-to-end from hardware configuration to software stack.

EDUCATION:

Master's degree or Ph.D. in Computer Science, Telecommunications Engineering, Network Engineering, or a related STEM field — or equivalent practical experience, including:

- 5+ years of hands-on engineering experience in Linux systems, cloud architecture, network engineering, and complex production systems.
- 3+ years of OpenStack production experience at scale (multi-site or strict SLA) — or equivalent experience operating a large-scale production platform, with the depth to become productive on OpenStack within the first months.
- 2+ years of direct vendor escalation accountability in enterprise or hyperscaler environments.

SKILLS AND COMPETENCES - CORE TECHNICAL SKILLS

Linux & systems engineering (deep):

- Expert Linux sysadmin (Rocky / RHEL, Ubuntu, SLES);
- Advanced troubleshooting strace, perf, ftrace / eBPF, gdb; comfortable reading the source of the services operated and contributing fixes upstream;
- Container platforms (Docker, Podman, Singularity) and their runtime internals.

OpenStack & cloud virtualization
- Strong, hands-on proficiency on Neutron, Nova, Ironic, Cinder, Keystone, Manila.
- Direct hands-on experience with Kayobe + Kolla-Ansible is a strong plus.

Advanced networking
- Working knowledge of BGP / OSPF / EVPN / VXLAN, Spine-Leaf datacenter fabrics, SDN, OVN / OVS.
- SR-IOV + VF-LAG, hardware offload, DPDK; Mellanox / low-latency networking.

Infrastructure as Code & Automation
- Production-quality automation in Bash, Python, or Go.
- Hands-on Ansible (modules / roles / collections); Terraform / OpenTofu modules; Helm.

Container orchestration (a plus)
- Production-grade Kubernetes (bare metal and over OpenStack), with focus on GPU-accelerated workloads; Cluster API, ArgoCD.

HPC ecosystem (a plus)
- Familiarity with Slurm, MPI, GPU stacks,



and parallel computing workloads on cloud-managed infrastructure.

Version Control & GitOps
- Solid Git operations (branching, rebasing, conflict resolution); GitLab CI or equivalent — design and review of pipelines for production change management.

Communication
- Fluent English (working language with Mellanox / NVIDIA / Dell engineering); clear written and verbal communication for vendor, customer, and internal audiences.

SKILLS AND COMPETENCES - SOFT SKILLS
- Technical Leadership — Ability to guide technical decisions and mentor engineers.
- Problem Solving — Strong analytical skills to resolve complex, cross-layer infrastructure issues.
- Collaboration — Ability to work effectively with internal teams and external partners.
- Ownership — Strong sense of responsibility and accountability for systems and outcomes.
- Decision Making — Ability to act decisively in high-pressure, production-critical environments.
- Adaptability — Comfortable working in fast-evolving, complex infrastructure environments.
- Knowledge Sharing — Commitment to mentoring and promoting best practices within the team.

WORKING ENVIRONMENT:
- Remote work — Flexibility to work from anywhere in Italy, with no hybrid mandate; on-site presence required for periodic group meetings and a few company events.
- Objective-based work — Clear, measurable objectives reviewed and updated throughout the year, with a structured growth path and access to leading-edge HPC and AI infrastructure in production.
- On-site missions — The role involves periodic on-site engagements at customer and partner facilities in Italy and across the EU (cluster bring-up, vendor PoC, co-deployment) and attendance at national and international events, with travel covered by the company.
- RAL: 40.000,00 - 50.000,00 €

GROWTH & DEVELOPMENT

While this position is targeted at senior engineers, we also welcome applications from mid-level engineers with a solid systems-engineering foundation — strong Linux fundamentals and hands-on experience with OpenStack, Kubernetes, or large-scale distributed systems — looking to grow into full ownership of production cloud infrastructure. What matters most is depth of reasoning and a genuine curiosity for how systems work under the hood; the specific stack is something we help you master.

Depending on experience level, successful candidates will grow into increased responsibility through hands-on work on large-scale AI and HPC infrastructures, supported by structured mentorship and knowledge sharing within the team.

» RISPONDI A QUESTO ANNUNCIO
Altri Annunci
Social & Content Architect
2026-09-16 07:05:04 - Testbusters - Milano
Testbusters è alla ricerca di una figura creativa e pragmatica per guidare la comunicazione social del brand, trasformando obiettivi di marketing e insight della community in contenuti informativi, coinvolgenti ed [...]
Magazziniere Dinamico - Opportunità di crescita
2026-09-16 07:05:01 - GiGroup H-6810 - Pagnacco
Gi Group Spa, filiale di Gemona, cerca un Addetto/a al Magazzino per la gestione quotidiana del magazzino presso Pagnacco. Orario a giornata: e dal lunedì al venerdì. Contratto a tempo determinato in somministrazi [...]
Ricevi nuove offerte di lavoro

Crea una Job Alert gratuita per senior infrastructure engineer — openstack & cloud systems / palermo

Chief of Staff for Private Fundraising Strategy
2026-09-16 07:04:59 - UNICEF - Roma
UNICEF seeks a Strategic Planning Manager to lead the Office of the PFP Director, coordinating strategy, execution and internal collaboration. You will advise the Director and manage cabinet support, ensuring alignm [...]
Sales Analytics Intern: Data & Dashboards
2026-09-16 07:04:57 - Stellantis - Torino
Stellantis is seeking an Sales Support Analyst Intern to join the Sales team in Torino, Italy. The role focuses on analysis of commercial performance, reporting activities, and operational support for the sales team [...]