Staff Platform Engineer, Ai/Ml Infrastructure

14 set - Catania
Pfizer

OverviewIn this role you will provide technical leadership for Pfizer's AI/ML infrastructure, shaping cloud platforms and deployment foundations to power enterprise-scale generative AI apps.
You will define patterns and standards across AWS/Kubernetes/serverless environments, focusing on reliability, scalability, observability, CI/CD, security, and developer enablement.
You will collaborate with software, AI engineering, security, and operations teams to raise platform maturity.
This is a hands-on, architecturally influential position with a strong emphasis on cross-team impact and platform-wide consistency.
ResponsabilitàDefine and drive the technical strategy for AI/ML platform infrastructure supporting generative AI applications, LLM integrations, model routing, and enterprise AI services
Architect, build, and operate scalable cloud platforms using AWS services such as EKS, ECS Fargate, Lambda, DynamoDB, S3, OpenSearch, Secrets Manager, CloudWatch, ALB, and MWAA
Establish reusable infrastructure patterns using CloudFormation, Helm, and Terraform for multi-environment/multi-region deployments
Lead CI/CD architecture with GitHub Actions, reusable workflows, OIDC-based AWS authentication, automated quality gates, deployment promotion, and environment approvals
Design and improve observability across AI platforms (CloudWatch, logs, Prometheus/Grafana, OpenSearch, Langfuse, and LLm metrics); build GenAI workload monitoring
Partner with software engineering teams to improve deployment reliability, rollback strategies, health checks, autoscaling, load testing, and runtime performance




Define and enforce security/compliance practices for infrastructure (IAM boundaries, Secrets Manager, secret scanning, audit logging, tagging, change-management)
Provide technical leadership for cost optimization, capacity planning, environment standardization, and resilience across environments
Mentor engineers, review architecture and infrastructure designs, influence platform engineering practices across teams
Requisiti fondamentali7+ years in DevOps, platform engineering, cloud infrastructure, SRE, or related roles
Strong hands-on experience with AWS/Azure/GCP infrastructure and services (container, serverless, networking, storage, observability, security)
Production experience on Kubernetes, ECS/Fargate, or equivalent container orchestration
Proficiency with infrastructure-as-code (CloudFormation, Terraform, Helm)
Strong CI/CD experience with GitHub Actions (workflows, testing, automation)
Experience building/operating observability solutions (CloudWatch, Prometheus, Grafana, OpenSearch)
Solid cloud security knowledge (IAM, secrets management, least privilege, audit logging, compliance)
Experience supporting distributed systems, microservices, APIs, and multi-environment deployments
Proven ability to lead technical design, mentor engineers, and influence practices across teams
Leadership and mentoring
Strong communication to diverse audiences
Collaborative, cross-functional mindset
AWS services: EKS, ECS Fargate, Lambda, DynamoDB, S3, OpenSearch, Secrets Manager, CloudWatch, ALB, MWAA
Infrastructure-as-code: CloudFormation, Terraform, Helm
CI/CD with GitHub Actions; OIDC-based authentication

Meccanico - 0149

16 set - Lissone
Adecco Italia

Pizzaiolo esperto 2.500

16 set - Somma Lombardo
SAMARCANDA

Ricevi nuove offerte di lavoro

Crea una Job Alert gratuita per staff platform engineer, ai/ml infrastructure / catania

Operatore telefonico in smart working CATANZARO

16 set - Catanzaro
Gierre Contact Call Center

Addetto produzione

16 set - Avigliana
Gelati PEPINO 1884