In this role you will provide technical leadership for Pfizer's AI/ML infrastructure, shaping cloud platforms and deployment foundations to power enterprise-scale generative AI apps. You will define patterns and standards across AWS/Kubernetes/serverless environments, focusing on reliability, scalability, observability, CI/CD, security, and developer enablement. You will collaborate with software, AI engineering, security, and operations teams to raise platform maturity.
This is a hands-on, architecturally influential position with a strong emphasis on cross-team impact and platform-wide consistency. Define and drive the technical strategy for AI/ML platform infrastructure supporting generative AI applications, LLM integrations, model routing, and enterprise AI services Architect, build, and operate scalable cloud platforms using AWS services such as EKS, ECS Fargate, Lambda, Dynamo DB, S3, Open Search, Secrets Manager, Cloud Watch, ALB, and MWAA Establish reusable infrastructure patterns using Cloud Formation, Helm, and Terraform for multi-environment/multi-region deployments Lead CI/CD architecture with Git Hub Actions, reusable workflows, OIDC-based AWS authentication, automated quality gates, deployment promotion, and environment approvals Design and improve observability across AI platforms (Cloud Watch, logs, Prometheus/Grafana, Open Search, Langfuse, and LLm metrics); build Gen AI workload monitoring Partner with software engineering teams to improve deployment reliability, rollback strategies, health checks, autoscaling, load testing,
and runtime performance Define and enforce security/compliance practices for infrastructure (IAM boundaries, Secrets Manager, secret scanning, audit logging, tagging, change-management) Provide technical leadership for cost optimization, capacity planning, environment standardization, and resilience across environments Mentor engineers, review architecture and infrastructure designs, influence platform engineering practices across teams 7+ years in Dev Ops, platform engineering, cloud infrastructure, SRE, or related roles Strong hands-on experience with AWS/Azure/GCP infrastructure and services (container, serverless, networking, storage, observability, security) Production experience on Kubernetes, ECS/Fargate, or equivalent container orchestration Proficiency with infrastructure-as-code (Cloud Formation, Terraform, Helm) Strong CI/CD experience with Git Hub Actions (workflows, testing, automation) Experience building/operating observability solutions (Cloud Watch, Prometheus, Grafana, Open Search) Solid cloud security knowledge (IAM, secrets management, least privilege, audit logging, compliance) Experience supporting distributed systems, microservices, APIs, and multi-environment deployments Proven ability to lead technical design, mentor engineers, and influence practices across teams Leadership and mentoring Strong communication to diverse audiences AWS services: EKS, ECS Fargate, Lambda, Dynamo DB, S3, Open Search, Secrets Manager, Cloud Watch, ALB, MWAA CI/CD with Git Hub Actions; OIDC-based authentication
📌 Staff Platform Engineer, Ai/Ml Infrastructure (Milano)
🏢 Pfizer
📍 Milano