Overview
In this role you will provide technical leadership for Pfizer’s AI/ML infrastructure, shaping cloud platforms and deployment foundations to power enterprise-scale generative AI apps. You will define patterns and standards across AWS/Kubernetes/serverless environments, focusing on reliability, scalability, observability, CI/CD, security, and developer enablement. You will collaborate with software, AI engineering, security, and operations teams to raise platform maturity. This is a hands-on, architecturally influential position with a strong emphasis on cross-team impact and platform-wide consistency.
Responsabilità
- Define and drive the technical strategy for AI/ML platform infrastructure supporting generative AI applications, LLM integrations, model routing, and enterprise AI services
- Architect, build, and operate scalable cloud platforms using AWS services such as EKS, ECS Fargate, Lambda, DynamoDB, S3, OpenSearch, Secrets Manager, CloudWatch, ALB, and MWAA
- Establish reusable infrastructure patterns using CloudFormation, Helm, and Terraform for multi-environment/multi-region deployments
- Lead CI/CD architecture with GitHub Actions, reusable workflows, OIDC-based AWS authentication, automated quality gates, deployment promotion, and environment approvals
- Design and improve observability across AI platforms (CloudWatch, logs, Prometheus/Grafana, OpenSearch, Langfuse, and LLm metrics); build GenAI workload monitoring
- Partner with software engineering teams to improve deployment reliability, rollback strategies, health checks, autoscaling, load testing, and runtime performance
- Define and enforce security/compliance practices for infrastructure (IAM boundaries, Secrets Manager, secret scanning, audit logging, tagging, change-management)
- Provide technical leadership for cost optimization, capacity planning, environment standardization, and resilience across environments
- Mentor engineers, review architecture and infrastructure designs, influence platform engineering practices across teams
Requisiti fondamentali
- 7+ years in DevOps, platform engineering, cloud infrastructure, SRE, or related roles
- Strong hands-on experience with AWS/Azure/GCP infrastructure and services (container, serverless, networking, storage, observability, security)
- Production experience on Kubernetes, ECS/Fargate, or equivalent container orchestration
- Proficiency with infrastructure-as-code (CloudFormation, Terraform, Helm)
- Strong CI/CD experience with GitHub Actions (workflows, testing, automation)
- Experience building/operating observability solutions (CloudWatch, Prometheus, Grafana, OpenSearch)
- Solid cloud security knowledge (IAM, secrets management, least privilege, audit logging, compliance)
- Experience supporting distributed systems, microservices, APIs, and multi-environment deployments
- Proven ability to lead technical design, mentor engineers, and influence practices across teams
- Leadership and mentoring
- Strong communication to diverse audiences
- Collaborative, cross-functional mindset
- AWS services: EKS, ECS Fargate, Lambda, DynamoDB, S3, OpenSearch, Secrets Manager, CloudWatch, ALB, MWAA
- Infrastructure-as-code: CloudFormation, Terraform, Helm
- CI/CD with GitHub Actions; OIDC-based authentication
Staff Platform Engineer, AI/ML Infrastructure datore di lavoro: Pfizer
Pfizer è un datore di lavoro eccezionale, che offre un ambiente di lavoro flessibile e inclusivo, dove i dipendenti possono crescere professionalmente e contribuire a trasformare la vita dei pazienti. Con opportunità di sviluppo personale e un forte impegno per la diversità, i collaboratori possono sentirsi valorizzati e supportati nel loro percorso professionale. La sede di Ascoli Piceno offre anche vantaggi come copertura sanitaria, piani di risparmio pensionistico e programmi di assistenza ai dipendenti, rendendo l'esperienza lavorativa ancora più gratificante.