- Location
- GRC - Thessaloniki, Chortiatis, Greece
- Type
- Full-time
- Department
- Engineering
- Seniority
- Manager
- Experience
- 4+ years
- Education
- Master
- Closing date
- Today
- Source
- Workday
Description
ROLE SUMMARY
At Pfizer we make medicines and vaccines that change patients' lives with a global reach of over 780 million patients. Pfizer Digital is the organization charged with winning the digital race in the pharmaceutical industry. We apply our expertise in technology, innovation, and our business to support Pfizer in this mission.
Our team, Platform & Reliability Engineering within Global Supply Engineering, is passionate about using software, data and AI to improve manufacturing processes. We build and run the platform behind three solution streams:
- Agentic Solutions (Digital Operations Center, Smart Factory, etc.)
- Manufacturing Operations Solutions
- AI & Predictive Operations Solutions
We partner with other Pfizer teams focused on:
- Manufacturing throughput efficiency and increased manufacturing yield
- Reduction of end-to-end cycle time and increase of percent release attainment
- Increased quality control lab throughput and more timely closure of quality assurance investigations
- Increased manufacturing yield of vaccines
- More cost-effective network planning decisions and lowered inventory costs
In the Manager, Platform & Reliability Engineer role, you will work with other Platform and Reliability Engineers, Product Owners, Software Engineers, Data Engineers and other key people to ensure our software is reliable, secure and performant.
You will own the reliability of production services and key parts of the SDLC and software release process. You will help build a self-service platform that lets product teams deliver safely and quickly, in a cost-effective, performant, and safe manner. You will support our solutions by focusing on automation, AI and continuous improvement.
Most of all, you will use your passion for automation, quality, and curiosity to dive into any reliability, performance or cost problem. You will figure out the cause of the problem and propose solutions to resolve it. Your drive for excellence will help us achieve new levels of reliability in our software and our AI solutions.
ROLE RESPONSIBILITIES
The Manager, Platform & Reliability Engineer's responsibilities include, but are not limited to:
- Ensure quality software and infrastructure is delivered on time within a fast-paced agile environment
- Build self-service, reusable CI/CD workflows (GitHub Actions) and infrastructure as code (Terraform) so teams can onboard on their own
- Operate and evolve our Kubernetes (Amazon EKS) platform through GitOps (FluxCD, Kustomize)
- Create and maintain observability and alerting solutions for our services (OpenTelemetry, Prometheus, Grafana, Zabbix)
- Operate AI and LLM services in production, including agent services, LLM gateways and LLM observability
- Support overall platform and reliability strategy, ensuring performance, scalability and reliability of systems
- Define and track service-level objectives (SLOs), error budgets and DORA metrics for critical services
- Lead incident response, on-call and blameless post-incident reviews, and drive fixes to completion
- Ensure database reliability, including Aurora PostgreSQL operations, migrations, backups and recovery
- Embed security and compliance as code: policy enforcement, secrets management, certificate automation, vulnerability remediation
- Maintain and optimize cloud infrastructure performance and cost (FinOps)
- Automate manual and repetitive processes, using AI agents to enhance operational efficiency
- Partner with cross-functional teams to design, develop and enhance new and existing applications
- Build and maintain the release and deployment knowledge base, usable by both engineers and AI agents
- Continuously optimize procedures and processes to reduce deployment time and increase reliability
- Mentor engineers and lead technical workstreams across teams
BASIC QUALIFICATIONS
- Education: Bachelor's degree or Master's degree in Computer Science, Engineering, or related discipline
- Minimum 4 years of experience in Site Reliability Engineering, Platform Engineering, DevOps, Software Development, or similar field
- Experience working with all of the following platforms and tools:
- Kubernetes (Amazon EKS preferred)
- AWS EC2, EKS, EBS, EFS, S3, RDS, MSK or equivalents on other Cloud Providers
- Git and GitHub Actions or similar Cloud based CI tooling
- Terraform or equivalent infrastructure-as-code tooling
- Docker
- Linux
- Experience working with at least 2 of the following tools:
- FluxCD/ArgoCD or Similar GitOps based Kubernetes CD tool
- Helm or Kustomize
- OpenTelemetry
- Prometheus / Grafana
- Envoy / Kubernetes Gateway API
- Kyverno or OPA Gatekeeper
- Proficiency in Bash, plus at least one other modern programming language such as Python, Go, Java, JavaScript, C#
- Operational experience with at least one SQL and one NoSQL database platform such as PostgreSQL, SQL Server, MongoDB, DocumentDB, Neo4j
- Good understanding of microservices architecture and cloud computing
- Experience supporting the development of complex software products, including the deployment of multiple versions / releases of a large product
- Experience with production incident response, on-call, and root-cause analysis
- Experience using AI coding assistants or agents (e.g. GitHub Copilot, Claude Code) in day-to-day engineering work
- Excellent communication skills (verbal & written)
- Proactive approach and goal-oriented mindset, research, problem solving and analytical skills
PREFERRED QUALIFICATIONS
- Pharmaceutical experience, including GxP / regulated environments
- Certified Kubernetes Administrator certification
- AWS Certified Solutions Architect or AWS Certified DevOps Engineer certification
- Experience operating LLM-based applications (e.g. Langfuse, LangChain, pgvector)
- FinOps Certified Practitioner, or hands-on cloud cost-optimization experience
- Experience mentoring engineers or leading a technical workstream
- Experience working with JIRA, Confluence
- Practical knowledge of Agile (Scrum), SAFe (Scaled Agile Framework)
Please apply by sending your CV in English.
Work Location Assignment: Hybrid
Purpose
Breakthroughs that change patients' lives... At Pfizer we are a patient centric company, guided by our four values: courage, joy, equity and excellence. Our breakthrough culture lends itself to our dedication to transforming millions of lives.
Digital Transformation Strategy
One bold way we are achieving our purpose is through our company wide digital transformation strategy. We are leading the way in adopting new data, modelling and automated solutions to further digitize and accelerate drug discovery and development with the aim of enhancing health outcomes and the patient experience.
Flexibility
We aim to create a trusting, flexible workplace culture which encourages employees to achieve work life harmony, attracts talent and enables everyone to be their best working self. Let’s start the conversation!
Equal Employment Opportunity
We believe that a diverse and inclusive workforce is crucial to building a successful business. As an employer, Pfizer is committed to celebrating this, in all its forms – allowing for us to be as diverse as the patients and communities we serve. Together, we continue to build a culture that encourages, supports and empowers our employees.
Disability Inclusion
Our mission is unleashing the power of all our people and we are proud to be a disability inclusive employer, ensuring equal employment opportunities for all candidates. We encourage you to put your best self forward with the knowledge and trust that we will make any reasonable adjustments to support your application and future career. Your journey with Pfizer starts here!
To learn more about acceptable and prohibited uses of AI during the recruitment process, please review our candidate AI-use guidelines available on Pfizer Careers.
Information & Business Tech