ROLE SUMMARY
Pfizers purpose breakthroughs that change patients lives is rooted in being science driven and a patient focused company. Digital and Technology is the driving force of data and AI innovation at Pfizer.
As part of the Data and AI Platforms organization you will join a team of Platform Engineers responsible for building and operating the enterprise procode AI Agent Platforms. We are expanding our forward-thinking Platform Engineering team focused on delivering secure scalable and resilient cloud-native infrastructure.
In the role of Site Reliability/Operations Engineer you will play a critical part in ensuring the reliability performance and operational excellence of our platforms. This position is well suited to a high-caliber well-rounded engineer who thrives in dynamic environments takes initiative and enjoys solving complex problems across infrastructure automation and observability. You will be part of a collaborative and inclusive team that values curiosity continuous learning and shared success with clear goals strong mentorship and a culture designed to help you succeed while taking ownership of meaningful challenges.
You will contribute to the design and development of a reliable transparent interoperable and sustainable enterprise-grade observability platform with a strong focus on security governance cost management and lifecycle this role you will help shape the platform roadmap collaborate closely with architecture security and product teams and drive the delivery of a best-in-class developer experience. You will also support and mentor engineers promoting engineering excellence and best practices across the team.
Success in this role will require building trusted partnerships with technology vendors delivery partners and internal stakeholders across the organization. You will apply both emerging and established technologies to enhance analytics and observability capabilities supporting the broader agentic platform strategy and enabling scalable high-performing and well-governed systems.
You will start your day by reviewing system health and platform metrics to ensure everything is operating as expected. From there you will collaborate with your platform engineer and operations teammates to prioritize work whether that involves deploying infrastructure refining automation or addressing new technical challenges. Some days will see you working deeply within Terraform modules or tuning infrastructure while others will focus on troubleshooting unexpected issues or supporting colleagues with complex problems. The role provides a balance of focused individual work and collaborative problem-solving all within a fast-paced yet supportive environment where your contributions directly influence platform reliability scalability and security.
In this role you will be responsible for establishing and continuously improving observability and operational practices across the platform for both internal platform engineers and external end-users of the platform services. This includes building robust monitoring logging tracing alerting and incident response capabilities alongside defining and maintaining SLOs and SLAs. You will proactively optimize performance availability and cost efficiency through techniques such as autoscaling right-sizing and resource planning.
Key responsibilities include:
BASIC QUALIFICATIONS
PREFERRED QUALIFICATIONS
PHYSICAL/MENTAL REQUIREMENTS
NON-STANDARD WORK SCHEDULE TRAVEL OR ENVIRONMENT REQUIREMENTS
Pfizer is an equal opportunity employer and complies with all applicable equal employment opportunity legislation in each jurisdiction in which it operates.
To learn more about acceptable and prohibited uses of AI during the recruitment process please review our candidate AI-use guidelines available onPfizer Careers.
Information & Business TechRequired Experience:
Manager
Manager, Platform Operations & Site Reliability Engineer • Chennai, Tamil Nadu, India