Site Reliability Engineer
Indexed description
Excellent opportunity for SRE to be part of our Cloud Infrastructure & Security services practice. Cognizant Infrastructure Services – Provides IT infrastructure & Cloud services for clients across industry verticals, including both Consulting/Professional and Managed Services, across Enterprise Computing, Cloud services, Security Services, DevOps, Data Centres, End User Computing, Service Desk, Network Services and Environment Management Services.
Key Responsibilities
- :
Act as a Site reliability engineer, Reliability as a feature and SRE concep - tsOwn and manage AWS-based IT environments supporting enterprise applications, STAP platform rollouts, and high-performance computing (HPC) workload
- s.End-to-end application rollout on STAP platforms, ensuring scalability, security, compliance, and performance readines
- s.Design, implement, and manage HPC platforms on AWS to support high end genomics, bioinformatics, and data intensive workload
- s.Drive AWS migration initiatives, including assessment, planning, execution, and post migration optimization of on premises and legacy system
- s.Collaborate with application, security, network, and compliance teams to ensure secure, compliant, and resilient cloud architecture
- s.Oversee operational stability, capacity planning, performance optimization, cost management, and incident/problem management across AWS platform
- s.Define governance models, operational standards, and best practices for AWS cloud usage across multiple workloads and platform
- s.Act as a key technical and managerial escalation point for platform issues and critical incident
s.
Key Skills and Experien
- ce:SRE with experience spanning development and operatio
- ns.Practical understanding of SRE concepts, including: Reliability as a feature Error budgets and risk based decision making, Toil reduction through automat
- ionGood experience with AWS cloud platfor
- ms.Proven experience in AWS General IT Management, including governance, operations, security, cost optimization, and stakeholder manageme
- nt.Good experience in application rollout and platform management, preferably on STAP or similar enterprise platfor
- ms.Strong experience in designing and managing HPC platforms on AWS supporting high end genomics, bioinformatics, or other data intensive workloa
- ds.Demonstrated experience in executing AWS-based migration programs, including assessment, planning, execution, and optimization of on premises to cloud migratio
- ns.Experience working in regulated and enterprise environments, with exposure to security, compliance, and data protection requiremen
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search