Associate Director Cloud Operations
Indexed description
Do you want to work on innovative projects, collaborate with a dynamic and supportive team, and receive investment in your professional development? At DTCC, we are at the forefront of innovation in the financial markets. We're committed to helping our employees grow and succeed. We believe that you have the skills and drive to make a real impact. We foster a thriving internal community and are committed to creating a workplace that looks like the world that we serve.
The Information Technology group delivers secure, reliable technology solutions that enable DTCC to be the trusted infrastructure of the global capital markets. The team delivers high-quality information through activities that include development of essential, building infrastructure capabilities to meet client needs and implementing data standards and governance.
Pay And Benefits
- Competitive compensation, including base pay and annual incentive
- Comprehensive health and life insurance and well-being benefits
- Pension
- Paid Time Off and Personal/Family Care, and other leaves of absence when needed to support your physical, financial, and emotional well-being.
- DTCC offers a flexible/hybrid model of 3 days onsite and 2 days remote (onsite Tuesdays, Wednesdays and a third day unique to each team or employee).
Your Primary Responsibilities
- Own incident response on your shift in a 24x7 hybrid cloud environment that supports critical financial market operations.
- Drive complex, cross-domain incidents to resolution across AWS, Azure, on-prem, and hybrid infrastructure, making mitigation decisions under pressure with incomplete information.
- Manage engineers on your shift rotation, developing them into independent decision-makers who can run incidents.
- Lead your team's build work on operational tooling and automation between incidents, with a focus on eliminating recurring manual work.
- Prioritize work based on incident patterns, translating recurring failures into automation that prevents recurrence.
- Enforce production change governance on your shift, including reconciliation of manual changes back into infrastructure-as-code within SLA.
- Maintain operational visibility through daily change posture awareness and pattern detection.
- Partner across Platform Engineering, Application Teams, and Delivery Engineering to ensure operational excellence drives improvements across the organization.
- Minimum of 8 years of related experience
- Bachelor's degree preferred or equivalent experience
- 3+ years managing people in an operational environment
- Deep experience in cloud operations, SRE, or infrastructure engineering
- Infrastructure knowledge spanning AWS, Azure, VMware, hybrid networking, storage, Kubernetes, Linux, and Windows
- Proficiency in Python, Terraform, infrastructure-as-code, and CI/CD pipelines
- Working knowledge of AI-assisted operations tooling such as Amazon Kiro, AWS AgentCore, or agentic automation frameworks
- Observability platform experience such as Grafana, Prometheus, Datadog, CloudWatch, or Dynatrace
- Incident management expertise across the full ITSM lifecycle, including root cause analysis that drives architectural change
- People leadership in an operational environment, including shift rotation design, on-call models, and performance development
- Experience operating in a regulated or financial services environment preferred
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search