American Express Enterprise Cloud team is looking for innovators to help us build world-class applications, Cloud platforms and infrastructure supported by integrated CICD, Observability and security capabilities.
The Director Infrastructure Engineering - Platform Security and Lifecycle Management is responsible for leading the strategy, governance, and execution of Platform as a Service and Middleware services lifecycle management across Private Cloud environments. This role leverages Gen AI/Agentic AI to drive fully automated platform upgrades, security posture management, capacity management, and operational resilience to ensure a secure, scalable, highly available, and compliant Private Cloud platform supporting mission-critical business applications.
The Director partners closely with Platform Engineering, Information Security, Architecture, Infrastructure, SRE, DevOps, and Application Development teams to drive platform modernization, automate operations, reduce technology risk, and enable cloud-native adoption at scale.
At American Express, our culture is built on a 175-year history of innovation, shared values and Leadership Behaviors, and an unwavering commitment to back our customers, communities, and colleagues. From delivering differentiated products to providing world-class customer service, we operate with a strong risk mindset, ensuring we continue to uphold our brand promise of trust, security, and service.
As part of Team Amex, you'll experience our powerful backing with comprehensive support for your holistic well-being and many opportunities to learn new skills, develop as a leader, and grow your career. Here, your voice and ideas matter, your work makes an impact, and together, you will help us define the future of American Express.
Bachelor's degree in Computer Science, Engineering, or related field (Master's preferred).
8+ years of experience in Platform Engineering & Operations, Site Reliability Engineering (SRE), Platform lifecycle management with a proven track record of leading teams in managing large-scale cloud infrastructure with a focus on automation, reliability and resilience.
Deep hands-on experience with any Kubernetes platform(multi-cloud preferred).
Experience building end to end platform upgrade and fleet management automation leveraging Gen AI / Agentic AI
Strong experience with:
Infrastructure as Code (Terraform, CloudFormation, ARM)
Infrastructure automation tools like Ansible
Container platforms (OpenShift/Kubernetes)
Monitoring tools (Prometheus, OTEL, LOKI)
CI/CD pipelines (Jenkins, GitHub Actions)
Open source based messaging, caching, and database technologies like Kafka, Redis, Elastic
Strong understanding of cloud networking, security, and architecture.
Experience managing large-scale, mission-critical production environments.
Relevant certifications preferred
Experience with DevOps practices and methodologies, including CI/CD pipelines, configuration management, and infrastructure as code.
Experience with observability tools such as Prometheus, Splunk, ELK, Dynatrace.
Strong analytical and problem-solving skills, with the ability to troubleshoot complex issues and drive resolution in a fast-paced environment.
Excellent communication and leadership skills, with the ability to effectively collaborate with cross-functional teams and influence decision-making at all levels of the organization.
Employment eligibility to work with American Express in the United States is required as the company will not pursue visa sponsorship for these positions.
Private Cloud Platform Upgrade Strategy & Modernization
Upgrades & Release Management
Security & Compliance
Capacity & Performance Management
Leadership & Team Management
Stakeholder & Vendor Management
Private Cloud Platform Upgrade Strategy & Modernization
Upgrades & Release Management
Security & Compliance
Capacity & Performance Management
Leadership & Team Management
Stakeholder & Vendor Management