Production Support Engineer

Diversity Nexus

  • Pittsburgh, PA
  • 9 days ago

    Highlights

    Oversee incident management for EOP sensitive client incidents, aged incidents, and critical production stability issues. Coordinate and represent in CAB, Pre-CAB, LCAB, GCAB, change freeze, and weekly change governance meetings.

    Numbers & Facts

    LocationPittsburgh, PA

    Description

    Looking for only Locals
    Location: Pittsburgh, PA

    CAB and change governance
    Release and deployment coordination
    Disaster recovery and resiliency planning
    Incident and production support managementL1/L2 and on-call coordination
    ServiceNow, JIRA, Confluence, SharePoint and xMatters
    Leadership reporting and stakeholder communication
    ITIL process knowledge

    Nice to have:
    Banking or financial services experience
    SRE, observability, SLIs/SLOs and error reduction
    Splunk, Dynatrace, Grafana or AppDynamics
    CI/CD and DevOps exposure
    Business continuity and regulatory resiliency knowledge
    ITIL, SRE or DR certification
    Coordinate and represent in CAB, Pre-CAB, LCAB, GCAB, change freeze, and weekly change governance meetings. Prepare and share weekly change reports, significant event updates, and change summaries for leadership stakeholders. Manage release governance activities, including release tracker updates, approval coordination, deployment support, and production readiness follow-up. Serve as Technology Resiliency Coordinator for applications, and related production services. Coordinate DCR, DR, SF DR/FOS, and other resiliency activities to ensure planning, execution, tracking, and closure. Coordinate between L1 and L2 production support coverage, US team rotations, Sunday SoD delegation, weekend support, and off-hours escalation coverage. Oversee incident management for EOP sensitive client incidents, aged incidents, and critical production stability issues. Provide operational guidance to the ECC team and support readiness for expanded critical monitoring responsibilities. Maintain, monitor and govern operational tools and repositories, including xMatters, Confluence, ServiceNow Knowledge, JIRA SDG, SharePoint, SNO dashboards Scale towards SRE adoption by practicing reliability engineering practices such as service health indicators, operational readiness standards, observability, error reduction, and proactive resilience engineering. ServiceNow, JIRA, Confluence, SharePoint
    SRE and observability practicesL1/L2 and on-call support coordination
    CI/CD and DevOps
    Banking and financial services domain knowledge
    Business continuity and regulatory compliance

    Work Experience 7-10Years

    Similar Jobs

    See more jobs