Description
Role: Senior SRE / Application Support Engineer
Duration: 12-18 Months
Rate: $65-$68/hour W2
Location: Charlotte, NC (Hybrid)
Project Overview:
The team is looking for an experienced SRE / Application Support Engineer to bring an SRE mindset and practices into an established Application Support organization. This person will help improve reliability, observability, incident management, and overall operational maturity across the application environment.
The ideal candidate has 5-10 years of overall technology experience, with 3+ years of recent SRE experience. Prior experience in Application Support or Production Support is highly relevant, particularly for candidates with a strong background in incident management, monitoring, troubleshooting, and reducing MTTR.
Responsibilities
- Bring SRE principles and practices into an Application/Production Support environment.
- Define, establish, and mature Service Level Indicators (SLIs) and Service Level Objectives (SLOs) for applications and services.
- Understand how SLI/SLO maturity evolves from basic service-level measurements toward more comprehensive reliability objectives across applications and business units.
- Help establish and monitor reliability metrics across a broader portfolio or business unit.
- Analyze application performance, availability, latency, errors, and other reliability indicators.
- Participate in incident management, production troubleshooting, root-cause analysis, and post-incident reviews.
- Drive improvements to Mean Time to Resolution (MTTR) and overall operational efficiency.
- Use monitoring and observability tools to identify issues, correlate events, and troubleshoot production incidents.
- Analyze application and infrastructure logs using tools such as Splunk, Dynatrace, or similar observability platforms.
- Identify opportunities to automate repetitive support activities and improve operational processes.
- Partner with application, infrastructure, engineering, and other technical teams to improve application reliability.
- Help establish consistent SRE and production-support practices across the environment.
Required Skills & Experience
- 5-10 years of overall IT/technology experience.
- 3+ years of recent, hands-on SRE experience.
- Prior Application Support / Production Support experience.
- Strong understanding of SRE principles and practices.
- Practical experience defining and using SLIs and SLOs.
- Understanding of SLI/SLO maturity and how reliability objectives can evolve from individual application metrics to broader business-unit-level objectives.
- Strong Incident Management experience.
- Experience improving MTTR and production reliability.
- Strong production troubleshooting and problem-solving skills.
- Experience with monitoring, observability, and log analysis.
- Hands-on experience with Splunk and/or Dynatrace; comparable monitoring/observability tools are also relevant.
- Ability to work across Application Support, Engineering, Infrastructure, and other technical teams.
The strongest candidates will be SREs first and foremost, with recent SRE experience and a prior foundation in Application or Production Support.
Candidates should be able to speak practically about:
- How they define an SLI and SLO.
- How they determine appropriate reliability targets.
- How SLI/SLO programs mature over time.
- How reliability metrics can be standardized or expanded across an entire application portfolio or business unit.
- How they use observability and monitoring to troubleshoot production issues.
- How they approach incident management and reduce MTTR.
- How they apply SRE principles to an existing Application Support organization.
Nice to Have
- Previous experience in the payments industry.
- Experience supporting high-volume or business-critical transaction processing environments.
- Experience establishing SRE practices within an existing production/application support organization.
- Experience with additional observability, monitoring, or logging platforms beyond Splunk and Dynatrace.