Leidos logo

Platform Operations Engineer

Leidos

  • Bethesda, Maryland
  • 8 days ago

    Highlights

    You also will partner with a multidisciplinary team of systems engineers, developers, integrators, and system administrators in the following areas: System Reliability & Performance — Ensuring uptime, performance, and capacity planning for a large scale big data production platform with a microservice architecture running on Kubernetes, Elasticsearch, PostgreSQL, Kafka, and technologies such as Java, Python, React, and low code tools like Appian. Leidos is excited to present an opportunity for a TS/SCI‑cleared Platform Operations Engineer to join a high‑impact team driving the design, development, and deployment of a modern technology stack supporting the DOMEX Technology Platform (DTP).
    Leidos

    Numbers & Facts

    LocationBethesda, Maryland
    IndustryEngineering Services
    Company Size10,000 employees or more
    Year Founded1969
    Websitehttp://leidos.com/

    Description

    Leidos is excited to present an opportunity for a TS/SCI‑cleared Platform Operations Engineer to join a high‑impact team driving the design, development, and deployment of a modern technology stack supporting the DOMEX Technology Platform (DTP). This role directly supports our customer’s mission to centralize and standardize the Tasking, Collection, Processing, Exploitation, and Dissemination (TCPED) of Open Source Intelligence (OSINT) across the Defense Intelligence Enterprise.

    You’ll be part of a mission‑focused, solutions‑oriented team that values inclusion, innovation, collaboration, and continuous professional growth. While the majority of work is performed on‑site at our customer location in Bethesda, MD, we offer a flexible schedule, and some tasks may be completed remotely.

    As a Platform Operations Engineer you will work with a team to ensure the availability, reliability, and performance of a full stack, containerized microservices platform. You also will partner with a multidisciplinary team of systems engineers, developers, integrators, and system administrators in the following areas:

    • System Reliability & Performance — Ensuring uptime, performance, and capacity planning for a large scale big data production platform with a microservice architecture running on Kubernetes, Elasticsearch, PostgreSQL, Kafka, and technologies such as Java, Python, React, and low code tools like Appian

    • Monitoring & Observability — Leveraging monitoring tools to proactively detect and resolve issues

    • Incident Response — Leading triage, troubleshooting, root cause analysis, and post incident reviews

    • SLIs & SLOs — Defining and tracking reliability metrics

    • SAFe Agile — Participating in release planning, scrums, design sessions, bug triage, and cross team coordination

    You bring enthusiasm, the ability to work well with people from different disciplines with varying degrees of technical experience, and meet the following qualifications:

                          

    • BS in Engineering, Computer Science, Systems Engineering, or related field (or equivalent experience) with 8+ years of relevant experience; 6+ years with a Master’s; additional experience may substitute for a degree

    • Active TS/SCI clearance with the ability to obtain and maintain a polygraph

    • At least one DoD 8570.01 M IAT Level II+ certification (e.g., Security+ CE, CySA+, CCNA Security, SSCP, CISSP (or Associate))

    • Ability to obtain Privileged User Account (PUA) certification

    • Experience with Kubernetes, GitLab pipelines, Linux, and containerized environments

    • Experience supporting enterprise scale production systems

    • Experience with cloud services (preferably AWS) and cloud infrastructure

    • Familiarity with Elasticsearch, PostgreSQL, Logstash, Kibana, and Keycloak

    • Demonstrated success in cross functional coordination and execution

    • Strong communication skills and the ability to perform under pressure during incidents

    You will stand out even more if you bring:

    • Experience with Agile methodologies

    • Experience with creating customized dashboards to track SLIs and other key performance indicators

    • Development experience (Bash, PowerShell, SALT, Python, Groovy, Java, etc.)

    • Experience with Appian or other low‑code platforms

    • Experience with technologies such as Kafka, AMQP/JMS, Prometheus/Grafana, GPU‑based Kubernetes, SALT automation, Nexus, or GraphQL

    • Knowledge of security best practices (authN/Z, secrets management, data protection)

    • Infrastructure‑as‑code experience (CloudFormation, Terraform, Pulumi)

    • AWS cloud certifications

    #NMECDTP-ALL

    #ASBA

    If you're looking for comfort, keep scrolling. At Leidos, we outthink, outbuild, and outpace the status quo — because the mission demands it. We're not hiring followers. We're recruiting the ones who disrupt, provoke, and refuse to fail. Step 10 is ancient history. We're already at step 30 — and moving faster than anyone else dares.

    Original Posting:

    July 29, 2026

    For U.S. Positions: While subject to change based on business needs, Leidos reasonably anticipates that this job requisition will remain open for at least 3 days with an anticipated close date of no earlier than 3 days after the original posting date as listed above.

    Pay Range:

    Pay Range $107,900.00 - $195,050.00

    The Leidos pay range for this job level is a general guideline only and not a guarantee of compensation or salary. Additional factors considered in extending an offer include (but are not limited to) responsibilities of the job, education, experience, knowledge, skills, and abilities, as well as internal equity, alignment with market data, applicable bargaining agreement (if any), or other law.

    About Company

    Everything we do is built on our commitment to do the right thing for our customers, our employees, and our communities. Learn more about the values and culture that are the foundations of our business.

    Similar Jobs

    See more jobs