Prepin
Log in
Leidos

engineering opportunity

Platform Operations Engineer

The Platform Operations Engineer will ensure the availability, reliability, and performance of a containerized microservices platform supporting OSINT tasking and collection. Responsibilities include monitoring system health, leading incident response, and collaborating with cross-functional teams to maintain large-scale production environments.

Bethesda, Maryland, United StateshybridFULL_TIME

Posted

About the role

What will you do at Leidos?

Leidos is excited to present an opportunity for a TS/SCI‑cleared Platform Operations Engineer to join a high‑impact team driving the design, development, and deployment of a modern technology stack supporting the DOMEX Technology Platform (DTP). This role directly supports our customer’s mission to centralize and standardize the Tasking, Collection, Processing, Exploitation, and Dissemination (TCPED) of Open Source Intelligence (OSINT) across the Defense Intelligence Enterprise.

You’ll be part of a mission‑focused, solutions‑oriented team that values inclusion, innovation, collaboration, and continuous professional growth. While the majority of work is performed on‑site at our customer location in Bethesda, MD, we offer a flexible schedule, and some tasks may be completed remotely.

As a Platform Operations Engineer you will work with a team to ensure the availability, reliability, and performance of a full stack, containerized microservices platform. You also will partner with a multidisciplinary team of systems engineers, developers, integrators, and system administrators in the following areas:

System Reliability & Performance — Ensuring uptime, performance, and capacity planning for a large scale big data production platform with a microservice architecture running on Kubernetes, Elasticsearch, PostgreSQL, Kafka, and technologies such as Java, Python, React, and low code tools like Appian

Monitoring & Observability — Leveraging monitoring tools to proactively detect and resolve issues

Incident Response — Leading triage, troubleshooting, root cause analysis, and post incident reviews

SLIs & SLOs — Defining and tracking reliability metrics

SAFe Agile — Participating in release planning, scrums, design sessions, bug triage, and cross team coordination

You bring enthusiasm, the ability to work well with people from different disciplines with varying degrees of technical experience, and meet the following qualifications:

BS in Engineering, Computer Science, Systems Engineering, or related field (or equivalent experience) with 8+ years of relevant experience; 6+ years with a Master’s; additional experience may substitute for a degree

Active TS/SCI clearance with the ability to obtain and maintain a polygraph

At least one DoD 8570.01 M IAT Level II+ certification (e.g., Security+ CE, CySA+, CCNA Security, SSCP, CISSP (or Associate))

Ability to obtain Privileged User Account (PUA) certification

Experience with Kubernetes, GitLab pipelines, Linux, and containerized environments

Experience supporting enterprise scale production systems

Experience with cloud services (preferably AWS) and cloud infrastructure

Familiarity with Elasticsearch, PostgreSQL, Logstash, Kibana, and Keycloak

Demonstrated success in cross functional coordination and execution

Strong communication skills and the ability to perform under pressure during incidents

You will stand out even more if you bring:

Experience with Agile methodologies

Experience with creating customized dashboards to track SLIs and other key performance indicators

Development experience (Bash, PowerShell, SALT, Python, Groovy, Java, etc.)

Experience with Appian or other low‑code platforms

Experience with technologies such as Kafka, AMQP/JMS, Prometheus/Grafana, GPU‑based Kubernetes, SALT automation, Nexus, or GraphQL

Knowledge of security best practices (authN/Z, secrets management, data protection)

Infrastructure‑as‑code experience (CloudFormation, Terraform, Pulumi)

AWS cloud certifications

#NMECDTP-ALL

#ASBA

If you're looking for comfort, keep scrolling. At Leidos, we outthink, outbuild, and outpace the status quo — because the mission demands it. We're not hiring followers.

We're recruiting the ones who disrupt, provoke, and refuse to fail. Step 10 is ancient history. We're already at step 30 — and moving faster than anyone else dares.

Original Posting:

July 29, 2026

For U.S. Positions: While subject to change based on business needs, Leidos reasonably anticipates that this job requisition will remain open for at least 3 days with an anticipated close date of no earlier than 3 days after the original posting date as listed above.

Pay Range:

Pay Range $107,900.00 - $195,050.00

The Leidos pay range for this job level is a general guideline only and not a guarantee of compensation or salary. Additional factors considered in extending an offer include (but are not limited to) responsibilities of the job, education, experience, knowledge, skills, and abilities, as well as internal equity, alignment with market data, applicable bargaining agreement (if any), or other law.

Which skills does this role require?

KubernetesElasticsearchPostgreSQLKafkaJavaPythonReactAppianGitLabLinuxAWSLogstashKibanaKeycloakSystem ReliabilityIncident ResponsePlatform OperationsTS/SCI ClearanceDoD 8570MicroservicesAgileSAFeCloud InfrastructureObservabilityBig DataTerraformPrometheusGrafanaGraphQLInfrastructure-as-code

Make your next move

Build a shortlist and prepare

Identify the requirements you can demonstrate, then choose examples from your work to discuss with the hiring team.

Review the responsibilities and requirements before adding an opening to your shortlist.

Role information can change. Confirm current details on the original application page.

Product

AI Candidate AgentCompaniesBrowse JobsDeep ProfileSkill AssessmentOpportunity Matching
Prepin.ai

© 2026 Prepin | All rights reserved.