Job Requirements
Bethesda, MD
Top Secret/SCI Polygraph
Career Level not specified
Salary not specified
Join Premium to unlock estimated salaries
Job Description
Platform Operations Engineer - TS/SCI
Bethesda, Maryland
Location: Bethesda, MD
Salary: $107,900.00 - $170,00
Category: Software Engineer
Travel Required: No
Remote Type: Hybrid
Clearance: TS/SCI
Client is looking for a Platform Operations Engineer to join a high‑impact team driving the design, development, and deployment of a modern technology stack supporting the DOMEX Technology Platform (DTP). This role directly supports our customer's mission to centralize and standardize the Tasking, Collection, Processing, Exploitation, and Dissemination (TCPED) of Open Source Intelligence (OSINT) across the Defense Intelligence Enterprise.
You'll be part of a mission‑focused, solutions‑oriented team that values inclusion, innovation, collaboration, and continuous professional growth. While the majority of work is performed on‑site at our customer location in Bethesda, MD, we offer a flexible schedule, and some tasks may be completed remotely.
As a Platform Operations Engineer you will work with a team to ensure the availability, reliability, and performance of a full stack, containerized microservices platform. You also will partner with a multidisciplinary team of systems engineers, developers, integrators, and system administrators in the following areas:
System Reliability & Performance - Ensuring uptime, performance, and capacity planning for a large scale big data production platform with a microservice architecture running on Kubernetes, Elasticsearch, PostgreSQL, Kafka, and technologies such as Java, Python, React, and low code tools like Appian
Monitoring & Observability - Leveraging monitoring tools to proactively detect and resolve issues
Incident Response - Leading triage, troubleshooting, root cause analysis, and post incident reviews
SLIs & SLOs - Defining and tracking reliability metrics
SAFe Agile - Participating in release planning, scrums, design sessions, bug triage, and cross team coordination
You bring enthusiasm, the ability to work well with people from different disciplines with varying degrees of technical experience, and meet the following qualifications:
BS in Engineering, Computer Science, Systems Engineering, or related field (or equivalent experience) with 8+ years of relevant experience; 6+ years with a Master's; additional experience may substitute for a degree
Active TS/SCI clearance with the ability to obtain and maintain a polygraph
At least one DoD 8570.01 M IAT Level II+ certification (e.g., Security+ CE, CySA+, CCNA Security, SSCP, CISSP (or Associate))
Ability to obtain Privileged User Account (PUA) certification
Experience with Kubernetes, GitLab pipelines, Linux, and containerized environments
Experience supporting enterprise scale production systems
Experience with cloud services (preferably AWS) and cloud infrastructure
Familiarity with Elasticsearch, PostgreSQL, Logstash, Kibana, and Keycloak
Demonstrated success in cross functional coordination and execution
Strong communication skills and the ability to perform under pressure during incidents
You will stand out even more if you bring:
Experience with Agile methodologies
Experience with creating customized dashboards to track SLIs and other key performance indicators
Development experience (Bash, PowerShell, SALT, Python, Groovy, Java, etc.)
Experience with Appian or other low‑code platforms
Experience with technologies such as Kafka, AMQP/JMS, Prometheus/Grafana, GPU‑based Kubernetes, SALT automation, Nexus, or GraphQL
Knowledge of security best practices (authN/Z, secrets management, data protection)
Infrastructure‑as‑code experience (CloudFormation, Terraform, Pulumi)
AWS cloud certifications
Bethesda, Maryland
Location: Bethesda, MD
Salary: $107,900.00 - $170,00
Category: Software Engineer
Travel Required: No
Remote Type: Hybrid
Clearance: TS/SCI
Client is looking for a Platform Operations Engineer to join a high‑impact team driving the design, development, and deployment of a modern technology stack supporting the DOMEX Technology Platform (DTP). This role directly supports our customer's mission to centralize and standardize the Tasking, Collection, Processing, Exploitation, and Dissemination (TCPED) of Open Source Intelligence (OSINT) across the Defense Intelligence Enterprise.
You'll be part of a mission‑focused, solutions‑oriented team that values inclusion, innovation, collaboration, and continuous professional growth. While the majority of work is performed on‑site at our customer location in Bethesda, MD, we offer a flexible schedule, and some tasks may be completed remotely.
As a Platform Operations Engineer you will work with a team to ensure the availability, reliability, and performance of a full stack, containerized microservices platform. You also will partner with a multidisciplinary team of systems engineers, developers, integrators, and system administrators in the following areas:
System Reliability & Performance - Ensuring uptime, performance, and capacity planning for a large scale big data production platform with a microservice architecture running on Kubernetes, Elasticsearch, PostgreSQL, Kafka, and technologies such as Java, Python, React, and low code tools like Appian
Monitoring & Observability - Leveraging monitoring tools to proactively detect and resolve issues
Incident Response - Leading triage, troubleshooting, root cause analysis, and post incident reviews
SLIs & SLOs - Defining and tracking reliability metrics
SAFe Agile - Participating in release planning, scrums, design sessions, bug triage, and cross team coordination
You bring enthusiasm, the ability to work well with people from different disciplines with varying degrees of technical experience, and meet the following qualifications:
BS in Engineering, Computer Science, Systems Engineering, or related field (or equivalent experience) with 8+ years of relevant experience; 6+ years with a Master's; additional experience may substitute for a degree
Active TS/SCI clearance with the ability to obtain and maintain a polygraph
At least one DoD 8570.01 M IAT Level II+ certification (e.g., Security+ CE, CySA+, CCNA Security, SSCP, CISSP (or Associate))
Ability to obtain Privileged User Account (PUA) certification
Experience with Kubernetes, GitLab pipelines, Linux, and containerized environments
Experience supporting enterprise scale production systems
Experience with cloud services (preferably AWS) and cloud infrastructure
Familiarity with Elasticsearch, PostgreSQL, Logstash, Kibana, and Keycloak
Demonstrated success in cross functional coordination and execution
Strong communication skills and the ability to perform under pressure during incidents
You will stand out even more if you bring:
Experience with Agile methodologies
Experience with creating customized dashboards to track SLIs and other key performance indicators
Development experience (Bash, PowerShell, SALT, Python, Groovy, Java, etc.)
Experience with Appian or other low‑code platforms
Experience with technologies such as Kafka, AMQP/JMS, Prometheus/Grafana, GPU‑based Kubernetes, SALT automation, Nexus, or GraphQL
Knowledge of security best practices (authN/Z, secrets management, data protection)
Infrastructure‑as‑code experience (CloudFormation, Terraform, Pulumi)
AWS cloud certifications
group id: 10290999