Senior Site Reliability Engineer, Vulcan (AI Security product)
Posted 2026-09-08 ยท Verified live 2026-09-21
See how your real experience scores against this role โ our AI drafts an honest, verifiable resume tailor. Nothing invented, ever. You approve everything.
Get matched โ join the waitlist Apply on company site โAbout this role
<h3><strong>Job Overview</strong></h3>
<p>We are looking for a hands-on infrastructure engineer to own the deployment, migration and troubleshooting of on-premise Kubernetes environments for enterprise and government clients, including airgapped, high-security data center environments where remote access is not possible. </p>
<p>This is a client-facing, on-site role: you will be the technical authority in the room, responsible for executing complex infrastructure changes correctly the first time, diagnosing failures independently under pressure and communicating clearly with client stakeholders throughout. </p>
<p>This role carries real ownership; you will be expected to understand the systems deeply enough to make sound judgment calls when things don't go to plan, without waiting on remote support.</p>
<p>In this role, you will play a vital part in supporting our Cybersecurity business, Vulcan. Vulcan is a cybersecurity solution for GenAI, providing red and blue team services to ensure compliance and security. </p>
<p>Learn more about us ๐</p>
<ul>
<li>Vulcan product: <a href="https://vulcanlab.ai/" rel="noopener noreferrer">https://vulcanlab.ai/</a></li>
<li>Vulcan LinkedIn: <a href="https://www.linkedin.com/company/vulcanlab-ai/" rel="noopener noreferrer">https://www.linkedin.com/company/vulcanlab-ai/</a></li>
<li>AIFT group: <a href="https://aift.io/" rel="noopener noreferrer">https://aift.io/</a></li>
</ul>
<p> </p>
<h3><strong>Responsibilities</strong></h3>
<ul>
<li>Plan and execute on-prem Kubernetes cluster deployments, upgrades and infrastructure migrations (including IP re-addressing, certificate rotation and cluster reconfiguration) in production and airgapped environments </li>
<li>Diagnose and resolve failures independently on-site </li>
<li>Own the full infrastructure stack end-to-end: Kubernetes control plane and data plane, PostgreSQL (primary/replica replication), distributed storage (e.g. SeaweedFS/Ceph/similar), private container registries and centralized logging (ELK or equivalent)</li>
<li>Validate deployment tooling (scripts, installers, automation) thoroughly in lab/staging environments before any client-facing execution </li>
<li>Represent the technical work directly to client stakeholders on-site: explain status, failures and remediation plans clearly </li>
<li>Travel to client data centers (including airgapped/restricted-access sites) as required, sometimes on short notice, for deployment and go-live support</li>
<li>Write clear, structured runbooks, decision trees and incident reports that others (including less experienced engineers) can follow under pressure </li>
<li>Escalate risks proactively to internal leadership, not just after something has gone wrong </li>
</ul>
<h3><strong>Requirements</strong></h3>
<p><strong>Technical:</strong> </p>
<ul>
<li>5-6 years of hands-on experience with Kubernetes in production, including at least one on-premise (not purely cloud-managed) deployment </li>
<li>Solid understanding of etcd internals. Quorum, peer membership, failure recovery, not just kubectl-level familiarity </li>
<li>Experience with kubeadm-based cluster bootstrapping and certificate management (SANs, CA rotation, renewal) </li>
<li>Working knowledge of PostgreSQL replication, Linux networking fundamentals (DNS, NTP, firewalls) and container registries (Docker Distribution or similar) </li>
<li>Comfortable working entirely from the Linux command line, writing and debugging bash scripts and reading unfamiliar automation tooling under time pressure </li>
<li>Experience with at least one distributed storage system (SeaweedFS, Ceph, MinIO or similar) is a strong plus </li>
<li>GPU-enabled Kubernetes nodes (NVIDIA device plugin, container toolkit) experience is a plus, not required</li></ul>
One of thousands of fresh listings refreshed nightly, built for cleared & defense careers.
Browse all jobs