Senior DevOps Engineer
About Circadence Circadence is a premier provider of cyber range solutions, boasting unparalleled proficiency in crafting, building, implementing, and managing tailored enterprise frameworks. Focused primarily on cybersecurity infrastructure, simulation, and education, Circadence excels across various tiers of complexity, seamlessly connecting academic, enterprise, and government use cases. Our unique approach to cybersecurity training stems from the power of gamification and active-learning models.Circadence’s cybersecurity training leverages gamified, cloud-based training and awareness platforms to provide personalized, agile capabilities at scale.Gamified-learning capabilities: Enjoyable, approachable, scalable, and an enduring practice for the next generation of professionals.NIST/NICE Framework alignment: Users can be confident they are training to industry-leading dimensions of cyber excellence.On demand: Accessible 24/7, and immersive, unlike other cyber training programs.The RoleCircadence builds and delivers software to create immersive experiences for cybersecurity training and assessment. Circadence is hiring a Senior DevOps Engineer to build and run the range-building capabilities of Project Ares, our cybersecurity training platform: the templates, images, networks, and provisioning systems that stand up our cyber ranges.Environments are built and destroyed continuously, in different Azure subscriptions and regions, and every one of them has to come up correct, come up fast, and come up unattended, because a learner is already waiting on it. That makes the failures the interesting kind: a template that deploys cleanly in one region and not the next, a guest agent that never reports in, a network path that passes every health check and still carries nothing. Most of this job is finding the real cause of that class of problem and then building the check that catches it before anyone else sees it.This is a senior individual contributor role with a domain of its own. You own our range-building capability: the features, the work, and the direction it takes. The work spans building and maintaining production infrastructure, creating prototype ranges, and integrating those changes into the platform alongside the rest of the engineering team. We care more that you work fluently with modern AI-native tooling and can get a capability from idea to something real than that you arrive knowing our stack.How You Will Add ValueRange infrastructureOwn the infrastructure-as-code for our range environments, along with the authoring standards that templates follow.Design how ranges are networked, including how remote consoles reach them, weighing isolation, per-environment cost, and the constraints content places on addressing.Own the image pipeline: how VM images are built, versioned, distributed across regions and subscriptions, and kept in sync with every environment that needs them.Research the options and make recommendations on where our range infrastructure goes next, backed by measurement rather than assumption.Reliability and correctnessBuild the checks that establish whether a deployed environment is correct: static analysis of templates before deployment, smoke deployments, and assertions that run against a live environment to confirm it is reachable and configured the way its content expects.Debug complex cloud infrastructure failures end to end, including disk and image problems, startup script and guest agent failures, quota and capacity limits, provisioning errors, and network paths that fail without reporting an error.Improve provisioning reliability and time to ready, and make the cost of an environment visible.Contribute to the platform services that provision and manage ranges, so the infrastructure and the code driving it stay in step.Partnership and enablementWork with our content development team as their infrastructure counterpart: review the range templates they author, unblock their deployments, and give them a foundation they can build on without needing you.Define and document what a range has to provide for the platform to run it, so content authors can work against a stable interface.Document what you build, including runbooks for the failure modes you have already worked through.Required Skills & Experience5 or more years in DevOps, cloud engineering, or infrastructure-focused software engineering, at a senior level.Deep hands-on experience running production infrastructure in a major public cloud. We run on Azure, and strong AWS or GCP experience with a track record of picking up a new platform quickly is equally welcome.Infrastructure-as-code at production scale (Terraform, Bicep, ARM, CloudFormation, Pulumi or similar), including reconciling a stack that has drifted from what is actually deployed without breaking what is running on it.Experience authoring cloud resource templates directly, not only consuming modules somebody else wrote.Practical networking: routing, DNS, NAT, firewalls, private connectivity between networks, and the ability to trace why traffic between two machines does not arrive.VM lifecycle and guest configuration on both Windows and Linux: image capture and generalization, cloud-init and custom script extensions, guest agents, remote command execution.Strong scripting and automation skills (PowerShell, Bash, Python).A habit of measuring before recommending, and of turning an incident into a check that keeps it from recurring.Fluency with current AI development tooling, including coding agents, used for real work: reading an unfamiliar codebase or specification as readily as writing and debugging, and a realistic sense of where it earns its place.Comfort owning a technical area end to end, and explaining a recommendation to people who will not be reading the code.Preferred Skills & ExperienceAzure in depth: compute, managed disks and images, virtual networks and peering, network security groups, private endpoints, Key Vault, quotas, and policy.TypeScript, and experience contributing to application code that drives infrastructure.Image distribution at scale: shared image galleries, Packer, or AMI and managed image pipelines.Remote access and console brokering: RDP, SSH, VNC, Apache Guacamole or similar.Multi-tenant or customer-managed subscription models, and the credential handling they require.Infrastructure testing and policy tooling: template linting, policy-as-code, deployment smoke suites.Cloud cost engineering: right-sizing, capacity planning, spend attribution.Background in cyber ranges, virtual labs, simulation environments, or CTF infrastructure.Personal AttributesExcellent verbal and written communication skills.Proactive problem-solver with strong analytical thinking.Self-motivated, with the ability to work effectively in a remote environment.Collaborative mindset, and comfortable being the infrastructure partner to a team that is not infrastructure-focused.Adaptable to changing priorities and requirements.Benefits & PerksMedical, dental and vision insurance, 401(k) match, Flexible PTO, Paid maternity/paternity leave, Disability insurance.Circadence Corporation is proud to be an equal opportunity employer.https://www.eeoc.gov/know-your-rights-workplace-discrimination-illegal