Reliability Engineer
Chenega Corp · Georgia
📍 Atlanta, GA💰 $95,000via icimsFirst listed here 2026-09-12
Apply on company site ↗
Career Moonshot pulls this listing straight from the employer's hiring system — no recruiter middleman, no reposts. Applying takes you directly to Chenega Corp.
Overview
Position contingent on contract award - d etails below are subject to change based on final award.
Come join a company that strives for Extraordinary People and Exceptional Performance ! Chenega Services & Federal Solutions, LLC , a Chenega Professional Services’ company, is looking for a Reliability Engineer. In this role, the Reliability Engineer will ensure the reliability, scalability, and operational health of EDAV’s Azure cloud environment. Terraform is central to this position: the engineer will independently design, build, review, and troubleshoot Infrastructure-as-Code for mission‑critical environments. The role involves close collaboration with platform engineers, developers, security teams, and product stakeholders to automate cloud infrastructure, improve Kubernetes operations, and resolve issues impacting the availability of EDAV data and analytics services.
Our company offers employees the opportunity to join a team where there is a robust employee benefits program, management engagement, quality leadership, an atmosphere of teamwork, recognition for performance, and promotion opportunities. We actively strive to channel our highly engaged employee’s knowledge, critical thinking, innovative solutions for our clients.
Responsibilities
Design, implement, maintain, and troubleshoot production Azure infrastructure using Terraform.
Support reliability, performance, and availability of workloads in Azure Kubernetes Service (AKS).
Troubleshoot cloud infrastructure, networking, Kubernetes, and application reliability issues.
Automate cloud operations to reduce manual work and improve consistency.
Collaborate with development and operations teams to enhance deployment and incident‑response practices.
Implement and refine monitoring, alerting, dashboards, and operational reporting.
Identify reliability risks and recommend improvements to cloud architecture and processes.
Document infrastructure, procedures, troubleshooting guidance, and operational runbooks.
Qualifications
4+ years in cloud infrastructure, systems engineering, DevOps, SRE, or similar roles
2+ years of hands‑on Microsoft Azure experience
2+ years of hands‑on Terraform expertise to design reusable modules, manage state, troubleshoot failures, and maintain production infrastructure
Experience administering/supporting Kubernetes (preferably AKS)
Experience supporting cloud‑hosted systems and automating cloud operations
Knowledge of cloud networking concepts and troubleshooting
Possession of strong analytical and problem‑solving abilities
Ability to work on-site in Atlanta, GA.
Ability to obtain/maintain Public Trust/Suitability clearance
Bachelor’s degree or equivalent experience
Nice to Haves:
Experience with certificate lifecycle management (maintaining, renewing, rotating certs)
Experience creating operational dashboards or reports (Power BI preferred)
Experience with observability platforms (Grafana, Prometheus, Elastic, Splunk)
Experience integrating Azure resources with Active Directory
Experience using agentic coding tools (Claude Code, Codex, GitHub Copilot) for applications or infrastructure automation
Azure, Kubernetes, or Terraform certifications
Estimated Salary/Wage USD $95,000.00/Yr. Up to USD $105,000.00/Yr.
More Georgia jobs
Georgia jobs · Browse all locations