Skip to main content

Bluestaq US External

Delivery Engineer - Cloud & Automation

Colorado SpringsFrom $140kmidAdded 1 month ago

About this role

Bluestaq seeks a Delivery Engineer to operate and evolve a Kubernetes-based platform on AWS, managing deployment automation, production health, and incident response for the Unified Data Library—a mission-critical data platform used in government and commercial space operations. This hands-on platform/SRE role involves infrastructure deployment, troubleshooting rollouts, monitoring performance, and directly remediating production issues.

What you'll do

  • Deploy and manage infrastructure using Terraform/Terragrunt across EKS environments
  • Monitor production health through Grafana dashboards and distributed tracing systems
  • Debug and resolve failed deployments and production incidents in Kubernetes
  • Perform incident response and remediation for mission-critical services
  • Improve runbooks, dashboards, and operational documentation
  • Triage alerts and independently resolve common production issues

What they're looking for

  • Kubernetes and EKS
  • Terraform/Terragrunt
  • AWS cloud infrastructure
  • Grafana monitoring and observability
  • Argo CD or similar deployment tools
  • Linux systems administration
  • Production incident response
  • Infrastructure-as-code practices
Apply with Autofill

Opens the application — the Jobs AI extension fills it for you. Set up autofill

Opens the official application on the employer’s site. No login required.

Bluestaq US External

Bluestaq builds a Unified Data Library, a mission-critical data platform serving government and commercial space operations on Kubernetes infrastructure. The company is hiring Delivery Engineers and platform/SRE roles to operate, evolve, and maintain this AWS-based system, focusing on deployment automation, production health, and incident response.

View all jobs at Bluestaq US External

Likely interview questions

  • Walk us through your experience with Kubernetes and EKS. Can you describe a time when you debugged and resolved a failed deployment or pod issue in production?
  • Tell us about your hands-on experience with Terraform or Terragrunt. How have you used Infrastructure as Code to automate deployments?