From ShyftLabs's posting. “We” and “our” refer to the employer.
Position Overview:
We are looking for a Lead DevOps Engineer to join our growing team. The DevOps Engineer will partner with application developers to automate and accelerate the testing, release, and deployment of applications into a runtime environment quickly and reliably while ensuring high availability and uptime. You will be responsible for working with a team to develop, optimize, and deliver cloud computing solutions, and development environments through the development life cycle.
ShyftLabs is a growing data product company founded in early 2020 and works primarily with Fortune 500 companies. We deliver digital solutions built to help accelerate the growth of businesses in various industries, by focusing on creating value through innovation.
Key Responsibilities
Infrastructure & Cloud Architecture
Design, implement, and maintain scalable, secure, and highly available infrastructure on AWS (VPC, EC2, ECS/EKS, RDS, S3, Lambda, IAM, CloudFront, Route53, etc.)
Own the infrastructure-as-code strategy using Terraform, including module design, state management, and multi-environment provisioning
Drive cost optimization, security hardening, and architectural best practices across all environments (dev, staging, production) • Lead disaster recovery, backup, and business continuity planning for critical systems CI/CD & Automation
Architect and maintain robust CI/CD pipelines (GitHub Actions preferred) for build, test, and deployment automation across multiple services/teams
Build automation tooling and scripts using Python and Shell to eliminate manual operational work
Implement GitOps practices and policy-as-code for infrastructure changes (plan/apply gates, automated compliance checks)
Own release management processes, deployment strategies (blue-green, canary, rolling), and rollback procedures Source Control & Collaboration
Manage and enforce GitHub best practices — branching strategy, code review workflows, repo structure, access controls, and GitHub Actions workflows
Collaborate closely with development, QA, and security teams to embed DevOps practices into the SDLC Reliability & Security
Set up and maintain monitoring, logging, and alerting (CloudWatch, Prometheus/Grafana, Datadog, or similar)
Lead incident response and root cause analysis for production issues; drive postmortems and long-term fixes
Implement and enforce security best practices — IAM least privilege, secrets management, vulnerability scanning, compliance guardrails Leadership & Mentorship
Lead and mentor a team of DevOps/Infrastructure engineers; conduct code/infra reviews and guide technical decision-making • Define and enforce infrastructure standards (tagging, naming conventions, module reuse, documentation)
Partner with engineering leadership on infrastructure roadmap, tooling decisions, and capacity planning
Own vendor/tool evaluations (build vs. buy) for infrastructure and DevOps tooling
Required Skills & Experience
7–10 years of experience in DevOps, SRE, or Infrastructure Engineering roles, with at least 2+ years in a lead or senior technical ownership capacity
Deep, hands-on expertise in AWS — networking, compute, storage, IAM, and security services
Strong proficiency in Terraform — module design, remote state management, workspaces, drift detection, and large-scale IaC governance
Solid experience with GitHub — Actions, branching strategies, repo governance, and workflow automation
Strong scripting/programming ability in Python and Shell/Bash for automation and tooling
Proven experience building and maintaining CI/CD pipelines end-to-end (build, test, deploy, rollback)
Experience with containerization and orchestration (Docker, ECS or Kubernetes/EKS)
Strong understanding of security best practices, IAM policies, and secrets management (Secrets Manager, Vault, or similar)
Experience with monitoring/observability tooling (CloudWatch, Prometheus, Grafana, ELK, Datadog, etc.)
Excellent troubleshooting skills across networking, compute, and application layers
Strong communication skills with the ability to lead technical discussions and mentor engineers Preferred / Nice-to-Have
Terraform Certified Associate or equivalent hands-on Terraform Enterprise/Cloud experience
Experience with policy-as-code tools (Sentinel, OPA, Checkov)
Experience with Kubernetes (EKS) at scale, including IRSA, autoscaling, and cluster upgrades
Experience in a regulated or high-compliance environment (SOC2, ISO 27001, HIPAA, etc.)
Familiarity with configuration management tools (Ansible, Chef, Puppet
About this listing
LAYIQ is an independent job-discovery service. This listing does not imply a partnership with or endorsement by the employer. Review the original posting for current details and availability.