About The Role
The DevOps Engineer builds and operates the infrastructure that runs production services at scale, with ownership spanning cloud platforms, deployment pipelines, observability, and incident response. The role focuses on creating reliable, repeatable systems across AWS, Kubernetes, and infrastructure-as-code environments.
You will partner with software engineers and SREs to improve release velocity without compromising availability or security. The team needs an engineer who can automate operational work, diagnose distributed-system failures, and turn production lessons into durable platform improvements.
Key Responsibilities
- Design and maintain highly available AWS infrastructure using Terraform, including networking, compute, IAM, databases, and monitoring resources
- Build and improve CI/CD pipelines with GitHub Actions, GitLab CI, or Jenkins to support automated testing, deployment, rollback, and artifact management
- Operate Kubernetes-based services, including cluster configuration, Helm releases, ingress, autoscaling, secrets management, and workload troubleshooting
- Implement observability across production systems using Prometheus, Grafana, CloudWatch, OpenTelemetry, or comparable tooling
- Automate operational workflows with Python, Go, or Bash, reducing manual intervention and improving repeatability across development and production environments
- Participate in on-call rotations, lead incident investigation and remediation, and document runbooks, postmortems, and reliability improvements
What We Are Looking For
- 3–8 years of experience in DevOps, SRE, platform engineering, or a closely related infrastructure role
- Hands-on experience operating production workloads in AWS, including core services such as EC2, EKS, VPC, IAM, S3, RDS, and CloudWatch
- Strong Kubernetes and Docker experience, including deployment configuration, service discovery, resource management, and troubleshooting
- Proficiency with Terraform or a comparable infrastructure-as-code tool and practical experience managing infrastructure through code review and automated workflows
- Solid understanding of Linux systems, networking, security fundamentals, Git, and CI/CD pipeline design
- Bachelor’s degree in computer science, engineering, information systems, or equivalent practical experience; Bonus: experience with Go, Python automation, service meshes, Argo CD, incident management, and compliance-focused infrastructure