We're looking for a Mid-Level DevOps Engineer who takes ownership of keeping our cloud infrastructure highly available, scalable, and running smoothly. Embedded within the Infrastructure Management team, you’ll work at the intersection of infrastructure reliability, pipeline automation, and developer enablement - helping our development squads move fast, deploy safely, and troubleshoot effectively.
This is a hands-on role where you'll be provisioning infrastructure as code, maintaining core network systems, optimizing deployment pipelines, and contributing to a modern engineering culture. You will have the opportunity to work with cutting-edge tools, tackle complex technical challenges, and help shape how we scale our infrastructure going forward.
Cloud Infrastructure & Network Operations
-
Maintain, patch, and optimize multi-environment cloud infrastructure to ensure system reliability and uptime in line with agreed SLAs.
- Monitor and manage core networking infrastructure, ensuring stable routing, firewalls, and redundant cloud connections (such as Site-to-Site VPNs and Direct Connect circuits).
- Ensure infrastructure configurations adhere to operational best practices and high-availability standards.
- Actively identify and execute opportunities for cloud cost optimization and infrastructure efficiency.
CI/CD Automation & & Infrastructure-as-Code (IaC)
-
Build, maintain, and optimize robust CI/CD pipelines to support fast and reliable code compilation and deployments.
- Create standardized, repeatable pipeline blueprints and templates that make onboarding new microservices seamless for development teams.
- Author robust, repeatable infrastructure blueprints and reusable templates using Terraform or CloudFormation to accelerate microservice delivery.
- Triage, debug, and resolve deployment failures, keeping our delivery pipelines green.
Developer Enablement & Self-Service
-
Work closely with software engineering teams to embed modern DevOps practices and guardrails throughout the development lifecycle.
- Build out clear documentation, runbooks, and self-service frameworks to empower developers to safely debug and resolve their own application-level pipeline issues.
- Act as a technical bridge, sharing operations and platform knowledge across teams through day-to-day collaboration and technical support.
Monitoring, Observability & Incident Response
-
Maintain and improve system observability, metrics collection, and alerting stacks to catch infrastructure performance issues before they impact users.
- Respond to infrastructure alerts, logs, and system events, investigating anomalies and escalating priority incidents calmly and quickly.
- Support regular testing and maintenance of infrastructure Business Continuity and Disaster Recovery (DR) frameworks.
Team & Continuous Improvement
-
Collaborate across the Infrastructure Management team to deliver on shared project objectives and sprints.
- Bring an Agile mindset to your daily work - looking for ways to continuously automate repetitive manual tasks and adopt modern engineering tooling.
From time to time, you may also be asked to:
-
Cover for other members of the Infrastructure Management team during periods of absence.
- Provide after-hours support as agreed - this may include scheduled on-call coverage, out-of-hours infrastructure releases, or priority incident resolution.
- Take on other reasonable tasks as directed that are consistent with the nature of this role.
- Cloud Infrastructure Experience: 3+ years of practical experience in a DevOps, SysOps, or CloudOps environment with strong, hands-on experience managing AWS and/or GCP environments.
- Kubernetes: Demonstrated experience configuring, managing and troubleshooting Kubernetes environments.
- Automation Mindset: A track record of automating repetitive operations tasks using scripting languages such as Python or Bash.
- Education: Bachelor's degree in Computer Science, IT, or equivalent practical, hands-on industry experience.
- Core Technical Stack: Strong technical proficiency with containerization and orchestration tools like Docker and Kubernetes (EKS/GKE).
- CI/CD Tools: Solid understanding of modern pipeline platforms (e.g., GitHub Actions, GitLab CI, Buildkite, or Bamboo).
- Networking Fundamentals: A solid understanding of core networking, including DNS, VPC routing, subnets, and establishing secure connections (VPN tunnels).
- Problem-Solving & Communication: Strong troubleshooting instincts with the ability to navigate unfamiliar infrastructure incidents quickly and explain technical outcomes clearly to other teams.
Canstar has a diverse mix of roles available across the business. Check out our current opportunities!