| Job Position and Company | Posted | Location | Salary | Tags |
|---|---|---|---|---|
| $73k - $110k | ||||
| $120k - $160k | ||||
| | $54k - $90k | |||
| $135k - $170k | ||||
ISO 9001 Certified | 400+ students | Learn more | by Metana | ||
| $73k - $100k | ||||
| $58k - $100k | ||||
| $112k - $165k | ||||
| $106k - $165k | ||||
| $106k - $165k | ||||
| $32k - $70k |
DevOps Engineer - Canada Wide - Remote
Role Overview:
We are searching  for a DevOps Engineer to improve how we build, deploy and run our systems. This role works across infrastructure, CI/CD, observability and operational tooling in an AWS-based environment spanning backend, frontend and internal services.
• Improve and maintain CI/CD, deployment workflows, and environment management across backend, web, and internal services
• Build, maintain and scale infrastructure across AWS and container based services
• Improve monitoring, alerting, logging, dashboards, tracing, and runbooks
• Work with engineers on safer deploys, rollback plans, and recovery from failures
• Automate repetitive operational work and improve internal tooling
• Maintain and improve infrastructure as code and deployment tooling
• Help improve failover planning, recovery procedures, and backup/restore testing for critical systems
• Support production systems and take part in on-call for critical services
• Manage and scale infrastructure across AWS, ECS, Docker, PostgreSQL, Redis, Celery, and Go/Python-based services
• Lead incident response and postmortems, and drive follow-up actions to reduce repeat issues
• Improve reliability, resilience, and operational readiness across critical systems
Who you are:
• Experience running production systems in AWS or a similar cloud environment
• Experience with CI/CD and infrastructure automation
• Strong understanding of AWS networking, including VPCs, subnets, route tables, security groups, load balancers, DNS and connectivity between services
• ​​Comfort with Linux, shell scripting, Python, and Go
• Experience with Docker and ECS or Kubernetes
• Experience with GitHub Actions, Pulumi, Terraform, or similar tooling
• ​​Experience with Datadog, Prometheus, Grafana, or similar observability tools
• Good understanding of PostgreSQL, Redis, queues, async workers, and scheduled jobs
• Familiarity with Cloudflare or similar edge, networking or traffic management tooling
• A practical approach to automation, reliability and day to day operational work
• Experience with on-call and incident response for business-critical systems
• Strong troubleshooting skills across application, infrastructure, and data layers