DevOps Engineer
Job in
Toronto, Ontario, C6A, Canada
Listed on 2026-09-04
Listing for:
Newton.co
Full Time
position Listed on 2026-09-04
Job specializations:
-
IT/Tech
AWS, SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, IT Infrastructure
Job Description & How to Apply Below
- We are searching for a Dev Ops Engineer to improve how we build, deploy and run our systems
- This role works across infrastructure, CI/CD, observability and operational tooling in an AWS-based environment spanning backend, frontend and internal services
- Improve and maintain CI/CD, deployment workflows, and environment management across backend, web, and internal services
- Build, maintain and scale infrastructure across AWS and container based services
- Improve monitoring, alerting, logging, dashboards, tracing, and runbooks
- Work with engineers on safer deploys, rollback plans, and recovery from failures
- Automate repetitive operational work and improve internal tooling
- Maintain and improve infrastructure as code and deployment tooling
- Help improve failover planning, recovery procedures, and backup/restore testing for critical systems
- Support production systems and take part in on-call for critical services
- Manage and scale infrastructure across AWS, ECS, Docker, PostgreSQL, Redis, Celery, and Go/Python-based services
- Lead incident response and postmortems, and drive follow-up actions to reduce repeat issues
- Improve reliability, resilience, and operational readiness across critical systems
- Strong understanding of AWS networking, including VPCs, subnets, route tables, security groups, load balancers, DNS and connectivity between services
- Good understanding of PostgreSQL, Redis, queues, async workers, and scheduled jobs
- Experience with on-call and incident response for business-critical systems
- Experience with CI/CD and infrastructure automation
- Familiarity with Cloudflare or similar edge, networking or traffic management tooling
- Experience with Docker and ECS or Kubernetes
- A practical approach to automation, reliability and day to day operational work
- Experience with Datadog, Prometheus, Grafana, or similar observability tools
- Experience with Git Hub Actions, Pulumi, Terraform, or similar tooling
- Comfort with Linux, shell scripting, Python, and Go
- Experience running production systems in AWS or a similar cloud environment
- Strong troubleshooting skills across application, infrastructure, and data layers
Note that applications are not being accepted from your jurisdiction for this job currently via this jobsite. Candidate preferences are the decision of the Employer or Recruiting Agent, and are controlled by them alone.
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
Search for further Jobs Here:
×