
5001 - 10000 employees
Founded 1997
💼 Consulting
📣 Marketing
☁️ SaaS
Consulting • Marketing • SaaS
Valtech is a global digital agency focusing on experience innovation. They strive to transform businesses through a combination of technology, marketing, and data strategies. Valtech helps companies elevate their digital presence and drive commerce strategies, enhance enterprise digital transformations, and unlock marketing and performance potential. They also utilize data and AI to help organizations harness the power of information. With offices around the world, Valtech partners with businesses to shape their digital futures, offering a range of services and insights designed to enhance customer experiences.
🔥 14 hours ago
🇨🇦 Canada – Remote
💵 $100k - $150k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
👻 Ghost score 0%
🗣️🇫🇷 French Required
Improve your chances of getting an interview by checking your resume score before you apply.

5001 - 10000 employees
Founded 1997
💼 Consulting
📣 Marketing
☁️ SaaS
Consulting • Marketing • SaaS
Valtech is a global digital agency focusing on experience innovation. They strive to transform businesses through a combination of technology, marketing, and data strategies. Valtech helps companies elevate their digital presence and drive commerce strategies, enhance enterprise digital transformations, and unlock marketing and performance potential. They also utilize data and AI to help organizations harness the power of information. With offices around the world, Valtech partners with businesses to shape their digital futures, offering a range of services and insights designed to enhance customer experiences.
• Define and implement observability strategies, standards, and governance across applications and platforms • Design and maintain monitoring, alerting, dashboarding, and reporting solutions using Dynatrace or equivalent observability platforms • Establish and drive SRE best practices, including SLIs, SLOs, error budgets, and symptom-based alerting • Partner with engineering and product teams to improve system reliability, performance, and operational maturity • Develop standards for tagging, ownership, dashboard design, access management, and alerting governance • Support non-specialized observability teams through guidance, coaching, and knowledge transfer • Lead technical workstreams, prioritize initiatives, and ensure delivery within defined timelines and budgets • Analyze distributed systems and troubleshoot complex production issues using monitoring and tracing data • Promote documentation, operational rigor, and continuous improvement across engineering teams • Collaborate within a distributed, multilingual environment
• Significant experience in Site Reliability Engineering within large-scale production environments • Deep understanding of Service Level Indicators (SLIs), Service Level Objectives (SLOs), error budgets, and symptom-based alerting • Proven expertise with enterprise observability platforms such as Dynatrace, Datadog, New Relic, or AppDynamics • Strong experience with Application Performance Monitoring (APM), Real User Monitoring (RUM), monitoring agents and instrumentation, alerting strategies, RBAC, SLO management, and tagging and governance models • Strong knowledge of OpenTelemetry (OTEL) and distributed tracing • Experience with composable, microservices-based architectures • Hands-on production experience with AWS and Kubernetes • Experience with infrastructure and operational automation • Practical knowledge of Terraform, Bash scripting, and Python scripting • Experience with CI/CD tools such as GitLab CI or equivalent pipeline/workflow platforms • Demonstrated ability to lead technical initiatives and workstreams • Experience working within complex operational and Agile environments • Strong stakeholder management and collaboration skills • Excellent communication skills in both French and English; French-speaking skills are needed • Strong documentation practices, organizational skills, and attention to detail • High degree of autonomy and ownership
• Comprehensive insurance plan with Gold, Silver, or Bronze modules; employer contribution up to 80%; short- and long-term disability coverage • Dialogue via Sun Life virtual healthcare services • Employee and Family Assistance Program • Complete mental health support program • $500 Personal Spending Account for healthcare reimbursements, gym memberships, public transit passes, office supplies, or RRSP contributions • RRSP retirement plan with 100% employer matching through DPSP, up to 4% • Flexible vacation policy, including 5 days during the probation period • $30/month Personal Technology Reimbursement, offered from day 1 • Winter holiday company closure • Flexible scheduling throughout the year • Growth opportunities, continuous learning, professional growth, and international career opportunities • Inclusion and accessibility support, including reasonable interview accommodations
Apply Now🕒 2 days ago
Senior DevOps Engineer helping JFrog customers build CI/CD platforms using JFrog’s liquid software tools. Designing cloud-native pipelines and guiding customers, communities, and internal teams.
🇨🇦 Canada – Remote
💵 $130k - $150k / year
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🕒 August 20
Senior Site Reliability Engineer deploying Kubernetes-based AI infrastructure on NVIDIA-certified hardware for Mirantis. Ensuring reliable, secure, scalable cloud operations and customer delivery.
🕒 August 20
DevOps Engineer building and maintaining cloud infrastructure, automation, and CI/CD pipelines for Calliere's software platform. Operating containers, observability tooling, and production workloads across public clouds.
🇨🇦 Canada – Remote
💵 $150k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🕒 August 19
Senior SRE scaling AWS, Kubernetes, Kafka, and CI/CD infrastructure for Blackpoint Cyber’s cybersecurity technology. Automating operations, observability, deployments, and incident response.
🇨🇦 Canada – Remote
💵 CA$131k - CA$164.3k / year
💰 $190M Series C on 2023-06
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🕒 August 18
Site Reliability Engineer operating Yelp’s Kafka-based streaming infrastructure. Automating upgrades, scaling, incident recovery, and reliable data pipelines across Canada.
🇨🇦 Canada – Remote
💵 $135k - $185k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)