Director of DevOps

🕒 June 29

🇺🇸 United States – Remote

💵 $220k - $260k / year

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Convoso

Convoso

201 - 500 employees

Founded 2006

💼 Consulting

📣 Marketing

📦 Logistics

Consulting • Marketing • Logistics

Convoso is a leading sales platform that leverages AI-powered call center dialer software to enhance sales performance. The company focuses on providing robust solutions for outbound and inbound sales, allowing businesses to automate workflows, improve compliance, and increase engagement. By integrating seamlessly with CRM systems, Convoso empowers sales teams to achieve higher conversion rates and meet their growth targets through efficient communication strategies.

📋 Description

• Own service reliability across Convoso’s platform, being a champion for quality and resilience, and drive adherence to internal and external SLAs, SLOs, and SLIs • Collaborate with other teams to establish key metrics and dashboards in order to drive quality in the SDLC • Own the incident management process, including monitoring and observability systems, and major incident response process, tracking and reducing MTTR continuously • Leverage chaos engineering principles and “game days” to proactively test the resilience of Convoso’s platform • Establish and maintain a CI/CD center of excellence • Drive automation for CI/CD processes across different tech stacks and cloud providers. • Recruit and lead DevOps teams across multiple product lines that excel at applying industry best practices • Develop and support automated, scalable solutions to deploy and manage our global infrastructure. • Maintain and improve the use of automation tools for infrastructure provisioning, configuration, and deployment. • Work with Development and Operations personnel to define CloudOps processes, introduce new insights and technologies so that we can stay on the cutting edge. • Provide mentorship and expertise on system options, risk and impact management, as well as cost vs. benefit analysis. • Uphold and ensure security requirements for tooling, systems, and environments are met and protect the assets of the company and our customers. • Lead troubleshooting of system and performance problems in Prod/QA/Dev environments. • Identify improvements to reduce technical debt. • Lead infrastructure focused on optimization and performance. • Lead DevOps Team • Author a dynamic training and skills improvement plan. • Acquire team level certifications applicable to AWS or Google Cloud • Constantly evaluate our services to ensure high quality • Conduct reviews to identify optimization opportunities in processes or systems • Contribute to the creation of departmental procedure documents and working instructions • Works with consultants and team members to design the solutions to be implemented • Delegate tasks related to solutions and monitors team members’ progress • Inspire and motivate teamwork for achieving goals • Provide mentoring and identify training opportunities to team members, staying current in the latest technologies and best practices.

🎯 Requirements

• BS in Computer Science, Computer Engineering, or related technical field • 5+ years experience as a Site Reliability Engineer leader leading multiple teams with a passion for automation and continuous improvement • 5+ years experience in operational management of SaaS production applications • 5+ years working with Agile/DevOps development teams • 5 years minimum related experience in infrastructure management in a Linux environment • Deep knowledge of cloud platforms like AWS or GCP and management of on-prem and hybrid environments • Track record as a player-coach who is comfortable both managing and doing • Experience deploying and operating containerized web applications • Experience with cloud automation/provisioning and configuration management tools using Ansible, Chef, Puppet, Docker or Salt • Experience with Containerization Tools such as Docker and Kubernetes • Experience in planning, creating, implementing and maintaining a scalable software development infrastructure • Knowledge of development, build, and CI/CD tools such as Git, GitHub and Jenkins • A security background that will come in handy as we navigate through the process of SOC2 certification • Good people and line management skills and the ability to recruit and develop a high performing team • Process-oriented with great documentation skills • Excellent oral and written communication skills.

🏖️ Benefits

• Competitive compensation package • Stock options • 100% covered premiums for employees; Medical, Dental, Basic life insurance, Long term disability • Affordable Vision plan and optional FSA • PTO, Paid Sick Time, Holidays, Bereavement time, Parental Leave • Your birthday off • 401k program with generous company match • No cost Employee Assistance Program and Travel Assistance • Monthly Gym membership reimbursement • Monthly credits toward food & beverage • Company Outings • On and offsite team building events • Paid training for departments • Apple laptop (most roles) • And a team of highly experienced and kind colleagues!

Apply Now

Similar Jobs

🕒 June 29

FluidStack

11 - 50

🤖 Artificial Intelligence

Principal Operations Engineer overseeing critical operations in data centers for Fluidstack. Leading on-call escalation, root cause analysis, and operational excellence in real-time situations.

🇺🇸 United States – Remote

💵 $150k - $250k / year

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 June 24

Redox

201 - 500

🏥 Healthcare

⚕️ Healthcare Insurance

☁️ SaaS

DevSecOps Engineer ensuring secure software development at Redox, enhancing healthcare data exchange. Collaborating with platform engineers to implement security best practices across the AWS/EKS infrastructure.

🕒 June 23

Lyric - Clarity in motion.

201 - 500

🏥 Healthcare

💼 Consulting

📦 Logistics

Azure DevOps Engineer at Lyric managing Azure infrastructure for healthcare technology solutions. Focus on security, reliability, and operational efficiency in a remote role.

🇺🇸 United States – Remote

💵 $150.3k - $225.4k / year

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 June 23

SAIC

10,000+ employees

☁️ SaaS

📣 Marketing

🏢 Enterprise

DevSecOps Engineer providing exceptional DevOps engineering for advancing CI/CD and automating pipelines. Must have deep proficiency in AWS, Azure, and DevSecOps tools.

🇺🇸 United States – Remote

🔥 Funding within the last year

💰 $500M Post-IPO Debt - SAIC on 2025-09

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 June 22

Kong Inc.

201 - 500

💼 Consulting

📦 Logistics

🔌 API

Staff Site Reliability Engineer for Kong's Volcano platform overseeing reliability and infrastructure scaling. Collaborating on SRE practices and emerging technology evaluations.

🇺🇸 United States – Remote

💵 $150k - $210k / year

💰 $100M Series D on 2021-02

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)