
10,000+ employees
Founded 1984
💼 Consulting
🏥 Healthcare
📚 Education
💰 Post-IPO Equity on 2015-07
Consulting • Healthcare • Education
CDW is a leading provider of technology solutions, offering a comprehensive range of products and services to address the needs of enterprises, government, healthcare, and education sectors. They specialize in providing IT infrastructure services, managed services, and security solutions. CDW partners with top technology brands to deliver customized solutions that enhance business productivity and performance across various industries. With a strong focus on digital transformation, security, and cloud solutions, CDW supports businesses in navigating the complexities of modern IT environments. Their expertise spans several domains, including digital workspaces, cybersecurity, and enterprise solutions, making CDW a key partner in driving technological advancements.
🔥 0 minutes ago
🇺🇸 United States – Remote
💵 $106k - $150.2k / year
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🦅 H1B Visa Sponsor
👻 Ghost score 0%
Ansible
AWS
Azure
Cloud
Distributed Systems
Google Cloud Platform
Grafana
Java
Kubernetes
Linux
Microservices
MongoDB
MySQL
Postgres
Prometheus
Python
Spring
Spring Boot
SpringBoot
Improve your chances of getting an interview by checking your resume score before you apply.

10,000+ employees
Founded 1984
💼 Consulting
🏥 Healthcare
📚 Education
💰 Post-IPO Equity on 2015-07
Consulting • Healthcare • Education
CDW is a leading provider of technology solutions, offering a comprehensive range of products and services to address the needs of enterprises, government, healthcare, and education sectors. They specialize in providing IT infrastructure services, managed services, and security solutions. CDW partners with top technology brands to deliver customized solutions that enhance business productivity and performance across various industries. With a strong focus on digital transformation, security, and cloud solutions, CDW supports businesses in navigating the complexities of modern IT environments. Their expertise spans several domains, including digital workspaces, cybersecurity, and enterprise solutions, making CDW a key partner in driving technological advancements.
• Improve reliability, scalability, performance, security, and operational excellence across Managed Services applications and customer-connectivity platforms • Serve as a senior technical escalation point for complex operational issues and resolve incidents exceeding the SRE team’s expertise • Design, develop, and maintain automation using Ansible and Python • Establish and mature reliability and observability practices, including SLIs, SLOs, Error Budgets, metrics, logs, traces, alerting, and dashboards • Lead major incident response, Problem Management, root cause analysis, post-incident reviews, and corrective actions • Perform operational readiness and resiliency reviews • Identify reliability risks, operational gaps, technical debt, and continuous-improvement opportunities • Troubleshoot complex issues across applications, Kubernetes environments, infrastructure, networking, identity services, databases, certificates, cloud services, and third-party integrations • Mentor engineers and partner with development, infrastructure, security, and business teams • Improve technical standards, operational practices, documentation, release quality, and production readiness • Participate in a scheduled primary and secondary on-call rotation after training and readiness assessment
• Bachelor’s degree in Computer Science, Software Engineering, Information Technology, or a related field and 7+ years of experience in Software Engineering, Site Reliability Engineering, DevOps, Platform Engineering, or a related discipline; or 10+ years of equivalent professional experience • 5+ years of experience administering Linux-based systems in enterprise environments • 5+ years of experience developing automation solutions using Ansible • 3+ years of experience developing automation and operational tooling using Python • 3+ years of experience supporting Kubernetes or other container orchestration platforms in production environments • Experience supporting business-critical production applications and distributed systems • Experience leading major incident response, Problem Management, root cause analysis, and corrective action initiatives • Experience working within ITIL-aligned Incident, Problem, and Change Management processes • Experience implementing observability solutions using metrics, logs, traces, alerting, and dashboards • Experience defining and measuring service reliability using SLIs, SLOs, and Error Budgets • Strong knowledge of networking, CI/CD pipelines, source control, certificates, secrets management, and modern software delivery practices • Demonstrated ability to troubleshoot complex technical issues, influence technical direction, mentor engineers, and communicate effectively with technical and non-technical stakeholders • Experience reviewing, troubleshooting, and making minor enhancements to existing Java-based applications; this is not primarily a Java application-development role • Experience with Spring Boot, REST APIs, messaging technologies, and microservices architectures, is a plus • Experience with OpenTelemetry, Dynatrace, Prometheus, Grafana, or similar observability platforms, is a plus • Experience supporting Azure, AWS, or GCP cloud services and cloud-native architectures, is a plus • Experience troubleshooting and supporting PostgreSQL, MySQL, MongoDB, DB2, IBMi (AS/400), or other enterprise database platforms, is a plus • Experience implementing OAuth 2.0, JWT, RBAC, and secure application practices, is a plus • Familiarity with Agile, Scrum, or SAFe delivery methodologies, is a plus • Experience supporting enterprise-scale Managed Services environments, is a plus • Kubernetes, cloud, security, or automation-related certifications, is a plus
• Annual bonus target of 10% subject to terms and conditions of plan • Benefits package (see benefits overview) • Remote work arrangement • AI learning and experimentation culture • Support for professional growth and development
Apply Now🔥 15 minutes ago
Senior DevSecOps Engineer securing AWS/GCP infrastructure, Kubernetes, and CI/CD for Fuze Health’s national pharmacy platform. Driving compliance, resilience, and secure engineering at scale.
🇺🇸 United States – Remote
💵 $128k - $160k / year
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🔥 38 minutes ago
Principal SRE directing reliability, observability, and automation for Merative’s medical imaging cloud platforms. Establishing SLOs, resilience, incident management, and infrastructure automation across teams.
🇺🇸 United States – Remote
💵 $174.7k - $262k / year
⏰ Full Time
🟠 Senior
🔴 Lead
⛑ DevOps & Site Reliability Engineer (SRE)
🦅 H1B Visa Sponsor
🔥 3 hours ago
Cloud Site Reliability Engineer modernizing cloud systems for GDIT’s federal court case-management program. Ensuring resilient operations, monitoring, automation, FinOps, and DevSecOps delivery.
🇺🇸 United States – Remote
💵 $128k - $173.2k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🦅 H1B Visa Sponsor
🔥 3 hours ago
DevOps Engineer II improving CI/CD, cloud infrastructure, and reliability for Granicus government technology platforms. Applying AI and AIOps to deployment, observability, and incident response.
🇺🇸 United States – Remote
💵 $75k - $105k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🦅 H1B Visa Sponsor
🔥 5 hours ago
1001 - 5000
Sr. DevOps Engineering Manager leading CI/CD, cloud automation, and operational resilience for Penn Mutual’s financial services business. Guiding engineers and modernizing AWS delivery practices.