
51 - 200 employees
Founded 2010
💼 Consulting
🎖️ Defense
📦 Logistics
Consulting • Defense • Logistics
Cognitive Medical Systems, Inc. is an IT and software engineering services firm dedicated to producing better outcomes within U. S. government healthcare programs. With a blend of real-world clinical experience and technical expertise, the company delivers healthcare IT solutions that enhance interoperability and improve care delivery. Their offerings include clinical knowledge management, analytics, custom software development, and secure cloud hosting solutions, aimed at converting health data into positive health outcomes while ensuring compliance and integration across various IT environments.
🔥 16 hours ago
🌵 Arizona, Colorado, +7 more states – Remote
💵 $90k - $115k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
Improve your chances of getting an interview by checking your resume score before you apply.

51 - 200 employees
Founded 2010
💼 Consulting
🎖️ Defense
📦 Logistics
Consulting • Defense • Logistics
Cognitive Medical Systems, Inc. is an IT and software engineering services firm dedicated to producing better outcomes within U. S. government healthcare programs. With a blend of real-world clinical experience and technical expertise, the company delivers healthcare IT solutions that enhance interoperability and improve care delivery. Their offerings include clinical knowledge management, analytics, custom software development, and secure cloud hosting solutions, aimed at converting health data into positive health outcomes while ensuring compliance and integration across various IT environments.
• Own operational reliability, observability, release coordination, and continuous delivery health for CMS Drug Data Processing System (DDPS) and Payment Reconciliation System (PRS) production environments • Coordinate and execute production releases, including deployment readiness validation, cutovers, post-implementation validation, and rollback execution • Maintain real-time monitoring and observability using AWS CloudWatch, Splunk, Splunk On-Call, New Relic, and Datadog • Drive reliability improvements, automate remediation, reduce manual operational tasks, and improve resilience and MTTR • Maintain alerting and on-call runbooks; coordinate incident response and escalation with cybersecurity and DevSecOps leadership • Plan and execute disaster recovery tests, failover/failback exercises, and validate RTO/RPO with cloud architecture • Manage release schedules, validation reports, and stakeholder communications • Provide coverage during peak operational windows and participate in the on-call rotation
• Bachelor's degree in Computer Science or related field • 4 or more years in site reliability engineering, DevOps, or release management on production federal systems • Hands-on experience with AWS CloudWatch (Logs, Events, Alarms), Splunk, and Splunk On-Call/VictorOps • Experience with New Relic and Datadog for application performance monitoring • Experience coordinating production releases, cutovers, and rollback execution on mission-critical systems • Experience developing and maintaining operational runbooks, alerting rules, and SRE playbooks • Familiarity with Linux environments, including RHEL, CentOS, and Amazon Linux 2 • Bash scripting for operational automation • Ability to pass CMS and internal required background checks for public trust • Must be eligible to work from one of the designated U.S. states: Virginia, District of Columbia, Maryland, Tennessee, Florida, Arizona, Colorado, Oregon, or Texas • Position contingent upon contract award
• 100% remote work • Supportive work/life balance • Opportunities for growth and development • Mission-driven healthcare IT work • Collaboration with innovative and passionate professionals
Apply Now🔥 16 hours ago
DevOps Engineer supporting Molina Healthcare’s large-scale Azure cloud environment. Automating infrastructure, CI/CD pipelines, and Kubernetes deployments remotely in the U.S.
🔥 19 hours ago
Senior DevOps Engineer building and operating AI cloud infrastructure for Bitdeer’s Bitcoin mining and AI computing platforms. Automating deployments, Kubernetes infrastructure, GPU workloads, reliability, and security.
🇺🇸 United States – Remote
💰 Post-IPO Equity on 2023-05
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🔥 20 hours ago
Sr. Staff Engineer operationalizing Databricks for Shield AI, a defense-tech company building autonomous aircraft and intelligent systems. Ensuring secure, scalable, observable production data infrastructure.
🇺🇸 United States – Remote
💵 $180k - $270k / year
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🕒 Yesterday
GCP DevOps Engineer designing secure infrastructure, CI/CD, and Kubernetes platforms for Sonatype. Improving developer delivery, reliability, observability, and software supply-chain security.
🇺🇸 United States – Remote
💰 $80M Private Equity Round on 2018-09
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🕒 Yesterday
Site Reliability Architect designing observability and reliability platforms for Summit’s regulated-industry application hosting and cloud services. Improving resilience, automation, and incident response across teams.
🇺🇸 United States – Remote
💵 $136k - $175k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🦅 H1B Visa Sponsor