
10,000+ employees
Founded 1986
💼 Consulting
🛡️ Insurance
📦 Logistics
Consulting • Insurance • Logistics
SS&C Technologies is a global leader in financial services and healthcare technology, recognized as the world’s largest independent hedge fund and private equity administrator, as well as the largest mutual fund transfer agency. The company offers a comprehensive range of solutions including asset management, banking, healthcare, insurance, and wealth management, leveraging proprietary software and extensive expertise to meet the operational needs of clients across multiple industries.
🔥 24 minutes ago
🌪️ Kansas, New Hampshire, +2 more states – Remote
💵 $110k - $120k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
Improve your chances of getting an interview by checking your resume score before you apply.

10,000+ employees
Founded 1986
💼 Consulting
🛡️ Insurance
📦 Logistics
Consulting • Insurance • Logistics
SS&C Technologies is a global leader in financial services and healthcare technology, recognized as the world’s largest independent hedge fund and private equity administrator, as well as the largest mutual fund transfer agency. The company offers a comprehensive range of solutions including asset management, banking, healthcare, insurance, and wealth management, leveraging proprietary software and extensive expertise to meet the operational needs of clients across multiple industries.
• Monitor the health, availability, performance, and security of production services. • Proactively identify emerging issues using telemetry, logs, metrics, and distributed tracing. • Investigate, troubleshoot, and resolve complex production incidents across application and infrastructure layers. • Act as the L3 escalation point for operational issues that cannot be resolved by L1 or L2 support. • Participate in an on-call rotation for critical production incidents. • Lead incident response activities, including coordination, communication, and post-incident reviews. • Perform root cause analysis and ensure corrective actions are implemented to prevent recurrence. • Develop and maintain operational runbooks, dashboards, alerts, and standard operating procedures. • Improve platform observability by enhancing monitoring, alerting, dashboards, and service-level indicators. • Work closely with software engineering teams to improve service reliability, scalability, and resilience. • Identify opportunities to automate operational tasks and eliminate repetitive manual work. • Support production deployments, infrastructure changes, and maintenance activities. • Assist with disaster recovery exercises, resilience testing, and operational readiness reviews. • Ensure operational activities comply with FedRAMP High security and compliance requirements. • Contribute to continuous improvement initiatives across reliability, performance, and operational excellence.
• U.S. Citizenship (required) • 3–6 years of experience in Site Reliability Engineering, Production Engineering, DevOps, Platform Engineering, or a senior production support role • Experience supporting mission-critical cloud-based production systems • Strong understanding of Linux operating systems and networking fundamentals • Experience troubleshooting distributed applications running in Kubernetes • Experience with public cloud platforms, preferably AWS • Experience with infrastructure as code and configuration management • Strong scripting or programming skills (e.g. Python, Bash, PowerShell, Go, or similar) • Experience using monitoring and observability platforms such as Prometheus, Grafana, CloudWatch, Datadog, Splunk, or OpenTelemetry • Experience analysing application logs, metrics, and traces to diagnose production issues • Understanding of incident management, problem management, and root cause analysis • Strong analytical and troubleshooting skills • Excellent written and verbal communication skills.
• Hybrid Work Model & a Business Casual Dress Code, including jeans • 401k Matching Program • Professional Development Reimbursement • Flexible Personal/Vacation Time Off • Sick Leave • Paid Holidays • Medical, Dental, Vision • Employee Assistance Program • Parental Leave • Discounts on fitness clubs, travel and more!
Apply Now🔥 33 minutes ago
DevSecOps Engineer supporting the design and deployment of secure software solutions within C5MI. Collaborating with teams to enhance system reliability and security compliance.
AWS
Azure
Cloud
Docker
Jenkins
Kubernetes
Microservices
Open Source
Python
🔥 33 minutes ago
Azure DevOps Administrator managing ADO environment and coaching delivery teams to align with SDLC standards. Ensuring configuration and governance for C5MI's Azure DevOps and SDLC compliance.
🇺🇸 United States – Remote
💵 $100k - $120k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
Azure
SDLC
🔥 1 hour ago
Sr. Utilities Engineer optimizing and managing utility systems at Smithfield Foods. Focusing on industrial refrigeration, boiler, and compressed air systems for manufacturing processes.
🔥 1 hour ago
Maintenance & Reliability Engineer leading enterprise-wide initiatives to enhance asset performance and reliability in a manufacturing environment. Collaborating with engineering and maintenance teams to improve operational efficiency.
🇺🇸 United States – Remote
💵 $115k - $150k / year
💰 $1G Post-IPO Debt - Dayforce on 2024-03
⏰ Full Time
🟠 Senior
🔴 Lead
⛑ DevOps & Site Reliability Engineer (SRE)
🔥 1 hour ago
Cloud Operations Engineer supporting Kofax Cloud Solutions technology stack. Assisting in system operation, management, and vendor communications.
🇺🇸 United States – Remote
💰 $1.9M Post IPO equity on 2015-02
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
AWS
Azure
Cloud
Linux
SQL