Site Reliability Engineer

🕒 June 6

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of OneStream Software

OneStream Software

1001 - 5000 employees

💸 Finance

🏢 Enterprise

💰 Series B on 2021-04

Software • Finance • Enterprise

OneStream Software is a financial software solutions provider that specializes in streamlining financial consolidation, reporting, and planning processes for businesses. Their platform is designed to enhance organizational efficiency by integrating various financial functions into a single, easily manageable system, thereby enabling businesses to make informed decisions based on real-time financial data.

📋 Description

• Implement application/infrastructure observability solutions to ensure desired application availability, reliability, and performance • Participate in regular On-Call rotations and share details related to incidents and their resolution through post-mortem reports and regular review meetings • Proactively partner with Product and Engineering teams to identify, develop, deploy, and maintain reliable systems and services • Influence and create new designs, architectures, standards, and methods for large-scale systems • Sustain a high level of reliability for key services and automated systems • Automate processes to improve reliability, performance, and availability • Update technical documentation, workflows, and knowledge base articles • Provide feedback in pull requests and peer coding reviews • Implement codified automated solutions that build integrations between Dynatrace, Azure DevOps and Jira • Solid knowledge in focused areas of OneStream Software • Ability to mentor others in several technical areas • Understanding practical use of SOC/FedRAMP controls to assist Compliance and Security teams

🎯 Requirements

• BS/BA in computer science, engineering, or technology-related field (or equivalent work experience) • Proven work experience as a Site Reliability Engineer or in a similar role • 6+ years of cloud infrastructure and software development experience • 2+ years hands on experience of Azure Kubernetes Services (AKS) with container-based deployment skills or other platforms such as OpenShift, GKS, EKS • Advanced understanding of APM and observability tools such as Dynatrace, AppInsights, DataDog, Log Analytics, New Relic, Prometheus and Grafana • Advanced understanding of Infrastructure-as-Code (IaC) concepts and tooling (Terraform, CloudFormation templates, Bicep or ARM templates) on Microsoft Azure, Amazon Web Services (AWS), or Google Cloud Platform (GCP) • Deep knowledge of Configuration Management/Orchestration utilities such as Ansible, PowerShell DSC, Chef, and Puppet • Advanced understanding of cloud concepts including elasticity, security, and identity management • Well versed familiarity with Agile Development methodologies utilizing Jira or Azure DevOps Boards • 6+ years of hands-on experience with the following technologies, tools, and concepts: Automating processes using PowerShell, Bash, CLI, REST APIs, python, ARM Templates or other scripting languages • Comfortable leveraging source control tools such as Git, Azure DevOps, or GitHub • Knowledge of container orchestration platforms such as Kubernetes, OpenShift, AKS, GKS or helm • Microsoft Azure, Amazon Web Services (AWS) or Google Cloud (GCP)

🏖️ Benefits

• Vision • Medical • Life • Dental • 401K

Apply Now

Similar Jobs

🕒 June 6

RAPIDFORT

51 - 200

💼 Consulting

🎖️ Defense

🏥 Healthcare

DevSecOps Engineer designing and maintaining secure cloud-native infrastructure. Delivering hardened software systems in collaboration with government clientele for the Department of War.

AWS

Azure

Cloud

Grafana

Jenkins

Kubernetes

Prometheus

🕒 June 6

Maxor National Pharmacy Services, LLC

1001 - 5000

💼 Consulting

📦 Logistics

🏥 Healthcare

DevOps Engineer II at VytlOne responsible for modernizing DevOps practices. Focused on CI/CD pipelines, environment management, and deployment processes.

🇺🇸 United States – Remote

💵 $105k - $120k / year

💰 Debt Financing on 2014-01

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

Azure

Cloud

🕒 June 5

SmithRx

51 - 200

💼 Consulting

🛡️ Insurance

🏥 Healthcare

Senior DevOps Engineer managing cloud-based infrastructure and CI/CD for health-tech firm SmithRx. Collaborating with teams to build a scalable and secure environment while ensuring regulatory compliance.

Amazon Redshift

AWS

BigQuery

Cloud

Groovy

Kubernetes

NoSQL

Perl

Postgres

Python

Redis

Ruby

SQL

Terraform

Go

🕒 June 5

SS&C Technologies

10,000+ employees

💼 Consulting

🛡️ Insurance

📦 Logistics

Sr. Network Operations Reliability Engineer at SS&C, a financial services and healthcare technology company. Working with complex network solutions and monitoring global network operations.

Ansible

Cloud

DNS

Firewalls

OpenShift

Python

ServiceNow

Switching

VMware

🕒 June 5

Equinix

5001 - 10000

📦 Logistics

💼 Consulting

🏭 Manufacturing

Senior Staff Engineer developing a scalable network model for Equinix’s digital infrastructure across global data centers. Collaborating to automate network management tasks and enhance customer services.

Cloud

Docker

Google Cloud Platform

GRPC

Kubernetes

MongoDB

Neo4j

Redis

Go