Staff Software Engineer – Reliability, Platform

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of One Identity

One Identity

501 - 1000 employees

🔒 Cybersecurity

☁️ SaaS

🏢 Enterprise

💰 $3M Venture Round on 2004-07

Cybersecurity • SaaS • Enterprise

One Identity is a company that specializes in identity and access management solutions, focusing on protecting digital identities and simplifying user access across organizations. They offer a comprehensive suite of products designed to secure privileged access, govern user identities, and streamline compliance with regulations through automation. Their platform integrates AI-driven insights to enhance security and operational efficiency, supporting both on-premises and cloud environments.

📋 Description

• Own production operability by debugging complex issues, improving system visibility, and eliminating recurring problems at the source • Improve mean time to detect (MTTD), mean time to resolve (MTTR), and recurrence rates for issues • Identify systemic issues and eliminate recurring problems through code fixes, architecture improvements, and better operational tooling • Improve observability across services — logs, metrics, and alerting — for faster diagnosis and resolution • Design and improve debugging workflows, runbooks, and internal tooling for engineers • Reduce operational burden by making systems easier to understand, operate, and troubleshoot • Partner closely with product teams to feed production learnings back into design and development • Reduce support and incident load by addressing root causes and improving system design, not just resolving individual issues

🎯 Requirements

• 4+ years of software engineering experience with ownership of production systems, reliability, or operational improvements • Strong backend development experience (Ruby, Node.js, or similar) • Solid understanding of REST APIs, service contracts, and software design principles • Experience building and operating services in AWS or similar cloud environments. • Good understanding of distributed systems, cloud-native architecture, and CI/CD. • Experience with observability, production debugging, and incident response. • Willingness to participate in a mandatory 24/7 on-call rotation. • Experience using, or strong interest in, AI-powered development tools (e.g., GitHub Copilot, ChatGPT, Cursor).

🏖️ Benefits

• Health insurance • Professional development opportunities

Apply Now

Similar Jobs

🕒 May 12

Stellar Cyber

51 - 200

🔒 Cybersecurity

🤖 Artificial Intelligence

🏢 Enterprise

Staff SRE Engineer driving reliability and scalability in production systems for global cybersecurity leader. Collaborating with teams and influencing architecture with advanced cloud technologies.

AWS

Azure

Cloud

Distributed Systems

ElasticSearch

Google Cloud Platform

Grafana

Kafka

Kubernetes

Linux

MongoDB

Prometheus

Python

Redis

Spark

Terraform