Search Remote Jobs

Senior Site Reliability Engineer, Cloud Platform

Job not on LinkedIn

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Salve.Inno

Salve.Inno

11 - 50 employees

Founded 2024

đź’Ľ Consulting

📣 Marketing

📦 Logistics

Consulting • Marketing • Logistics

Salve. Inno is a recruitment and consulting firm that connects exceptional talent with businesses through personalized hiring strategies and global remote sourcing. The company specializes in recruitment for roles across sectors such as marketing, forex, and iGaming, offering candidate sourcing, screening, and career-site driven hiring experiences while emphasizing DE&I, communication, and innovative process building. Founded in 2024 and headquartered in Gdańsk, Poland, Salve. Inno operates with a small team and a global footprint via remote job listings and consulting services.

đź“‹ Description

• Maintain the reliability, availability, and performance of production and pre-production environments • Monitor platform health and improve alerting, automation, and operational processes • Respond to production incidents, participate in root cause analysis, and implement long-term improvements • Design, build, and enhance observability solutions using metrics, logs, traces, and dashboards • Partner with software engineers to improve application reliability throughout the development lifecycle • Develop and maintain operational documentation, troubleshooting guides, and runbooks • Automate repetitive operational tasks to improve efficiency and reduce manual intervention • Participate in on-call rotations while continuously improving incident response processes • Promote reliability engineering principles, operational excellence, and continuous improvement across engineering teams

🎯 Requirements

• Bachelor's or Master's degree in Engineering, Computer Science, or a related field • Strong experience operating Kubernetes or other container orchestration platforms • Experience supporting large-scale production services • Hands-on experience with AWS • Experience with Prometheus, Grafana, and ELK • Strong scripting skills in Bash, Python, or Go • Experience administering Linux-based production environments • Experience with Infrastructure as Code or configuration management tools such as Terraform or Ansible • Solid understanding of networking fundamentals, including TCP/IP, DNS, load balancing, and routing • Excellent troubleshooting, communication, and collaboration skills • Proactive mindset with a passion for automation and reliability • Experience with SIP or VoIP technologies (nice to have) • Familiarity with MySQL or PostgreSQL (nice to have) • Experience with Redis or other NoSQL databases (nice to have)

🏖️ Benefits

• Long-term, full-time collaboration • Flexible remote working environment • Professional development opportunities, including training and technical learning • Opportunity to work on innovative cloud technologies used by customers worldwide • Collaborative engineering culture focused on knowledge sharing and continuous improvement • Modern Apple equipment provided • Inclusive, respectful workplace

Apply Now

Similar Jobs

🔥 48 minutes ago

NVIDIA

10,000+ employees

🏥 Healthcare

🏭 Manufacturing

🤖 Artificial Intelligence

Technical Marketing Engineer building NVIDIA enterprise infrastructure solutions for global marketing and sales teams. Designing reference architectures, demos, data center environments, and technical documentation.

🔥 1 hour ago

Astreya

1001 - 5000

đź’Ľ Consulting

📦 Logistics

📣 Marketing

Data center engineer designing POP infrastructure, rack layouts, power systems, and deployment documentation. Supporting Astreya’s global IT managed services through network infrastructure projects and vendor coordination.

🔥 3 hours ago

Planet Depos

201 - 500

⚖️ Legal

🤝 B2B

Lead DevOps Engineer owning compliant AWS environments, databases, CI/CD, and observability for legal-industry software. Mentoring platform engineers and enabling secure, reliable product delivery.

🔥 11 hours ago

Guardian Industries - DeWitt

-

🏭 Manufacturing

🏗️ Construction

Georgia-Pacific reliability engineer improving asset performance across its mailer manufacturing network. Leading maintenance strategies, failure analysis, and reliability standardization across U.S. facilities.

🔥 17 hours ago

Omnissa

1001 - 5000

🤖 Artificial Intelligence

🏢 Enterprise

🏥 Healthcare

Senior DevOps Engineer building secure, automated cloud infrastructure for Omnissa's UEM SaaS offerings. Improving reliability, compliance, vulnerability management, and deployment automation.