Private Cloud Engineer – Infrastructure Operations

Job not on LinkedIn

🔥 1 hour ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Cloudera

Cloudera

1001 - 5000 employees

Founded 2008

💼 Consulting

🏥 Healthcare

📦 Logistics

💰 $4.1M Venture Round on 2013-01

Consulting • Healthcare • Logistics

Cloudera is a leading enterprise data cloud company that empowers businesses to manage and analyze data across any environment. Offering a hybrid data platform, Cloudera facilitates modern data architectures with solutions like open data lakehouse, scalable data mesh, and unified data fabric, designed for artificial intelligence, data engineering, and machine learning. Key industries served include financial services, telecommunications, healthcare, and more, where Cloudera's platform enables secure, scalable, and effective data management. By leveraging AI and advanced analytics at scale, Cloudera helps organizations transform their data into actionable insights.

📋 Description

• Monitor platform health, utilisation, and capacity across all private cloud environments • Manage compute provisioning, configuration, and lifecycle operations • Track resource utilisation and deliver reporting - driving efficiency and accountability • Manage quotas, resolve capacity issues, and plan for growth • Provide first-line operational support - triage incidents, apply known fixes, escalate appropriately • Enforce operational standards, access controls, and governance policies • Manage environment intake and provisioning for consumers • Support compute and storage infrastructure operations (server lifecycle, Ceph storage, shared services) • Maintain operational documentation and runbooks • Participate in on-call rotation for infrastructure services • Contribute to infrastructure efficiency and automation initiatives

🎯 Requirements

• Relevant studies / BS or MS in related field or equivalent experience • 3+ years operating cloud or virtualised infrastructure at scale • Strong Linux systems administration (RHEL/CentOS/Ubuntu) • Hands-on experience with distributed storage (Ceph preferred) or SAN operations • Experience with monitoring stacks, dashboards, alerting, and capacity planning • Virtualisation experience (KVM, VMware, or similar) • A governance mindset - comfortable enforcing standards, reporting compliance, holding consumers accountable for waste • Ability to author and maintain clear runbooks, procedures, and operational documentation • Strong communication - can translate technical state into clear reporting for stakeholders

🏖️ Benefits

• Generous PTO Policy • Support work life balance with Unplugged Days • Flexible WFH Policy • Mental & Physical Wellness programs • Phone and Internet Reimbursement program • Access to Continued Career Development • Comprehensive Benefits and Competitive Packages • Paid Volunteer Time • Employee Resource Groups

Apply Now

Similar Jobs

🕒 2 days ago

OpsMill

11 - 50

☁️ SaaS

🤖 Artificial Intelligence

🤝 B2B

Product Reliability Engineer enhancing on-prem deployment reliability for OpsMill's Infrahub while partnering with customers on troubleshooting and improvements. Building diagnostics tools and resolving complex issues in Kubernetes environments.

Distributed Systems

Kubernetes

Python

Rust

Go

🕒 3 days ago

Talentgrator

11 - 50

🎯 Recruiter

👥 HR Tech

🎲 Gambling

Experienced DevOps Engineer responsible for server administration, automation, and system management at an iGaming company. Collaborating with development teams and implementing CI/CD pipelines for improved workflows.

Ansible

DNS

Docker

ElasticSearch

Flux

Grafana

Kafka

Kubernetes

MongoDB

NGINX

Postgres

Prometheus

Python

RabbitMQ

Redis

TCP/IP

Terraform

Go

🕒 3 days ago

Talentgrator

11 - 50

🎯 Recruiter

👥 HR Tech

🎲 Gambling

DevOps Tech Lead overseeing high-load infrastructure design for iGaming company. Leading DevOps team and shaping architectural decisions with a focus on reliability and security.

Cloud

Distributed Systems

Kubernetes