
501 - 1000 employees
đŒ Consulting
đ„ Healthcare
đŠ Logistics
Consulting âą Healthcare âą Logistics
Mirantis is a company that specializes in container management and cloud infrastructure solutions. It offers a range of products, including Mirantis Kubernetes Engine (MKE), Mirantis OpenStack for Kubernetes (MOSK), and Mirantis Container Cloud (MCC), which provide enterprise-level Kubernetes and container management platforms. Mirantis also develops tools for secure software supply chains, such as the Mirantis Container Runtime (MCR) and Mirantis Secure Registry (MSR). As an advocate for open source technologies, Mirantis supports various projects and provides resources like Lens Desktop, a popular Kubernetes IDE, and technical support for enterprises adopting cloud-native technologies. Their solutions cater to sectors such as public services, financial services, and broader SaaS and technology services industries.
đ„ 0 minutes ago
đ Kazakhstan, Poland â Remote
đ” âŹ75k - âŹ95k / year
â° Full Time
đĄ Mid-level
đ Senior
đïž Platform Engineer
đ» Ghost score 9%
Improve your chances of getting an interview by checking your resume score before you apply.

501 - 1000 employees
đŒ Consulting
đ„ Healthcare
đŠ Logistics
Consulting âą Healthcare âą Logistics
Mirantis is a company that specializes in container management and cloud infrastructure solutions. It offers a range of products, including Mirantis Kubernetes Engine (MKE), Mirantis OpenStack for Kubernetes (MOSK), and Mirantis Container Cloud (MCC), which provide enterprise-level Kubernetes and container management platforms. Mirantis also develops tools for secure software supply chains, such as the Mirantis Container Runtime (MCR) and Mirantis Secure Registry (MSR). As an advocate for open source technologies, Mirantis supports various projects and provides resources like Lens Desktop, a popular Kubernetes IDE, and technical support for enterprises adopting cloud-native technologies. Their solutions cater to sectors such as public services, financial services, and broader SaaS and technology services industries.
âą Design, build, and operate observability platform components for metrics, logging, distributed tracing, and alerting âą Build telemetry pipelines handling high-cardinality, high-volume data from large infrastructure fleets âą Optimize telemetry pipelines for cost, retention, and query performance âą Define and implement SLO/SLI frameworks and alerting strategies âą Partner with service delivery and operations teams to understand incident-response needs âą Integrate observability tooling with incident-management workflows, including root-cause analysis and post-incident review data âą Improve detection speed and reduce MTTD and MTTR across the platform âą Contribute to the roadmap for AI-assisted operations tooling, including automated triage, anomaly detection, and engineer-assist tooling âą Own the reliability, scalability, and security of the observability stack âą Document architecture, runbooks, and operational practices âą Enable operations teams to diagnose incidents faster with automatically surfaced data âą Scale the observability platform with infrastructure growth while controlling cost and performance
âą Proven experience designing and building observability platforms for large-scale, production infrastructure environments âą Strong hands-on experience with metrics, logging, and distributed tracing tooling, such as Prometheus, Grafana, OpenTelemetry, Loki, Thanos/Cortex/Mimir, Elasticsearch/OpenSearch, and Jaeger/Tempo, or equivalents âą Experience with high-volume telemetry pipelines and tradeoffs involving cardinality, retention, cost, and query latency âą Strong software engineering skills in at least one language commonly used in this space, such as Go, Python, or Rust âą Experience with Kubernetes and cloud-native infrastructure âą Solid understanding of SLO/SLI/error-budget practices and low-noise alerting design âą Comfortable working in a fast-moving environment where the platform is built alongside the infrastructure it monitors âą Strong communication skills and ability to work directly with operations/service delivery teams âą Experience building observability for GPU/HPC infrastructure or other specialized, high-performance compute environments (preferred) âą Experience with eBPF-based observability tooling (preferred) âą Familiarity with AIOps/ML-based anomaly detection or automated triage systems (preferred) âą Experience operating in a managed services or MSP context (preferred) âą Contributions to open-source observability projects (preferred)
âą Work with an established Silicon Valley leader in the cloud infrastructure industry âą Work with exceptionally passionate, talented and engaging colleagues âą Help Fortune 500 and Global 2000 customers implement next-generation cloud technologies âą Be part of cutting-edge, open-source innovation âą Professional development and training âą Attend conferences and working groups âą Company outings, happy hours, hackathons, and tech talks âą Competitive compensation package with a strong benefits plan
Apply Now