Observability Engineer

🕒 5 days ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Miratech

Miratech

501 - 1000 employees

Founded 1989

💰 Private Equity Round on 2022-04

Miratech helps visionaries to change the world. We are a global IT services and consulting company that brings together global enterprise innovation and start-up innovation. Today we support digital transformation for the largest enterprises on the planet.

📋 Description

• Design and implement end-to-end observability solutions across applications, infrastructure, and cloud environments. • Develop dashboards, alerts, and telemetry frameworks to provide real-time visibility into system health and performance. • Build automation solutions to eliminate repetitive operational tasks and improve efficiency. • Enable runbook automation, self-healing capabilities, and automated incident triage workflows. • Define and implement SLIs, SLOs, and alerting strategies to improve service reliability. • Drive improvements in MTTD and MTTR through actionable alerts and telemetry-driven insights. • Implement proactive monitoring, anomaly detection, and predictive alerting to identify issues before customer impact. • Leverage AIOps capabilities for alert correlation and intelligent incident response. • Integrate observability platforms with CI/CD pipelines, cloud services, and ITSM tools such as ServiceNow. • Collaborate with engineering, product, and operations teams to establish observability standards and operational readiness practices. • Mentor teams and drive adoption of observability best practices across the organization.

🎯 Requirements

• 5+ years of experience in Observability Engineering, Site Reliability Engineering, or related domains. • Hands-on experience with observability platforms such as Dynatrace, Splunk, Grafana, and OpenTelemetry. • Strong expertise in AWS and GCP, with familiarity with cloud-native architectures. • Proficiency in Python for automation and operational tooling. • Experience implementing metrics, logs, events, and distributed tracing (MELT) across distributed systems. • Hands-on experience with Terraform and Infrastructure as Code practices. • Strong understanding of SLIs, SLOs, alerting strategies, and incident response frameworks. • Excellent troubleshooting, communication, and collaboration skills. • Bachelor's degree in Computer Science, Information Technology, or a related field (or equivalent experience).

🏖️ Benefits

• Culture of Relentless Performance : join an unstoppable technology development team with a 99% project success rate and more than 30% year-over-year revenue growth. • Competitive Pay and Benefits : enjoy a comprehensive compensation and benefits package, including health insurance, language courses, and a relocation program. • Work From Anywhere Culture : make the most of the flexibility that comes with remote work. • Growth Mindset : reap the benefits of a range of professional development opportunities, including certification programs, mentorship and talent investment programs, internal mobility and internship opportunities. • Global Impact : collaborate on impactful projects for top global clients and shape the future of industries. • Welcoming Multicultural Environment : be a part of a dynamic, global team and thrive in an inclusive and supportive work environment with open communication and regular team-building company social events. • Social Sustainability Values : join our sustainable business practices focused on five pillars, including IT education, community empowerment, fair operating practices, environmental sustainability, and gender equality.

Apply Now

Similar Jobs

🕒 6 days ago

Twilio

5001 - 10000

Supportability Engineer focused on embedding customer supportability in product development lifecycle. Collaborating between Customer Experience and R&D to drive product excellence at Twilio.

Java

JavaScript

Node.js

Python

Splunk

SQL

Tableau

VoIP

🕒 July 11

Fox Corporation

5001 - 10000

📱 Media

Software Engineer for Fox Digital Video Platform team designing and supporting video workflows. Collaborating with engineers to enhance system performance and video streaming reliability.

AWS

Cloud

JavaScript

React

SDLC

Go

🕒 July 11

Milliman

1001 - 5000

🤝 B2B

⚕️ Healthcare Insurance

💸 Finance

Identity Engineer supporting Milliman’s global Identity & Access Management service by engineering and operating identity capabilities. Collaborating across teams to improve identity services and reduce risk.

Cloud

Cyber Security

DNS

🕒 July 6

Veltris

501 - 1000

🤖 Artificial Intelligence

🤝 B2B

Senior Infra Engineer at Veltris designing, building, and operating platform services. Focused on infrastructure automation, Kubernetes engineering, observability in cloud environments.

Cloud

Grafana

Hadoop

Kubernetes

Prometheus

Redis

Terraform

🕒 July 3

Databricks

1001 - 5000

🤖 Artificial Intelligence

🏢 Enterprise

☁️ SaaS

Forward Deployed Engineer building and productionizing solutions with customers using Databricks. Lead architecture and design decisions in a hands-on, customer-facing role.

Apache

AWS

Azure

Bootstrap

Cloud

Google Cloud Platform

JavaScript

Python

Scala

Spark

TypeScript