Observability Technical Lead

🕒 vor 20 Tagen

🐊 Florida – Remote

infoinfo

💵 $96.000 - $192.000 / Jahr

⏰ Vollzeit

🟠 Senior

🧑‍💻 Full-Stack-Entwickler

🦅 H1B-Visum-Sponsor

infoinfo

👻 Geisterscore 0%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of Carrier

Carrier

10.000+ Mitarbeiter

Gegründet 1915

🏗️ Bauwesen

🏥 Gesundheitswesen

📦 Logistik

Construction • Healthcare • Logistics

Carrier ist ein weltweit führender Anbieter von Lösungen für Gebäude und die Kühlkette und engagiert sich für Innovationen, die gesunde, sichere, nachhaltige und intelligente Umgebungen schaffen. Das Unternehmen ist im Bereich HVAC (Heizung, Lüftung, Klima) und Kältetechnik tätig und konzentriert sich darauf, mit seinen fortschrittlichen Technologien die Gesundheit und Sicherheit von Innenräumen zu fördern und die globale Versorgung mit Lebensmitteln und Arzneimitteln zu sichern. Carrier setzt sich zudem dafür ein, den Klimawandel anzugehen, und arbeitet mit Partnern zusammen, um Nachhaltigkeit und Energieeffizienz in der Infrastruktur voranzutreiben. Durch die Entwicklung von Smart-Building-Lösungen und Resilienz der Energieversorgung macht Carrier wichtige Fortschritte auf dem Weg in eine Net-Zero-Zukunft.

Beschreibung

• Serve as the senior SME for enterprise observability and define architectures, standards, integration patterns, and reusable solutions • Design and optimize observability across cloud, on-premises infrastructure, networks, Kubernetes, containers, applications, APIs, databases, middleware, and enterprise platforms • Standardize metrics, logs, traces, events, topology, dashboards, alerting, instrumentation, and service health • Provide technical leadership for LogicMonitor, Splunk, OpenTelemetry, Grafana, Prometheus, Tempo, VictoriaMetrics, Loki, and related technologies • Lead OpenTelemetry adoption, including instrumentation, collectors, telemetry pipelines, distributed tracing, context propagation, and vendor-neutral standards • Design telemetry pipelines routing metrics, logs, and traces across multiple platforms • Develop API-driven and Observability-as-Code capabilities for onboarding, configuration, validation, and lifecycle management • Establish governance for RBAC, telemetry standards, alerting, retention, integrations, configuration management, and data lifecycle • Improve observability coverage, telemetry quality, scalability, reliability, performance, and cost efficiency • Drive automation and self-service through APIs, Infrastructure-as-Code, CI/CD, GitOps, and reusable observability patterns • Integrate Edwin AI and AIOps capabilities for anomaly detection, event correlation, root-cause analysis, investigation, and operational intelligence • Connect telemetry and operational context to deliver end-to-end visibility • Establish reliability standards including SLIs, SLOs, error budgets, monitoring coverage, alert quality, MTTD, and MTTR • Reduce alert fatigue through intelligent correlation, dynamic thresholds, suppression, automation, and event-management practices • Evaluate eBPF, continuous profiling, Kubernetes observability, dependency mapping, and auto-instrumentation • Advance operations from reactive monitoring toward proactive and predictive operations • Collaborate with engineering and operations teams across North America, Europe, and Asia • Lead architecture reviews, platform evaluations, workshops, and technical working sessions • Influence observability strategy and standards across teams without direct authority • Mentor engineers and promote observability, SRE, and reliability best practices • Evaluate emerging technologies and recommend adoption based on interoperability, scalability, business value, and cost • Translate strategy into standards, reference architectures, and reusable implementation patterns

🎯 Anforderungen

• Bachelor’s Degree • 7+ years of experience in observability, monitoring, APM, SRE, DevOps, platform engineering, or related professional experience • 7+ years of experience designing, implementing, operating, and maintaining enterprise-scale observability platforms using observability-as-code, monitoring-as-code, infrastructure-as-code, GitOps, and CI/CD practice • Experience with AWS, Azure, GCP, and large-scale Kubernetes environments • Experience designing OpenTelemetry Collector architectures and telemetry pipelines • Expertise with Grafana, Tempo, Loki, and Prometheus-compatible platforms • Experience with VictoriaMetrics or other large-scale time-series databases • Knowledge of eBPF, continuous profiling, auto-instrumentation, and cloud-native telemetry • Experience integrating observability platforms with ServiceNow, ITSM, CMDB, incident management, and automation platforms • Understanding of SRE practices including SLIs, SLOs, error budgets, and incident management • Experience enabling developer self-service and internal developer platform integrations • Knowledge of RBAC, secrets management, governance, compliance, and telemetry data protection • Reasonable schedule flexibility to support global collaboration

🏖️ Vorteile

• Short-term cash incentives, subject to plan requirements • Medical, Dental, and Vision benefits • Wellness incentives • Retirement benefits • Paid vacation days, up to 15 days • Paid sick days, up to 5 days • Paid personal leave, up to 5 days • Paid holidays, up to 13 days • Birth and adoption leave • Parental leave • Family and medical leave • Bereavement leave • Jury duty leave • Military leave • Purchased vacation • Short-term and long-term disability • Life Insurance and Accidental Death and Dismemberment • Health Savings Account • Health Care Spending Account • Dependent Care Spending Account • Tuition Assistance

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 20 Tagen

Togal.AI

11 - 50

🏗️ Bauwesen

💼 Beratung

🤖 Künstliche Intelligenz

GTM Engineer owning predictive models, campaign infrastructure, and AI personalization. Building revenue systems for an AI-powered construction pre-estimating platform.

🇺🇸 Vereinigte Staaten – Remote

💰 €7.000.000 Seed Round - Togal.AI im 2023-03

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

🧑‍💻 Full-Stack-Entwickler

🗣️🇺🇸🇬🇧 Englisch erforderlich

Apollo

SQL

🕒 vor 20 Tagen

PrizePicks

201 - 500

🎮 Gaming

⚽ Sport

Software Engineer III building scalable Ruby on Rails services for PrizePicks’ daily fantasy sports platform. Owning rewards and loyalty features, architecture, code quality, and junior engineer guidance.

🇺🇸 Vereinigte Staaten – Remote

💵 $145.000 - $165.000 / Jahr

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

🧑‍💻 Full-Stack-Entwickler

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 21 Tagen

GE Vernova

10.000+ Mitarbeiter

💼 Beratung

📦 Logistik

🏭 Fertigung

Technical Leader designing SCADA, cybersecurity, and grid automation solutions for GE Vernova’s energy-sector customers. Leading projects from presales and technical design through testing, commissioning, and customer training.

🇺🇸 Vereinigte Staaten – Remote

💵 $98.400 - $164.000 / Jahr

⏰ Vollzeit

🟠 Senior

🧑‍💻 Full-Stack-Entwickler

🗣️🇺🇸🇬🇧 Englisch erforderlich

Cyber Security

Firewalls

🕒 vor 21 Tagen

Vontier

5001 - 10000

🚘 Automobilindustrie

⚡ Energie

🔧 Hardware

Software Engineer II developing automated tank gauge and fueling infrastructure software for Gilbarco Veeder-Root. Designing solutions, applying modern technologies, improving Agile practices, and mentoring engineers.

🇺🇸 Vereinigte Staaten – Remote

💵 $96.800 - $120.000 / Jahr

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

🧑‍💻 Full-Stack-Entwickler

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 21 Tagen

General Motors

10.000+ Mitarbeiter

🚘 Automobilindustrie

🏭 Fertigung

🚗 Transport

Principal AV safety engineering lead shaping GM’s safety strategy for Level 3–4 autonomous driving systems. Defining behavior validation, safety cases, and launch readiness for millions of customers.

🇺🇸 Vereinigte Staaten – Remote

💵 $250.600 - $384.600 / Jahr

💰 €500.000.000 Grant im 2024-07

⏰ Vollzeit

🟠 Senior

🧑‍💻 Full-Stack-Entwickler

🦅 H1B-Visum-Sponsor

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

AWS

Azure

Cloud

Docker

Google Cloud Platform

Java

Kubernetes

Numpy

Pandas

PySpark

Python

PyTorch

Scikit-Learn

SQL

Tableau

Tensorflow